

Join the discussion
Write your take first — we'll ask for email only when you're ready to publish.
- Hacker News
- Hope it goes well for you Andy. Have loved your lectures over the years.
Best of luck!
by Tostino - Hi Andy, that’s very cool news. Will you be at the HPTS workshop in October?by adrianco
- TBDby apavlo
- What has been your experience with HPTS? It will be my first time presenting.by laluser
- That's great news for Clickhouse and for Andy I guess, they are both obsessed with databases and are at the cutting edge.by throwaw12
- Andy, go ahead and correct the https://dbdb.io/db/bigquery How did you come to the conclusion that BigQuery's data model is "Document / XML" :joy:
This entry is not only wrong, it's way outdated, therefore misinforming. I'm wondering about the overall quality of this dbdb.io database at large.
- There are several out of date entries in there. We used to have students help with writing entries. I just did a major refresh of the site this summer. I need to go through all the entries and figure out how to update them.by apavlo
- > The goal of ClickHouse Labs is to establish a best-in-class industry research organization focused on databases. It will not operate as an isolated research organization that throws ideas over the wall to engineering. Instead, we will work closely with ClickHouse engineers, customers, collaborators, and industry partners to develop and disseminate new ideas that keep ClickHouse at the bleeding edge.
This is cool, though bittersweet that the public research infrastructure (universities) is not really configured to support this kind of high-impact research any more.
by pphysch - What do you mean? There’s plenty of high impact research on databases happening in academia. (Including, for example, Andy’s group at CMU)
Also keep in mind that the part you quoted is partially marketing copy.
by Ar-Curunir - Andy taught me a lot about trolling. I also hear he's banned from every post office within 50 miles of Baltimore.by bhickey
- Ok now I need to know.by mkw5053
- Clickhouse just became the hottest talent-attraction on the market.
Congrats Andy, hope you enjoy the ride =)
by gavinray - Congrats, that's incredible news. Watched apavlo@'s lectures while studying in the university and finished my bachelor thesis implementing features and doing research at ClickHouse. Surprised to see these worlds being together now!by danlark1
- Its refreshing to see corporate research labs in a non AI area. Clickhouse has been a huge beneficiary of the AI wave and its good to see some of the value going into advancing fundamental research in infrastructureby adi2907
- Always enjoyed his lecture series from CMU, hopefully those continue in a sponsored format from Clickhouse.by tomsanbear
- They will continue. New seminar series starts next month (announcement coming this week).by apavlo
- I'm very curious about the convergence of the best in class fast OLAP products (StarRocks, ClickHouse) with Trino. It sounds like everybody is going for decoupled compute/storage, using S3 or similar as the storage layer, and thus forgoing colocated joins (ok I know ClickHouse joins suck)...
So what does this mean for ingestion (and indexing)? Iceberg V3? Paimon? Bespoke ingestion through the DB engine to do the indexing?
by a34729t - Clickhouse joins have been improving almost every month for the last couple of years. Maybe they still suck but a lot lessby mrits
- I’ve historically read this as ‘open format compatible’ but ‘native preferred’ - where this opens up market space and dev velocity - but it’ll be interesting to see if native storage differentiation gets dumped entirely. It just seems like ‘fork and optimize for our engine’ would always be tempting enough that you’d want a native play for when you don’t need the decoupling.by efromvt
- Opinionated take: ingestion should be treated as a streaming reorganization workload, separate from whatever a "database" is.
You do not even need to change the Iceberg spec, although you could. Just rejig things a bit to produce better statistics (think liquid clustering but on the write path). You still write the same dataset queried the same way, but automatically get better query performance. Workload owners should be able to decide the tier at which this streaming reorganization happens: they know their data staging and bandwidth constraints the best.
by mrlongroots - It's also interesting how Clickhouse / Starrocks can now also act as a query planner and executor on top of non-native formats (ex. Iceberg).
I assume the native formats will always be faster / more optimized but the need for Trino as a separate executor while running either of these databases seems to be close to gone.
by jimmyl02 - Looks like Andy's here, so if you see this - please also try to convince ClickHouse to consider funding DB research in academia. With all the money being poured into AI and the chaos in government funding, there is almost nothing for DB research anymore.by remywang
- Remy - Good to hear from you. Yes things are mess right now. Let me get settled with setting up the lab and we'll figure out how we want to engage with academia.by apavlo