A Preview of DuckDB v2.0
ibotty
588 points
106 comments
August 17, 2026
Related Discussions
Found 5 related stories in 43.4ms across 4,128 title embeddings via pgvector HNSW
- Choose DuckDB rather than SQLite rubenvanwyk · 85 pts · July 29, 2026 · 67% similar
- DuckDB V2 PEG-based SQL parser karma_daemon · 62 pts · August 21, 2026 · 66% similar
- DuckPGQ – A DuckDB community extension for graph workloads rzk · 49 pts · July 24, 2026 · 61% similar
- Show HN: SirixDB 1.0 Beta – Git-Like Versioning, Diffs, Time-Travel Queries lichtenberger · 12 pts · July 15, 2026 · 51% similar
- Logseq 2.0 Beta (DB version) is here karencarits · 94 pts · July 13, 2026 · 49% similar
Discussion Highlights (20 comments)
c9cf35860db4
The last year of DuckDB enhancements feel like the shift from in-process execution engine (which it is phenomenal at) to an engine that can serve as the foundation of a cloud data warehouse. I know the founders were reticent about not wanting to build that, but I have a feeling it is in the works.
markhalonen
Was hoping to see procedural functionality like PL/pgSQL... regardless, an astonishing project overall.
srameshc
I <3 DuckDB. It has become one of my go to tools for storing, data processing , integrations and now even graph. More importantly it's fun to use because it is so portable. Looking forward to v2.
est
This is cool What about the runtime size? I care this because I intend to run a stripped WASM version of DuckDB in browser.
jtbaker
DuckDB is one of the things I've been most excited about in a long time. Introduced it to projects at 3 companies since 2023, greatly lowering resource requirements and running it in a variety of environments. Just having the ability to do out of core bigger than memory data processing on lower end consumer grade hardware is remarkable. Thanks to the team for everything!
sv123
DuckDB is so cool, game changer when it comes to local data processing.
hnlb53nrpg
Same problem, different day
anentropic
Please document the new "extensible PEG-based parser" for extension authors
badatnames
Looks like an awesome release, but the smell of AI from that post is horrid. Here is a wild idea: is it really so hard to edit out sentences structured and punctuated like this - it's so painfully obvious and distracts from the content. The effect is real.
jeffbee
"We reimplemented ICU" U+1F631 FACE SCREAMING IN FEAR
logancbrown
Funny to think one of my favorite software projects this decade is basically "lets make it easy to host your own OLAP database".
amluto
If I could have a pet feature added to DuckDB, it would be some form of native ordered table. In a database like Clickhouse or any of the dedicated time series DBMSes or log stores, there’s a built-in concept that a table might have an order, and the database will optimize based on the order. But, for databases that are logically just bags of rows (traditional DBMSes and also DuckDB [0]), you either need an index or you need to rely on full table scans or at least scans of big blocks. DuckDB does the latter really well, but I think it would be quite nice for some workflows to have explicit ordering. Also, I bet compression could work a lot better with ordering hints. All that being said, I’m quite excited about DuckDB 2.0. I want to give the improved VARIANT support a try. [0] Documentation on DuckDB’s native format is rather sparse AFAICT. But the DDL has nothing resembling an ordered table.
aleda145
Excited about a stable C++ API for extensions! I made a dry run extension a few months ago ( https://github.com/aleda145/duckdb-dryrun ), will be so nice to build it just once and know that it will always work. Also urge anyone to make an extension, the template makes it quite smooth: https://github.com/duckdb/extension-template
otter-in-a-suit
Super excited about Quack (partially due to the name). I use duckdb for both analytics and runtime, but I do have to serve/handle/manage a giant, multi-GiB duckdb file as effectively a runtime artifact[1]. I'm aware that this isn't the _perfect_ database for this, but the mix of it being fast, having spatial support, sane coding interfaces, great dbt integration, and me being able to do everything between "run a giant several hundred step dbt pipeline" to "query the output of said pipeline" to "read/query a csv on disk" with the exact same tool is just so nice. If I could centrally manage said asset more akin to a traditional database, I'd be very happy. I've partially solved this with separate databases for different steps in the data pipeline(s) and have even experimented with Clickhouse as a complete alternative, but I really like way too many things about duckdb to replace it. [1]: If you care: https://skaldmaps.com/blog/2026/07/zip-codes-are-a-bad-spati...
drannex
Really looking forward to that new Async system, especially when reading/querying against thousands of parquet files. This is going to monumentally affect me and my work - I have to query against millions of massive parquet files and the speed has already been rather wonderful, but if those metrics are to be even 100% in range, this is going to make life so much better. DuckDB is seriously an incredible utility.
giovannibonetti
Disappointed, since I was expecting they would rewrite the implementation from C++ to Zig. I bet that would increase the number of positive pull requests they get, since most developers prefer to stay away from C++ nowadays.
brunoborges
How does DuckDB compares with PostgreSQL / MariaDB ?
thejosh
I've been working on a demo database project, and have been really impressed by the UI. So glad they decided to put more effort into it, it has made building a "follow along" tutorial really nice.
dzonga
well done to the duckDB team - one of the features I'm waiting for is real time materialized views.
d33
Are there improvements in how memory_limit works? I often had DuckDB get OOM killed because it went beyond its limit. It's definitely one of the reasons why I usually have an AI tune the environment for my datasets.