DuckDB
is an analytical in-process SQL database management system.
🦆 A curated list of awesome DuckDB resources
This page lists names, links and short descriptions. The original list on GitHub is the source and belongs to its authors.
Official DuckDB documentation.
Official DuckDB blog.
Feed for the official DuckDB blog.
Client APIs for DuckDB.
The DuckDB documentation as a single PDF file.
The DuckDB documentation as a single Markdown file.
A lakehouse format from the team behind DuckDB.
Official Docker image for the DuckDB CLI.
GitHub Action to install DuckDB in CI.
Running DuckDB over a data lake on S3 using lambda.
Collection of snippets curated by MotherDuck.
DuckDB's entry in tldr pages, available in CLI via the tldr duckdb command.
Run DuckDB in AWS Lambda functions.
Run DuckDB in AWS Lambda functions using Python.
Use DuckDB as API with Amazon API Gateway and AWS Lambda.
Use DuckDB to repartition data in S3-based Data Lakes.
Notebooks using DuckDB on the Observable data visualization platform.
Example uses of DuckDB with Nextflow.
A CLI tool to create an ER Diagram from DuckDB database files.
SQL notebooks by TimerStored powered by DuckDB.
Visualizing and understanding DuckDB EXPLAIN plans made easy.
A curated list of awesome DuckLake tools and resources.
A collection of scientific papers building on DuckDB.
Cross-platform installer and version manager for DuckDB.
Snap package of DuckDB, e.g., for Ubuntu Linux.
Chocolatey package for Windows.
Monthly newsletter by MotherDuck.
Newsletter by Tobias Müller.
Definite pulls all your data into a single place for analytics and dashboards. No engineering or SQL required. Get a managed data warehouse (DuckDB), ELT, data modeling / transformations and BI in a single platform. (pricing)
Generate reports using SQL and markdown. The DuckDB connector allows querying across DuckDB, CSV, Parquet and JSON. (Evidence Cloud)
Tool for effortlessly transforming data sets into powerful, opinionated dashboards using SQL. (Rill Cloud)
Open-source SQL-driven data dashboards, powering Taleshape, built on DuckDB. (managed hosting)
Open-source, DuckDB-powered BI workspace for AI-assisted analysis, SQL, charts, and dashboards.
Duck Powered simple analysis and business intelligence dashboard. Load excel, csv, parquet files, instantly query your data in the query analyzer tool, make multiple simple dashboards. Data stays local in OPFS local browser storage which is native to Edge and Chrome. (proprietary; free to use)
Business Intelligence Done Right - ReportBurster uses DuckDB as its in-process analytics database engine, and for larger workloads, supports ClickHouse too. (SSPL v1, source-available)
Desktop app (Windows, macOS) for a first look at CSV, Excel, JSON and Parquet files: plain-language questions are turned into DuckDB SQL by a language model that runs on the user's own machine. Can work on a computer with no network connection. (pricing)
Dual visual/textual programming language and analytics platform, allowing DuckDB to be used as a data engine for its visual analytics and low-code workflows. (Enso Cloud)
Latitude uses DuckDB to power data snapshots. Drop a CSV file and query it with SQL at the speed of light. (paid cloud)
A tool that provides interactive visualizations for large embeddings. Uses DuckDB.
Live visualization tool that enables cloud system architects to answer specific instance selection questions, powered by DuckDB-Wasm.
Beautiful visualization and analytics right in the browser. (proprietary app with a free non-commercial tier; its cosmos.gl library is open source)
The privacy-first data analysis toolkit. (proprietary; free tier)
A collaborative data analysis tool. (proprietary; free to use)
Spreadsheet-style data tool with visual canvas, inline editing, and AI commands for SQL transformations. (proprietary; free to use)
A serverless data transformation platform for data lakes. (pricing)
A next-generation data transformation and modeling framework with support for DuckDB connections for state, transformations & running unit tests locally. (Tobiko Cloud)
YAML-based data pipeline framework that runs both locally and fully in-browser designed for data engineers, ML teams, and SaaS developers who need flexible, SQL-powered pipelines.
Load, explore, transform your datasets and expose them via API. Integration with external APIs, S3, PostgreSQL and ChatGPT.
Local-first visual ETL/ELT studio. Drag sources, transforms and sinks onto a canvas; it compiles to plain DuckDB SQL and runs entirely on DuckDB. Open source desktop app, with a built-in MCP server for generating and running pipelines from natural language.
DuckDB-powered ETL tool written in Go, inspired by evidence.dev's syntax. It uses a structured Markdown config where heading levels define nested blocks, yaml code blocks specify metadata, and sql code blocks handle data interactions. Enables clean, code-light orchestration with minimal setup.
The smallest DuckDB SQL orchestrator on Earth.
Low-code data pipelines for structured and unstructured data. SQL transformations are powered by DuckDB. (Elastic License 2.0, source-available)
Serverless data analytics overlay on top of S3 Data Lakes. (pricing)
Database IDE and migration tool with DuckDB-powered federated SQL. (pricing)
Routes your Snowflake queries to a DuckDB powered warehouse to reduce costs and speed up queries. (pricing)
Desktop SQL client that runs one query across PostgreSQL, MySQL, DuckDB and CSV/Excel files, allowing multi-source JOINs and data versioning powered by a local-first DuckDB engine. (pricing)
Time-series data warehouse built on DuckDB. (commercial license)
A unified SQL query interface and portable runtime to locally materialize (using an embedded DuckDB), accelerate, and query datasets from any database, data warehouse, or data lake. (Spice.ai Cloud)
Virtual warehouse over cloud Parquet. SQL shell, Jupyter/Marimo notebooks, AI natural language queries, and local cache — all powered by DuckDB.
An data mesh platform and high-performance GraphQL backend powered by DuckDB.
Serverless OLAP API/UI built on top of DuckDB with basic ClickHouse API compatibility and MotherDuck support.
An implementation of Snowflake API, enables running queries on Snowflake tables locally with DuckDB without a running warehouse.
Manage with SQL, like for creating topics (tables) and derived topics (materialised views) - all landing on object storage in DuckLake as optimised Parquet files. (proprietary core shipped as binaries, free-forever community tier; SDKs are open source)
Governed, multi-tenant MCP access to customer data. Connects DuckDB (and Snowflake, BigQuery, Postgres) to AI agents as a secure, per-customer MCP server. (pricing)
Census's dataset diffing for incremental syncs is powered by DuckDB. (pricing — Census was acquired by Fivetran)
DuckDB can be used as a caching layer or a data connector in VulcanSQL, a Data API framework for data folks to create REST APIs by writing SQL templates.
Open-source, privacy-first web analytics for traffic, funnels, ecommerce, Search Console, and AI visibility. Runs as a single Go binary with embedded DuckDB. (HitKeep Cloud)
Plug and play analytics for Phoenix applications, powered by DuckDB.
Open-source, self-hosted, warehouse-native product analytics. Runs funnels, retention, and paths on DuckDB (and Postgres, Snowflake, ClickHouse, Databricks).
A browser-based geospatial analysis tool leveraging DuckDB-Wasm. (pricing)
Kepler.gl is a powerful open-source geospatial analysis tool for large-scale data sets, now embeds duckdb wasm to create geospatial layers.
Fast, accurate, open-source geocoding in Python, using DuckDB.
Local, contract-governed data explorer. Federates files and databases (Postgres, Snowflake, BigQuery, Excel, and more) through DuckDB — SQL editor, charts, profiling — then hands AI agents a PII-masked, read-only query surface over MCP. Apache-2.0.
Data quality platform for data engineers, data quality teams and data operations. (Business Source License 1.1, source-available)
No code tool built on top of DuckDB-Wasm and Pyodide that helps build pivot tables from databases of any size with a few clicks.
Preview csv/tsv, json, and Parquet files in the yazi file manager using duckdb. View the raw data, or a "summarized" view with data-types, min, max, avg etc. for all columns.
Blazing-fast & intuitive pivot tables on Parquet, CSV, JSON files and DuckDB tables in the browser based on DuckDB-Wasm. open-source (MIT). Zero install!
Visual Studio Code extension for exploring Parquet files with SQL, powered by DuckDB.
Hex's Dataframe SQL cells are powered by DuckDB. (pricing)
Mode uses DuckDB for their in-memory data engine. (pricing)
A command line tool to efficiently show end-of-life dates for a number of products in your terminal using the endoflife.date API, makes it possible to export the whole endoflife.date database as a fully featured DuckDB file.
Malloy is an experimental language for describing data relationships and transformations. Malloy connects to BigQuery, Snowflake, Trino, and Postgres, and natively supports DuckDB.
Python transpiler that translates between 24 different SQL dialects including DuckDB.
A SAS language engine (DATA step, macros, PROCs) that translates programs to DuckDB SQL — runs natively or fully in the browser on DuckDB-Wasm. (no published license; free to use)
Play the New York Times Connections Puzzle with DuckDB.
Business Intelligence Done Right - ReportBurster uses DuckDB as its in-process analytics database engine, and for larger workloads, supports ClickHouse too. (SSPL v1, source-available)
Virtual warehouse over cloud Parquet. SQL shell, Jupyter/Marimo notebooks, AI natural language queries, and local cache — all powered by DuckDB.
Open-source, self-hosted, warehouse-native product analytics. Runs funnels, retention, and paths on DuckDB (and Postgres, Snowflake, ClickHouse, Databricks).
Database IDE and migration tool with DuckDB-powered federated SQL. (pricing)
Local-first visual ETL/ELT studio. Drag sources, transforms and sinks onto a canvas; it compiles to plain DuckDB SQL and runs entirely on DuckDB. Open source desktop app, with a built-in MCP server for generating and running pipelines from natural language.
Open-source, privacy-first web analytics for traffic, funnels, ecommerce, Search Console, and AI visibility. Runs as a single Go binary with embedded DuckDB. (HitKeep Cloud)
YAML-based data pipeline framework that runs both locally and fully in-browser designed for data engineers, ML teams, and SaaS developers who need flexible, SQL-powered pipelines.
A SAS language engine (DATA step, macros, PROCs) that translates programs to DuckDB SQL — runs natively or fully in the browser on DuckDB-Wasm. (no published license; free to use)
Local, contract-governed data explorer. Federates files and databases (Postgres, Snowflake, BigQuery, Excel, and more) through DuckDB — SQL editor, charts, profiling — then hands AI agents a PII-masked, read-only query surface over MCP. Apache-2.0.
Open-source, DuckDB-powered BI workspace for AI-assisted analysis, SQL, charts, and dashboards.
Desktop SQL client that runs one query across PostgreSQL, MySQL, DuckDB and CSV/Excel files, allowing multi-source JOINs and data versioning powered by a local-first DuckDB engine. (pricing)
Desktop app (Windows, macOS) for a first look at CSV, Excel, JSON and Parquet files: plain-language questions are turned into DuckDB SQL by a language model that runs on the user's own machine. Can work on a computer with no network connection. (pricing)
Self-hosted analytics stack with a semantic model, DAX-inspired queries compiled to DuckDB SQL, DuckLake storage and dashboards.
A fully-functional todo list application that demonstrates DuckDB WASM OPFS (Origin Private File System) persistence using a pure functional programming approach. (no LICENSE file in repo)
a TypeScript-based Docker image containing DuckDB, and a Hono framework REST API with JSON or streaming Arrow responses.
A Python-based server that runs a local DuckDB instance and supports queries over Web Sockets or HTTP, returning data in either Apache Arrow or JSON format.
A Rust-based server that runs a local DuckDB instance and supports queries over Web Sockets or HTTP/HTTPS, returning data in either Apache Arrow or JSON format.
Declarative YAML ETL engine that can run each transformation step on DuckDB or pandas.
A DataFrame API for interacting with DuckDB (and other compute engines).
Lightweight and extensible compatibility layer between dataframe libraries, supports DuckDB.
Implements the PySpark DataFrame API in order to enable running transformation pipelines directly on database engines such as DuckDB.
An extensible framework for linking databases and interactive views.
A unified interface for distributed computing. Fugue executes SQL, Python, Pandas, and Polars code on Spark, Dask and Ray without any rewrites.
A free Python library for fast, accurate data deduplication and record linkage.
Easy-to-use and high-performance JavaScript library for data analysis.
DuckDB Foreign Data Wrapper for PostgreSQL.
A context manager for React and DuckDB-Wasm.
A Python library for downloading and transforming raw OpenStreetMap data into GeoParquet files.
A Python library that turns your dataframe into an interactive UI for data visualization.
An API Framework that heavily relies on the power of DuckDB and DuckDB extensions. Ready to build performant and cost-efficient APIs on top of BigQuery or Snowflake for AI Agents and Data Apps.
A distributed data processing framework by DeepSeek built on DuckDB and 3FS.
PostgreSQL read replica optimized for analytics, using DuckDB.
Rewrite BigQuery, Redshift, Snowflake and Databricks queries into DuckDB-compatible SQL.
An open-source react framework for single-node data analytics powered by DuckDB.
A Python library for efficient data management that wraps the APIs of SQLite and DuckDB and offers a high-level interface analytical tasks that involve fast storage, processing and retrieval of data.
Lightweight DuckDB query-building client for C#.
A lightweight Snowflake emulator built with Go and DuckDB for local development and testing.
Entity Framework Core provider for DuckDB and DuckLake, with LINQ, writes, migrations, bulk ingestion, and Parquet-backed tiered storage for .NET.
A multimodal-native engine for AI workloads, built on a DuckDB fork with Python and SQL interfaces.
Online DuckDB shell powered by DuckDB-Wasm.
DuckDB-Wasm based SQL Workbench for running queries on local or remote data, being able to show data as tables or visually as graphs, and sharing queries via URLs.
Online DuckDB shell powered by DuckDB WASM and Ghostty.
A lightweight JavaScript library that turns SQL code blocks into interactive, browser-based database environments. Powered by DuckDB WASM.
Query your local Parquet, CSV, JSON. Your data will not be sent out of the device you are using.
Embed executable code snippets directly into your product documentation, online course or blog post.
Open-source in-browser DuckDB SQL playground and editor.
Sidequery is a privacy-preserving DuckDB-powered query editor & data exploration tool for local & remote data.
Duck-UI is a web-based interface for interacting with DuckDB with a SQL editor, data import/export, data explorer, query history, theme toggle and keyboard shortcuts.
An open-source react framework for single-node data analytics powered by DuckDB.
Open-source, 100% client-side data exploration tool that enables users to analyze local and remote data using SQL. Zero-copy direct access to local datasets sets PondPilot apart from similar tools. It runs entirely in the browser—no servers, no cloud uploads, and no setup required.
WASM packager for Python-based interactive data apps.
Self-hostable, privacy-focused website analytics.
DuckDB workbench for native & browser.
Privacy-first local data analytics with a modern SQL editor, multi-format support (CSV, Excel, JSON, Parquet), and parameterized saved queries. Available as browser app and Tauri desktop client.
Open-source visual SQL workbench to query local files (CSV/Excel/Parquet/JSON) and remote databases (MySQL/PostgreSQL) in one cross-source JOIN, plus AI text-to-SQL. Browser demo runs on DuckDB-Wasm; the full version self-hosts via Docker.
Free SQL-on-CSV in the browser via DuckDB-Wasm — no signup, no upload, files stay on-device.
Browser-based Parquet viewer, SQL workbench and converter powered by DuckDB-Wasm. Fully client-side — files never leave your device.
Tabular data visualizer and DuckDB SQL editor on DuckDB-Wasm: open CSV/Parquet/JSON/Arrow locally, chart results with a grammar-of-graphics VISUALIZE syntax, and embed it anywhere via npm component or a one-line iframe.
Browser-based, read-only SQL playground powered by DuckDB-Wasm, with synthetic structured-event datasets.
Browser-based data explorer on DuckDB-Wasm: open CSV, Excel, JSON, Parquet, Arrow, Avro, SQLite and DuckDB files up to 1GB, chain filter/join/pivot/aggregate steps that each emit DuckDB SQL, write SQL directly, chart, and export. Files never leave the device.
Browser-based Parquet viewer on DuckDB-Wasm. Reads the schema, row groups, compression codecs and column statistics, runs read-only SQL over the file and exports results as CSV. A JSONL/NDJSON viewer uses the same engine. Files are opened through a browser file handle and are not uploaded.
Browser-based Parquet viewer on DuckDB-Wasm. Shows the schema, row groups and column chunk statistics, runs SQL over the file, and exports to CSV or JSON. Files are read locally and are not uploaded.
The DuckDB IDE (TUI) for your terminal.
A free SQL tool specialized for data analysts. It runs on every operating system and allows easy browsing of tables and charting of results.
Free DuckDB SQL Tools for VS Code IDE. Premium version available with advanced features.
Free open-source VSCode extension to query and explore your DuckDB databases with latest DuckDB support.
DBeaver is a universal database access and development tool that can be used to connect almost any type of database.
Paid SQL IDE by JetBrains that supports many different database technologies, including DuckDB.
A fast viewer for CSV/Parquet files and DuckDB/SQLite, based on Tauri.
CLI for DuckDB, LibSQL, MariaDB, MySQL, PostgreSQL, SQLite3 and SQL Server.
A lightweight, commercial SQL IDE that supports different DBMS, including DuckDB. The focus is on performance and special DBMS features.
Simple easy-to-use database manager, supports DuckDB, PostgreSQL, MySQL, SQL Server, SQLite etc.
A database TUI with experimental support for DuckDB.
Unified data & AI platform that lets you work with SQL and Python with native support for DuckDB.
Run SQL in Jupyter/IPython via a %sql and %%sql magics, with support for DuckDB.
DuckDB native %dql and %%dql magics for Jupyter/IPython.
The Modern Codex for Local Data Analysis. A high-performance, local-first IDE built specifically for DuckDB.
A Rust TUI for fast, in-depth analytics on large datasets (CSV, JSON, Parquet, Excel, SQLite) powered by DuckDB.
A fast, free, cross-platform tabular data viewer application powered by DuckDB.
Paid desktop database workbench with DuckDB support, including isolated transactions, transactional migrations, physical backups, and remote import from Parquet and CSV.
Read-only DuckDB database file viewer for JetBrains IDEs: schema tree, paged tables of any size, constant memory. Sibling Lens viewers cover SQLite, Parquet, Excel and Jupyter files.
Paid cross-platform desktop client supporting DuckDB alongside two dozen other databases, including relational, document, key-value, vector and search engines.
Browser-based SQL IDE that opens DuckDB database files alongside PostgreSQL, MySQL, ClickHouse and other engines, with a schema explorer and JSON EXPLAIN plans. MIT licensed, runs as a container or Helm chart.
Free, open-source SQL editor and database manager with native support for DuckDB, MySQL, Postgres, SQLite, SQL Server, and more.
Monte Carlo simulation of the NBA season, leveraging Meltano, dbt, DuckDB and Evidence.
Open-source and local-friendly data platform to collaborate on Open Data using DuckDB, Dagster, dbt, and Quarto.
Daily dumps of endoflife.date data.
Curated football datasets from Transfermarkt.
A search engine for DuckDB that uses embedding vectors to find similar documents.
Live dashboard of PyPI downloads using DuckDB, dbt, Evidence and MotherDuck with code source to build your own.
Python package that supports data preparation for object-centric process mining.
DuckDB powered WASM app where you can search how much students spend on books at Georgia State University.
A Slack data analysis agent powered by DuckDB and Claude Code.
A minimal DuckDB WASM build for browsers and serverless environments like Cloudflare Workers.
Archive a lifetime of email and chat. Offline search, analytics, and AI query over your full message history. Powered by DuckDB.
Generates random data, allowing to define dependencies between individual fields and varying/definable distribution of field values.
Extract all Ukrainian POIs from Overture Maps releases into CSV or Parquet with a single DuckDB query.
End-to-end pipeline over seven public data sources in a single DuckDB file, using dlt, dbt, Polars, Dagster and Evidence. Rebuilds from live sources on every push and publishes the DuckDB file and per-table Parquet monthly.
Detective game where each case file is a set of Parquet tables queried with SQL, running DuckDB-WASM in the browser.
Self-hostable research workspace that compiles boolean literature-search queries to SQL and runs them offline with DuckDB over versioned Parquet snapshots of a scholarly corpus.
Data-quality grades for 2,400+ public transit (GTFS) feed records. DuckDB builds the cross-agency query layer and a public Parquet table that DuckDB can query over HTTP, and the site's SQL page runs DuckDB-WASM in the browser.
Resumable local Product Hunt catalogue and analytics pipeline that normalizes records into DuckDB and exports verified Parquet snapshots.
DuckDB dbt adapter.
Extract and load data from APIs to DuckDB using dlt.
Load data to DuckDB based on Singer spec.
Load data to DuckDB with Airbyte.
Run queries with DuckDB to schedule data transformations and process automations and run event-driven anomaly detection pipelines.
Enables SQL-based stream processing, powered by DuckDB.
This plugin provides support for interacting with SQL databases in Nextflow scripts.
The platform for customizing AI from enterprise data. MindsDB integrates with DuckDB, making data from DuckDB accessible to a diverse range of AI/ML models.
A CLI tool to convert SQLite database to DuckDB.
NoSQL Database Connector for R, providing a common API across Elasticsearch, CouchDB, MongoDB, SQLite, PostgreSQL, and DuckDB.
Drop-in replacement for dplyr in R that uses DuckDB for performance.
In-Memory Analytics for Kafka using DuckDB.
DuckDB Tableau connector.
DuckDB Power Query Custom Connector.
Metabase DuckDB Driver shipped as 3rd party plugin.
Excel addin to run DuckDB queries in Excel.
Allows connecting to a DuckDB database or a MotherDuck-hosted DuckDB database through a GraphQL API.
Allows to create Virtual Knowledge Graphs directly from DuckDB.
Type safe querying of DuckDB (and many other RDBMS) from Java. A transpiler from and to DuckDB is also available.
Use native DuckDB SQL of any complexity directly & type-safely in Java source with comprehensive IntelliJ support.
SAS/ACCESS engine support for DuckDB.
Teradata connector.
Supports reading from DuckDB databases using JDBC.
A semantic layer with DuckDB integration.
Connect to Cloudflare R2 Data Catalog with DuckDB.
Lightweight event ingestion that stores JSON events as Parquet in R2, queryable directly from DuckDB.
Excel/VBA integration for DuckDB using the native C API through a lightweight DLL bridge. Supports Range/Array ingestion, dictionary lookups, Parquet/CSV/JSON workflows, SQLite/PostgreSQL connectivity, and Access-to-DuckDB migration.
Open-source semantic sidecar that compiles YAML semantic models to optimized SQL across 8 engines including DuckDB. Ships with an ob-duckdb driver, REST + Arrow Flight SQL + Postgres wire surfaces, and a baked-in DuckDB quickstart on Colab.
Native Excel XLL add-in for DuckDB with parameter binding, async execution and Excel range table functions.
Data pipeline CLI that runs SQL and Python transformations with built-in quality checks, using DuckDB as one of its supported platforms.
CLI tool to copy data between databases and SaaS sources, with DuckDB supported as both a source and a destination.
Polyglot data loader for ETL, warehousing and more, based on dlt. Supports 140+ source/destination adapters. Derived from ingestr v0 (MIT).
Data quality monitoring CLI that profiles tables and detects anomalies (volume, freshness, schema drift, NULL rates, distributions) without configuration; DuckDB is one of its connectors and backs its built-in demo.
Fully managed DBaaS based in PostgreSQL integrated with DuckDB.
A serverless cloud data warehouse powered by DuckDB.
A server wrapping DuckDB with MySQL and PostgreSQL wire protocol support.
PostgreSQL extension embedding DuckDB-in-PostgreSQL for fast on-disk and remote object storage analytics from Postgres. Built as a Foreign Data Wrapper with full query pushdown to DuckDB. Integrates easily with ParadeDB.
DuckDB-powered PostgreSQL for high-performance apps & analytics.
A PostgreSQL extension that adds native column store tables with DuckDB.
A C++ implementation of the Arrow Flight SQL protocol that runs in a client-server setup with DuckDB or SQLite as backends.
A Go-based implementation of a DuckDB Arrow Flight SQL Server.
DuckDB CLI client for the Termux Android terminal emulator.
pg_lake integrates Iceberg and data lake files into Postgres. Uses DuckDB to execute queries.
A MySQL branch originated from Alibaba Group. Integrates DuckDB as a native storage engine.
Managed analytical platform pairing DuckDB compute with Iceberg storage on S3. Postgres wire protocol, CLI, and Python API.
Managed DuckDB hosting with a browser query console, REST API and CLI, alongside 17 other database engines under one account. Databases scale to zero when idle and wake on connect.
Aggregate Parquet, CSV and JSON in object storage from SQL, with DuckDB embedded in MySQL.
A zero-copy data integration between Apache Arrow and DuckDB.
For reading Avro files.
For handling AWS credentials.
For using the Azure Blob storage.
For Delta Lake support.
For DuckLake support.
To support full-text search.
For reading Iceberg tables.
For storing and handling IPv4 and IPv6 Internet addresses.
For reading and writing JSON data.
To read from and write to MySQL databases.
For reading and writing Parquet data.
To read from and write to PostgreSQL databases.
Enables geospatial processing.
To read from and write to SQLite databases.
Add support for vector similarity search.
Integrates DuckDB with DeepSeek 3FS distributed file system.
Embeds AI agents such as Claude Code inside of DuckDB via Agent Client Protocol.
Integrates DuckDB with Google BigQuery, allowing direct querying and management of BigQuery datasets.
Adds a read caching layer to duckdb filesystem to improve query performance and reduce egress cost.
A Preloads table data blocks into the buffer pool or OS page cache, inspired by PostgreSQL's pg_prewarm extension.
ClickHouse SQL Dialect macros for DuckDB.
Cryptographic hash functions and HMAC.
Enhanced HTTP file system with connection pooling, HTTP/2 support, and asynchronous I/O operations.
Fully local data canvas and dashboarding app within DuckDB.
Distributed execution for DuckDB queries.
Add supports for SQL/PGQ (Property Graph Queries) introduced in the SQL:2023 standard.
Query Elasticsearch indices directly using SQL.
Evaluates the Rhai scripting language as part of SQL.
Performs fuzzy string matching for autocompletion.
A DuckDB extension for working with Kaggle datasets.
GPU-accelerated plain DuckDB SQL on Apple Silicon Metal and NVIDIA CUDA: GROUP BY, joins, filters and top-k run on the GPU when measured faster, with the same answers.
Read and write Google Sheets using SQL.
Adds support for the H3 discrete global grid system.
Navigate and explore the local filesystem using SQL.
DuckDB HTTP API Server and Query Interface.
A DuckDB extension for in-database inference.
Linearization/Delinearization, Z-Order, Hilbert and Morton Curves.
Parsing, extracting, and analyzing domains, URIs, and paths with ease.
I/O observability for DuckDB filesystems with latency statistics and external file cache access insights.
A DuckDB extension for graph data analytics.
Read block-indexed PFC-compressed JSONL logs with timestamp filtering — 25% smaller than gzip with minimal S3 egress.
Run PRQL commands directly within DuckDB.
Read Microsoft PST files in-place with rich schemas for emails, contacts, appointments, tasks, and more.
Caches query conditions to improve performance for repeated-query workloads.
A set of aggregation functions and data scanners on financial data.
Allows shell commands to be used for input and output.
Statistics for tabular and clinical data: descriptive tables (table_one), linear models with robust/clustered standard errors, meta-analysis, bootstrap, and a grammar-of-graphics VISUALIZE clause that turns queries into Vega-Lite charts.
ULID data type for DuckDB. A ULID is similar to a UUID except that it also contains a timestamp component.
Implements Measures in SQL paper as a DuckDB extension for centralized metric definitions / en embedded semantic layer.
SQLAlchemy driver for DuckDB.
A Zig & Nix toolkit template for building extensions against multiple versions of DuckDB using Zig, C or C++.
DuckDB extension to read JFR (Java Flight Recorder) files directly.
Plugin for querying encoded protobuf messages (both sequences and individual messages per file).
DuckDB extension to allow running SQL on arbitrary data sources.
DuckDB SAP connector using RFC, ODP, or BICS.
Scan DuckDB tables in Kùzu, an embeddable property graph database management system.
Integrate Lance (modern columnar data format for ML implemented in Rust) with DuckDB.
DuckDB extension to read data directly from databases supporting the ODBC interface.
Plugin for reading DuckDB spatial tables in QGIS software.
Proof-of-concept extension combining the delta extension with Unity Catalog.
Integrate language model (LLM) capabilities directly into your queries and workflows.
ERPL Web is a DuckDB extension that connects API-based ecosystems via standard interfaces like OData, GraphQL, and REST.
The infamous DuckDB quack extension rewritten in C and built with Zig. Proof that you can develop DuckDB extensions without drowning in boilerplate.
Build native DuckDB extensions in C#.
A template for developing DuckDB extensions in Zig using DuckDB's C API.
A Go library that mounts io/fs file systems as DuckDB virtual file systems, sandboxing all I/O through the Go runtime.
Repository that contains DuckDB extensions on GitHub. Refreshed daily.
Hannes Mühleisen.
Hannes Mühleisen and Mark Raasveldt.
Hannes Mühleisen and Mark Raasveldt.
Hannes Mühleisen.
Hannes Mühleisen.
Hannes Mühleisen & Mark Raasveldt.
Pedro Holanda & Sam Ansmink.
Hannes Mühleisen.
Mark Needham.
Mehdi Ouazza.
Hannes Mühleisen.
Mark Raasveldt.
Hannes Mühleisen.
Series in Disseminate, the Computer Science Research Podcast, with host Jack Waudby.
Adding concurrent read/write to DuckDB with Arrow Flight.
Fast, free, and open-source Modern Data Stack deployed on a laptop using the combination of DuckDB, Meltano, dbt, and Apache Superset.
How DuckDB can transform data, mask sensitive PII information, detect anomalies in event-driven workflows, and streamline reporting use cases.
What are the key differences between them, and when to choose each of these options.
For Nix users and Zig developers familiar with DuckDB looking to extend its capabilities with custom extensions.
Example project using DuckDB to persist API data, but also explains how to use DuckDB as a versatile data manipulation tool in data wrangling scripts.
How DuckDB can provide a view over data stored in S3.
How to set up DuckDB and how to work with extensions in an offline (and potentially sensitive) environment.
Improved DuckDB support in JetBrains' Datalore collaborative data science platform
Demo of using Cloudflare R2 hosting and a WASM DuckDB application to store and query data
An Arrow Flight SQL Datalake Service Built on DuckDB + DuckLake
Why a static site ships the single-threaded build: SharedArrayBuffer, COOP/COEP, and the cross-origin isolation tradeoff.
How to go from 0 to a production-ready DuckDB extension for Outlook PSTs, including table functions, MAPI schema serialization, projection/statistics pushdown, concurrent planning, and late materialization.
DuckDB in Action will show you how to quickly get your hands dirty with DuckDB.
A practical guide for accelerating your data science, data analytics, and data engineering workflows.
Building an analytics stack on one machine with DuckDB, Parquet and Arrow: performance, partitioning, data quality and orchestration, with runnable code. Chapter one is free to read online.
VoltAgent/awesome-openclaw-skills
The awesome collection of OpenClaw skills. 5,400+ skills filtered and categorized from the official OpenClaw Skills Registry.🦞
awesome-dsh-plugin/awesome-dsh-plugin
A curated list of plugins for DeepSeek Harness (dsh) · DeepSeek Harness 插件精选列表
Kristories/awesome-guidelines
Programming style, best practices, and coding conventions.
sindresorhus/awesome
😎 Awesome lists about all kinds of interesting topics [NOTE: Pull requests are temporarily disabled until I have a chance to catch up with the existing ones]
ai-boost/awesome-prompts
Curated list of chatgpt prompts from the top-rated GPTs in the GPTs Store. Prompt Engineering, prompt attack & prompt protect. Advanced Prompt Engineering papers.
matiassingers/awesome-readme
A curated list of awesome READMEs