Open highlighted repo slot
Put your repository first
Promote a GitHub repo at the top of Awesome repository list views for 7 days.
GitHub projects from awesome lists
Search names, descriptions, topics, tags, and stacks, then tune results by ecosystem, freshness, health, and cross-list signal.
Open highlighted repo slot
Promote a GitHub repo at the top of Awesome repository list views for 7 days.
The unified workspace where open-source models get things done for you.
Python SQL Parser and Transpiler
AliSQL is a MySQL branch originated from Alibaba Group. Fetch document from Release Notes at bottom.
🐸 a database tool for the terminal
A lightweight data processing framework built on DuckDB and 3FS.
Preswald is a WASM packager for Python-based interactive data apps: bundle full complex data workflows, particularly visualizations, into single files, runnable completely in-browser, using Pyodide, DuckDB, Pandas, and Plotly, Matplotlib, etc. Build dashboards, reports, and notebooks that run offline, load fast, and share like a document.
Embedded property graph database built for speed. Vector search and full-text search built in. Implements Cypher.
ingestr is a CLI tool to copy data between any databases with a single command seamlessly.
Scalable and efficient data transformation framework - backwards compatible with dbt.
DuckDB-powered Postgres for high performance apps & analytics.
Add a real-time analytics node to your operational database. Spice is a portable, accelerated SQL query, search, and LLM-inference engine in Rust for data-grounded AI apps and agents.
DuckLake is an integrated data lake and catalog format
The fastest business intelligence tool for humans and agents.
Manifold is a Java compiler plugin, its features include Metaprogramming, Properties, Extension Methods, Operator Overloading, Templates, a Preprocessor, and more.
Fast, accurate and scalable probabilistic data linkage with support for multiple SQL backends
A unified interface for distributed computing. Fugue executes SQL, Python, Pandas, and Polars code on Spark, Dask and Ray without any rewrites.
Archive a lifetime of email and chat. Offline search, analytics, and AI query over your full message history. Powered by SQLite and DuckDB
Real-time analytics on Postgres tables
Lightweight and extensible compatibility layer between dataframe libraries!
Build data pipelines with SQL and Python, ingest data from different sources, add quality checks, and build end-to-end flows.
pg_lake: Postgres with Iceberg and data lake access
Open-source Snowflake & Fivetran alternative, with Postgres compatibility.
visual data prep powered by python
dbt adapter for DuckDB
Open-source ETL/ELT you deploy on your own servers or cloud. Built on DuckDB: no-code/low-code visual pipelines or SQL, 385 components, dbt, CDC, data quality, reverse ETL, lineage, MCP for AI agents. No vendor cloud, no per-row billing.
Visualize and share your data. All in SQL. Powered by DuckDB.
Ergonomic bindings to duckdb for Rust
DuckDB for streaming data
Bindings and ADO.NET Provider for DuckDB
Open, SQL-native time-series database for telemetry you need to keep. 34M+ records/sec ingestion, 8M+ rows/sec queries. InfluxDB Line Protocol and Telegraf compatible. Open Parquet on your storage. Single binary. S3/Azure native. Air-gap ready. AGPL-3.0.