Open highlighted repo slot
Put your repository first
Promote a GitHub repo at the top of Awesome repository list views for 7 days.
GitHub projects from awesome lists
Search names, descriptions, topics, tags, and stacks, then tune results by ecosystem, freshness, health, and cross-list signal.
Open highlighted repo slot
Promote a GitHub repo at the top of Awesome repository list views for 7 days.
Stop renting your intelligence. Own it with AnythingLLM. Everything you need for a powerful local-first agent experience
A lightning-fast search engine API bringing AI-powered hybrid search to your sites and applications.
Ready-to-run cloud templates for RAG, AI pipelines, and enterprise search with live data. 🐳Docker-friendly.⚡Always in sync with Sharepoint, Google Drive, S3, Kafka, PostgreSQL, real-time data APIs, and more.
LlamaIndex is the leading document agent and OCR platform
Milvus is a high-performance, cloud-native vector database built for scalable vector ANN search
📑 PageIndex: Document Index for Vectorless, Reasoning-based RAG
Qdrant - High-performance, massive-scale Vector Database and Vector Search Engine for the next generation of AI. Also available in the cloud https://cloud.qdrant.io/
Cognee is the open-source AI memory platform for agents. Give your AI agents persistent long-term memory across sessions with a self-hosted knowledge graph engine.
This repository showcases various advanced techniques for Retrieval-Augmented Generation (RAG) systems. Each technique has a detailed notebook tutorial.
Weaviate is an open-source vector database that stores both objects and vectors, allowing for the combination of vector search with structured filtering with the fault tolerance and scalability of a cloud-native database.
Memory layer for AI Agents. Replace complex RAG pipelines with a serverless, single-file memory layer. Give your agents instant retrieval and long-term memory.
A lightweight, lightning-fast, in-process vector database
💡 All-in-one AI framework for semantic search, LLM orchestration and language model workflows
[MLsys2026 Best Paper]: https://arxiv.org/abs/2506.08276. RAG on Everything with LEANN. Enjoy 97% storage savings while running a fast, accurate, and 100% private RAG application on your personal device.
LangChain4j is an idiomatic, open-source Java library for building LLM-powered applications on the JVM. It offers a unified API over popular LLM providers and vector stores, and makes implementing tool calling (including MCP support), agents and RAG easy. It integrates seamlessly with enterprise Java frameworks like Quarkus and Spring Boot.
Developer-friendly OSS embedded retrieval library for multimodal AI. Search More; Manage Less.
🌌 A complete search engine and RAG pipeline in your browser, server or edge network with support for full-text, vector, and hybrid search in less than 2kb.
Data Agent Ready Warehouse : One for Analytics, Search, AI, Python Sandbox. — rebuilt from scratch. Unified architecture on your S3.
One Postgres for your application data, full-text search, vector retrieval, and aggregations. Home of the pg_search extension.
Deeplake is AI Data Runtime for Agents. It provides serverless postgres with a multimodal datalake, enabling scalable retrieval and training.
Open Source Deep Research Alternative to Reason and Search on Private Data. Written in Python.
Postgres with GPUs for ML/AI apps.
Open-source framework for building agentic apps in JavaScript, Go, Dart, and Python, built and used in production by Google
A query and indexing engine for Redis, providing secondary indexing, full-text search, vector similarity search and aggregations.
HelixDB is an OLTP graph database with native vector and full-text search built in Rust on Object Storage.
The AI-native database built for LLM applications, providing incredibly fast hybrid search of dense vector, sparse vector, tensor (multi-vector), and full-text.
Local persistent memory store for LLM applications including claude desktop, github copilot, codex, antigravity, etc.
CrateDB is a distributed and scalable SQL database for storing and analyzing massive amounts of data in near real-time, even with complex queries. It is PostgreSQL-compatible, and based on Lucene.
Local persistent memory store for LLM applications including claude desktop, github copilot, codex, antigravity, etc.
Jupyter Notebooks to help you get hands-on with Pinecone vector databases