🔎 Open source

RAG & Vector Search

Vector databases and retrieval stacks for pointing a model at your own documents.

12repositories
413kstars combined
AN

anything-llm

Mintplex-Labs

Stop renting your intelligence. Own it with AnythingLLM. Everything you need for a powerful local-first agent experience

★ 64k · JavaScript · MITagent-computeragent-harnessagent-orchestration
updated today View on GitHub →
LL

llm-app

pathwaycom

Ready-to-run cloud templates for RAG, AI pipelines, and enterprise search with live data. 🐳Docker-friendly.⚡Always in sync with Sharepoint, Google Drive, S3, Kafka, PostgreSQL, real-time data APIs, and more.

★ 59k · Jupyter Notebook · MITchatbothugging-facellm
updated 17 days ago View on GitHub →
ME

meilisearch

meilisearch

A lightning-fast search engine API bringing AI-powered hybrid search to your sites and applications.

★ 59k · Rustaiapiapp-search
updated today View on GitHub →
MI

milvus

milvus-io

Milvus is a high-performance, cloud-native vector database built for scalable vector ANN search

★ 45k · Go · Apache-2.0annscloud-nativediskann
updated today View on GitHub →
PA

PageIndex

VectifyAI

📑 PageIndex: Document Index for Vectorless, Reasoning-based RAG

★ 34k · Python · MITagentic-aiagentsai
updated today View on GitHub →
QD

qdrant

qdrant

Qdrant - High-performance, massive-scale Vector Database and Vector Search Engine for the next generation of AI. Also available in the cloud https://cloud.qdrant.io/

★ 33k · Rust · Apache-2.0ai-searchai-search-engineembeddings-similarity
updated today View on GitHub →
CO

cognee

topoteretes

Cognee is the open-source AI memory platform for agents. Give your AI agents persistent long-term memory across sessions with a self-hosted knowledge graph engine.

★ 29k · Python · Apache-2.0agent-memoryagent-skillsai
updated today View on GitHub →
RA

RAG_Techniques

NirDiamant

This repository showcases various advanced techniques for Retrieval-Augmented Generation (RAG) systems. Each technique has a detailed notebook tutorial.

★ 29k · Jupyter Notebookagentic-ragaiembeddings
updated 8 days ago View on GitHub →
WE

weaviate

weaviate

Weaviate is an open-source vector database that stores both objects and vectors, allowing for the combination of vector search with structured filtering with the fault tolerance and scalability of a cloud-native database​.

★ 17k · Go · BSD-3-Clauseapproximate-nearest-neighbor-searchgenerative-searchgrpc
updated today View on GitHub →
ME

memvid

memvid

Memory layer for AI Agents. Replace complex RAG pipelines with a serverless, single-file memory layer. Give your agents instant retrieval and long-term memory.

★ 16k · Rust · Apache-2.0aicontextembedded
updated 8 days ago View on GitHub →
ZV

zvec

alibaba

A lightweight, lightning-fast, in-process vector database

★ 15k · C++ · Apache-2.0agent-skillsdbembedded
updated today View on GitHub →
TX

txtai

neuml

💡 All-in-one AI framework for semantic search, LLM orchestration and language model workflows

★ 13k · Python · Apache-2.0agentsaiai-agents
updated today View on GitHub →

Where this comes from: GitHub search: topic:vector-database stars:>1000, sorted by stars. Last refreshed 22 July 2026. Nothing on this page is a paid placement.