开源精选 · 向量数据库
最好用的向量数据库
每日更新。60 个项目,按 star 排序,并核对 fork 与维护活跃度。
向量数据库 镇场榜
成熟且仍在更新的项目。GitHub 上标注了 vector-database, vector-search。
-
For developers, who are building real-time data-driven applications, Redis is the preferred, fastest, and most feature-rich cache, data structure server, and document and vector query engine.
-
Stop renting your intelligence. Own it with AnythingLLM. Everything you need for a powerful local-first agent experience
-
A lightning-fast search engine API bringing AI-powered hybrid search to your sites and applications.
-
Ready-to-run cloud templates for RAG, AI pipelines, and enterprise search with live data. 🐳Docker-friendly.⚡Always in sync with Sharepoint, Google Drive, S3, Kafka, PostgreSQL, real-time data APIs, and more.
-
LlamaIndex is the document processing platform for AI
-
Milvus is a high-performance, cloud-native vector database built for scalable vector ANN search
-
📑 PageIndex: Document Index for Vectorless, Reasoning-based RAG
-
Qdrant - High-performance, massive-scale Vector Database and Vector Search Engine for the next generation of AI. Also available in the cloud https://cloud.qdrant.io/
-
Open Source AI Platform - AI Chat with advanced features that works with every LLM
-
Open-source LLM knowledge platform: turn raw documents into a queryable RAG, an autonomous reasoning agent, and a self-maintaining Wiki.
-
A modern replacement for Redis and Memcached
-
Cognee is the open-source AI memory platform for agents. Give your AI agents persistent long-term memory with small models for free
-
This repository showcases various advanced techniques for Retrieval-Augmented Generation (RAG) systems. Each technique has a detailed notebook tutorial.
-
The #1 AI Harness for Building Resumes, PDFs, Cover Letters & more, locally with 100+ LLMs support.
-
TencentDB Agent Memory is a team-level memory hub for AI Agents - turning conversations, docs, and code into four reusable memory assets (Chat Memory, Skill, LLM-Wiki, Code-Graph) that are governed, shared, and equipped across agents and frameworks.
-
Open Source alternative to Algolia + Pinecone and an Easier-to-Use alternative to ElasticSearch ⚡ 🔍 ✨ Fast, typo tolerant, in-memory fuzzy Search Engine for building delightful search experiences
-
A vector index built on TurboQuant, written in Rust with Python bindings
-
Weaviate is an open-source vector database that stores both objects and vectors, allowing for the combination of vector search with structured filtering with the fault tolerance and scalability of a cloud-native database.
-
Memory layer for AI Agents. Replace complex RAG pipelines with a serverless, single-file memory layer. Give your agents instant retrieval and long-term memory.
-
LangChain4j is an idiomatic, open-source Java library for building LLM-powered applications on the JVM. It offers a unified API over popular LLM providers and vector stores, and makes implementing tool calling (including MCP support), agents and RAG easy. It integrates seamlessly with enterprise Java frameworks like Quarkus and Spring Boot.
-
[MLsys2026 Best Paper]: https://arxiv.org/abs/2506.08276. RAG on Everything with LEANN. Enjoy 97% storage savings while running a fast, accurate, and 100% private RAG application on your personal device.
-
💡 All-in-one AI framework for semantic search, LLM orchestration and language model workflows
-
Code search MCP for Claude Code. Make entire codebase the context for any coding agent.
-
Developer-friendly OSS embedded retrieval library for multimodal AI. Search More; Manage Less.
-
Refine high-quality datasets and visual AI models
-
🌌 A complete search engine and RAG pipeline in your browser, server or edge network with support for full-text, vector, and hybrid search in less than 2kb.
-
OceanBase is the unified distributed database for the AI era - open-source, multi-model, one engine for your most demanding workloads.
-
Data Agent Ready Warehouse : One for Analytics, Search, AI, Python Sandbox. - rebuilt from scratch. Unified architecture on your S3.
-
One Postgres for your application data, full-text search, vector retrieval, and aggregations. Home of the pg_search extension.
-
MariaDB server is a community developed fork of MySQL server. Started by core members of the original MySQL team, MariaDB actively works with outside developers to deliver the most featureful, stable, and sanely licensed open SQL server in the industry.
-
Open Source Deep Research Alternative to Reason and Search on Private Data. Written in Python.
-
Semantic cache for LLMs. Fully integrated with LangChain and llama_index.
-
The AI search platform
-
Open-source framework for building agentic apps in JavaScript, Go, Dart, and Python, built and used in production by Google
-
A query and indexing engine for Redis, providing secondary indexing, full-text search, vector similarity search and aggregations.
-
HelixDB is an OLTP graph database with native vector and full-text search built in Rust on Object Storage.
-
Go Open Source, Distributed, Simple and efficient Search Engine
-
MineContext is your proactive context-aware AI partner(Context-Engineering+ChatGPT Pulse)
-
A distributed approximate nearest neighborhood search (ANN) library which provides a high quality vector index build, search and distributed online serving toolkits for large scale vector search scenario.
-
The AI-native database built for LLM applications, providing incredibly fast hybrid search of dense vector, sparse vector, tensor (multi-vector), and full-text.
向量数据库 新秀榜
近 90 天新建、涨势最猛的项目。
-
Local-first search across your workspace, built for humans and AI agents.
-
Open-source infrastructure that turns scattered SKILL.md files into curated, retrieval-ready agent-skill corpora - with retrieval and evaluation tooling included.
-
Encrypted, fully offline agentic memory. One click install, GUI w/ memory map, all OS and agents. Superior memory creation, storage and retrieval.
-
A free, self-paced 24-week AI engineering course: Python, machine learning, LLMs, RAG, fine-tuning, agents and MCP, Azure and Vertex and Bedrock, and Databricks. 43 runnable notebooks, one continuous case study. MIT licensed, no signup. By Zorost Intelligence AI Lab.
-
Not another ChatGPT wrapper. This is a local-first tool that turns your notes, web pages, PDFs and audio/video into a personal knowledge graph - running entirely on your own machine.
-
Engineering deterministic, production-grade systems around non-deterministic LLMs - FSM, durable execution, retries, DAGs, agent runtimes, model routing, edge inference, RAG, memory, multi-agent orchestration, security, and observability. 14 runnable proof-of-concept phases.
-
Spotlight-style local search for everything you saved and forgot: GitHub stars, local files, images, and bookmarks. Privacy-first, no full-disk scanning, fully on-device.
-
A dsh plugin that periodically scans reddapi.dev for new Reddit leads matching your one-sentence ICP, dedupes what you've already seen, and writes a dated markdown report. Also forwards 6 read-only Reddit search/lookup MCP tools.
-
🧠 The memory that dreams - cross-session memory for DeepSeek Harness. Offline & private, auto-consolidates in its sleep (autoDream), visualized in a memory panel.
-
Drop-in AI memory layer with 2x faster retrieval and 10x lower cost. Fully compatible with Mem0 API. Migrate in 5 minutes without any code changes. Self-host for free.
-
Billion-scale embedded vector database built entirely on Parquet and Arrow.
-
Comparison of memory systems for agents
-
Local AI search for podcast and video archives on Apple Silicon: Whisper speech, Apple Vision OCR and SigLIP-2 visual search fused into one ranked list, with a trim editor and FCPXML export. FastAPI, vanilla JS, Tauri. Source-visible, all rights reserved.
-
Agentic event-venue operator demo built on MongoDB Atlas. Showcases three-layer memory (long-term, short-term, shared), hybrid retrieval with Voyage AI multimodal embeddings, optional Langfuse traces, and persona-segmented agent decisions during a rain delay scenario at a.
-
Local document Q&A in your terminal - powered by on-device LLMs.
-
Shared, persistent memory for AI agents. Self-hosted MCP server with semantic search, vector RAG, and live updates. Works with Claude, Cursor, Codex, and any MCP client.
-
A zero-to-100 learning path for applied AI engineering - RAG, embeddings, vector search, agents, MCP, and the production engineering around them. 56 pages, built as a searchable site.
-
Zero downtime embedding upgrades
-
Open-source vector database and sub-microsecond key-value store in one Go engine - embed it as a library, run it standalone, or replicate it across a Raft cluster. HNSW/IVF/Vamana indexes, quantization, hybrid dense+sparse and BM25 search, WASM stored procedures. Apache-2.0.
-
Gomaa - Autonomous Agent Memory OS. Persistent memory system for AI agents with Obsidian vault integration, hybrid RRF search, knowledge graphs, security gates, and MCP server.
这份榜单怎么来的
候选来自 GitHub 话题标签 vector-database, vector-search。star 数与周增长取自 GitHub 官方的 star 历史接口,因此数字与 GitHub 自己的口径一致。榜单收录活跃、话题数量至多 20 个的仓库。某个项目的 star 数相对 fork 数明显偏高时会被标注,可结合仓库活动进一步了解原因。