Open source vector databases, installed and queried
Vector databases you can run yourself. The list is ordered by whether each project came up when Argusic installed it fresh, with the recording of that attempt one click away.
Tested between and . Each row shows its own test date; a project can change after that day.
30 of 32 tested projects run. 52 more waiting for a test.
In short: 25 of the 32 tested projects started as-is on a fresh machine: anything-llm, milvus, PageIndex, cognee, zvec, langchain4j, LongMemory, and crate, and 17 more. 5 more started once a stand-in replaced a service they expect, such as a database: claude-context, MineContext, ava-whatsapp-agent-course, LEANN, and deep-searcher. 2 could not be verified: autoflow and orbit; the log shows where each one stopped.
Measured by Argusic on a fresh machine every time. Every number links to its evidence.
| # | project | verdict | Argusic Score | language | stars | tested on |
|---|---|---|---|---|---|---|
| 1 | anything-llm | Runs | 100 / 100 | JavaScript | 66,820 | |
Stop renting your intelligence. Own it with AnythingLLM. Everything you need for a powerful local-first agent experience What the test found: AnythingLLM installs and runs in development mode on port 3001; the test suite passes all 1119 tests across 62 suites covering server and collector modules. 17 minutes. | ||||||
| 2 | milvus | Runs | 100 / 100 | Go | 46,341 | |
Milvus is a high-performance, cloud-native vector database built for scalable vector ANN search What the test found: pymilvus 3.0.2 with milvus-lite 3.2.1 runs embedded vector database operations (create, insert 1000 vectors, search, query, delete, drop) in the container. Go toolchain 1.26.6, Rust 1.84, and Conan 2.25.1 are installed. The C++ core cannot build without root for system library packages. 31 minutes. | ||||||
| 3 | PageIndex | Runs | 100 / 100 | Python | 38,976 | |
📑 PageIndex: Document Index for Vectorless, Reasoning-based RAG What the test found: PageIndex SDK v0.2.10 installed in a venv, all 796 unit tests pass, the CLI parses correctly, and the package imports without error. 3 minutes. | ||||||
| 4 | cognee | Runs | 100 / 100 | Python | 31,642 | |
Cognee is the open-source AI memory platform for agents. Give your AI agents persistent long-term memory with small models for free What the test found: cognee 1.6.1 installs, builds, starts, and runs: the local LLM-free remember/recall pipeline produces search results from ingested text, the CLI demo loads a bundled knowledge graph, the FastAPI server answers HTTP 200 on /health, and 388 unit tests pass across 5 test suites with zero failures. 33 minutes. | ||||||
| 5 | zvec | Runs | 100 / 100 | C++ | 16,077 | |
A lightweight, lightning-fast, in-process vector database What the test found: Zvec v0.7.0 prebuilt wheel and from-source build both work: the Python SDK successfully creates vectors collections, inserts documents, and returns ranked similarity search results, and the 1543-test suite passes with zero failures. 58 minutes. | ||||||
| 6 | langchain4j | Runs | 100 / 100 | Java | 13,216 | |
LangChain4j is an idiomatic, open-source Java library for building LLM-powered applications on the JVM. It offers a unified API over popular LLM providers and... What the test found: JDK 17 installed from Adoptium tarball, Maven wrapper mvnw boots correctly, full project compiles without errors, and tests across langchain4j-core (1413), langchain4j (1724), and 10+ provider/integration modules all pass with 0 failures. 36 minutes. | ||||||
| 7 | LongMemory | Runs | 100 / 100 | TypeScript | 4,523 | |
Local persistent memory store for LLM applications including claude desktop, github copilot, codex, antigravity, etc. What the test found: The LongMemory Hydrograph TypeScript package builds, its CLI (init, ingest, recall, doctor) works, the library API (createMemory, ingest, recall) returns correct results, the HTTP server starts on port 7331 and responds 200 on /health, and the CI benchmark suite passes all 11 smoke-test gates with 100% retrieval... 9 minutes. | ||||||
| 8 | crate | Runs | 100 / 100 | Java | 4,444 | |
CrateDB is a distributed and scalable SQL database for storing and analyzing massive amounts of data in near real-time, even with complex queries. It is... What the test found: CrateDB 6.5.0 compiles, installs, launches as a single-node cluster, accepts HTTP SQL queries, and returns correct query results. 27 minutes. | ||||||
| 9 | xerj | Runs | 100 / 100 | Rust | 3,229 | |
XERJ is the new way for AI to search data. Its autoindex capability activates agents to know your data without the token waste of grep and sed. One command... What the test found: xerj v1.0.0-rc.77 is installed at /home/runner/.local/bin/xerj, started in 690ms on --insecure mode, successfully autoindexed /work/repo/docs (18 datasets, 9681 records, 21s wall time), and serves queries on http://127.0.0.1:9200 with ES 8.13 wire protocol. No build step needed (binary install from GitHub releases)... 3 minutes. | ||||||
| 10 | vearch | Runs | 100 / 100 | Python | 2,331 | |
Distributed vector search for AI-native applications What the test found: Vearch v3.5.9 binary (81MB) and gamma engine (.so) built from source, server starts with master+router+PS on localhost:8817/9001, REST API responds, db/create/list works, cluster health/stats/members endpoints return valid JSON, tests pass basic cluster operations but space creation times out due to PS heartbeat... 59 minutes. | ||||||
| 11 | matrixone | Runs | 100 / 100 | Go | 2,040 | |
AI-native HTAP database with Git-for-Data and built-in vector search, serving as the data and memory backbone for intelligent agents and applications. What the test found: MatrixOne server built from source (mo-service, 275MB binary), launched on port 6001, accepting MySQL protocol connections, queries returning correct results including data manipulation operations. 34 minutes. | ||||||
| 12 | mcp-memory-service | Runs | 100 / 100 | Python | 1,995 | |
Open-source persistent memory for AI agent pipelines (LangGraph, CrewAI, AutoGen) and Claude. REST API + knowledge graph + autonomous consolidation. What the test found: mcp-memory-service 11.13.0 installed from source via pip install -e '.[dev]', the uvicorn server starts on 127.0.0.1:9791 (HTTP) with sqlite_vec backend, health endpoint returns 200, memory CRUD operations work end to end, and the full test suite passes except for one benchmark test that fails due to hash-embedding... 20 minutes. | ||||||
| 13 | SeekStorm | Runs | 100 / 100 | Rust | 1,918 | |
SeekStorm: vector & lexical search - in-process library & multi-tenancy server, in Rust. What the test found: Workspace builds cleanly with Rust 1.99.0, all 45 seekstorm library integration tests and all 7 seekstorm_client E2E tests pass, and the release seekstorm_server binary serves the live endpoint (HTTP 200) and completes the API key creation, index creation, document indexing, commit, and query flow returning the... 63 minutes. | ||||||
| 14 | LLPhant | Runs | 100 / 100 | PHP | 1,716 | |
LLPhant - A comprehensive PHP Generative AI Framework using OpenAI GPT 4. Inspired by Langchain What the test found: Static PHP 8.2.32 and Composer 2.7.9 installed; all 119 Composer dependencies resolved; 223 unit tests pass in 0.3s; PHPStan static analysis passes with 0 errors; Laravel Pint style check passes on 294 files. 6 minutes. | ||||||
| 15 | OpenContracts | Runs | 100 / 100 | Python | 1,502 | |
The open document intelligence platform for builders and hackers - DMS for the agentic world What the test found: Python dependencies installed in virtual environment, embedded PostgreSQL 18.6 provisioned and running, Django app boots with all migrations applied, and 613+ tests pass verifying admin, agents, annotations, authentication, analysis pipeline, badge system, base services, and artifact service against a real PostgreSQL... 77 minutes. | ||||||
| 16 | pymilvus | Runs | 100 / 100 | Python | 1,415 | |
Python SDK for Milvus Vector Database What the test found: PyMilvus SDK installed via uv sync, imports successfully, get_commit returns '290d76f', and the full unit test suite passes 5255/5255 with 3 skipped. 36 minutes. | ||||||
| 17 | NGT | Runs | 100 / 100 | C++ | 1,374 | |
Nearest Neighbor Search with Neighborhood Graph and Tree for High-dimensional Data What the test found: NGT v2.8.1 C++ library, command-line tool, and four sample programs build and run successfully; the CLI creates indexes, appends data, searches with L2/cosine/hamming distance, and reports index statistics; the Python ctypes binding (ngt.base) inserts 5000 128-d vectors and returns correct nearest-neighbor results. 11 minutes. | ||||||
| 18 | qdrant-client | Runs | 100 / 100 | Python | 1,370 | |
Python client for Qdrant vector search engine What the test found: The qdrant-client Python library installed without errors. All local-mode tests pass. When a Qdrant server v1.19.1 runs in the same session, all REST and gRPC integration tests pass. The in-memory client creates collections, upserts points, and searches vectors successfully. 32 minutes. | ||||||
| 19 | memory-os | Runs | 96 / 100 | Python | 1,372 | |
A 7-layer memory operating system for Hermes Agent, persistent memory with Qdrant, structured facts, fabric recall, auto-curated wiki, and surgical context... What the test found: Python dependencies install cleanly in a venv, all 7 icarus module imports succeed, all 3 standalone test suites pass with 0 failures, SQLite databases are initialized with correct FTS5 schema, Icarus plugin tools produce valid output, and the Docker-compose file, worker services, and cron scripts are syntactically... 8 minutes. | ||||||
| 20 | txtai | Runs | 94.7 / 100 | Python | 12,998 | |
💡 All-in-one AI framework for semantic search, LLM orchestration and language model workflows What the test found: txtai 9.14.0 installed in venv, all core embeddings/graph/workflow/agent/cloud/vector/scoring tests pass, API responds to search/index/count endpoints with real model inference, ~470 automated tests pass across 30 test suites. 83 minutes. | ||||||
| 21 | helix-db | Runs | 60 / 100 | Rust | 6,117 | |
HelixDB is an OLTP graph database with native vector and full-text search built in Rust on Object Storage. What the test found: The HelixDB workspace builds and all crate test suites pass. The server binary starts and responds to HTTP POST /v2/query with HTTP 200. The CLI binary prints version 3.3.0 and help output. 51 minutes. | ||||||
| 22 | infinity | Runs | 60 / 100 | C++ | 4,734 | |
The AI-native database built for LLM applications, providing incredibly fast hybrid search of dense vector, sparse vector, tensor (multi-vector), and full-text. What the test found: Infinity v0.7.3 nightly (x86_64-v2) standalone server starts and accepts Thrift connections on port 23817 and HTTP on port 23820; Python SDK can create tables, insert vectors, run dense vector search queries, and return results via Polars DataFrames. 12 minutes. | ||||||
| 23 | arcadedb | Runs | 60 / 100 | Java | 1,189 | |
ArcadeDB Multi-Model Database, one DBMS that supports SQL, Cypher, Gremlin, HTTP/JSON, MongoDB and Redis. ArcadeDB is a conceptual fork of OrientDB, the first... What the test found: ArcadeDB 26.10.1-SNAPSHOT is fully built and functional: the server starts in development mode listening on port 2480, serving the Studio web UI (HTTP 200), and the engine test suite passes ~14624 tests with 0 errors and 1 pre-existing flaky timeout test that fails under heap contention during concurrent runs. 85 minutes. | ||||||
| 24 | neuron-ai | Runs | 58.6 / 100 | PHP | 2,125 | |
The Agentic Framework of the PHP ecosystem to build production-ready AI driven applications. Connect components (LLMs, Tools, vector DBs, memory) to agents... What the test found: Composer install with all 145 dependencies succeeded; vendor/bin/phpunit reports 1373 tests, 5460 assertions passed, 0 failures, 0 errors, 78 skipped (external services); vendor/bin/phpstan analyse reports OK, 0 errors; bin/neuron CLI reports help with 7 commands. 33 minutes. | ||||||
| 25 | lancedb | Runs | 48 / 100 | Rust | 11,617 | |
Developer-friendly OSS embedded retrieval library for multimodal AI. Search More; Manage Less. What the test found: Rust core crate lancedb compiles and passes all 73 integration tests; Python bindings install, import, and pass 1368 tests; the local (in-process) embedding store works end-to-end with vector search queries returning correct results. 84 minutes. | ||||||
| 26 | claude-context | Runs with mocks | 92 / 100 | TypeScript | 12,593 | |
Code search MCP for Claude Code. Make entire codebase the context for any coding agent. What the test found: Install, build, typecheck, lint config, and all 35 tests pass across the monorepo. The core indexing pipeline runs end-to-end with mock providers. The MCP server starts but requires real OpenAI API key and Milvus credentials for production use. 15 minutes. | ||||||
| 27 | MineContext | Runs with mocks | 92 / 100 | Python | 5,537 | |
MineContext is your proactive context-aware AI partner(Context-Engineering+ChatGPT Pulse) What the test found: MineContext server starts successfully, serves all 12 HTTP routes (GET and POST), initializes ChromaDB vector store and SQLite document store, and the CLI help command prints usage information. 18 minutes. | ||||||
| 28 | ava-whatsapp-agent-course | Runs with mocks | 92 / 100 | Python | 1,679 | |
Meet Ava, the WhatsApp Agent What the test found: uv sync installs all dependencies, 30 Python source files compile, LangGraph workflow graph compiles with 9 nodes, FastAPI WhatsApp webhook answers GET verification with HTTP 200, Chainlit chat interface starts on port 8000 and serves HTTP 200, mock API keys produce 401 when POSTing messages as expected since real... 19 minutes. | ||||||
| 29 | LEANN | Runs with mocks | 89.5 / 100 | Python | 13,012 | |
[MLsys2026 Best Paper]: https://arxiv.org/abs/2506.08276. RAG on Everything with LEANN. Enjoy 97% storage savings while running a fast, accurate, and 100%... What the test found: LEANN Python packages installed, HNSW backend with custom-built Faiss SWIG module works, CI tests pass, README HNSW examples pass with mock embeddings, DiskANN backend blocked by missing libaio system dependency. 87 minutes. | ||||||
| 30 | deep-searcher | Runs with mocks | 56 / 100 | Python | 8,325 | |
Open Source Deep Research Alternative to Reason and Search on Private Data. Written in Python. What the test found: DeepSearcher package installs with uv sync and pip install -e . and all 422 unit tests pass after fixing 2 test files with stale default model values. 5 minutes. | ||||||
| 31 | autoflow | Could not verify | 25 / 100 | TypeScript | 2,978 | |
pingcap/autoflow is a Graph RAG based and conversational knowledge base tool built with TiDB Serverless Vector Storage. Demo: https://tidb.ai What the test found: could not be verified; the log shows where it stopped. 29 minutes. | ||||||
| 32 | orbit | Could not verify | 10 / 100 | Python | 353 | |
Self-hosted AI gateway for private RAG, natural-language data access, and tool-calling agents. What the test found: could not be verified; the log shows where it stopped. 28 minutes. | ||||||
| - | dingo | Not yet tested | - | Java | 1,703 | - |
| - | infinispan | Not yet tested | - | Java | 1,355 | - |
| - | SeaGOAT | Not yet tested | - | Python | 1,312 | - |
| - | VectorDBBench | Not yet tested | - | Python | 1,189 | - |
| - | chromem-go | Not yet tested | - | Go | 1,063 | - |
| - | FlashRank | Not yet tested | - | Python | 1,008 | - |
| - | chatWeb | Not yet tested | - | Python | 918 | - |
| - | rag-api | Not yet tested | - | Python | 911 | - |
| - | iai-personal-memory-engine | Not yet tested | - | Python | 902 | - |
| - | OpenSwarm | Not yet tested | - | TypeScript | 858 | - |
| - | automem | Not yet tested | - | Python | 821 | - |
| - | vector-db-from-scratch | Not yet tested | - | Rust | 807 | - |
| - | mcp-server-elasticsearch | Not yet tested | - | Rust | 717 | - |
| - | RAGLight | Not yet tested | - | Python | 672 | - |
| - | embedJs | Not yet tested | - | TypeScript | 601 | - |
| - | Compartment | Not yet tested | - | Python | 578 | - |
| - | aisearch-openai-rag-audio | Not yet tested | - | Python | 563 | - |
| - | next-plaid | Not yet tested | - | Rust | 556 | - |
| - | unbody | Not yet tested | - | TypeScript | 522 | - |
| - | wikipedia-semantic-search | Not yet tested | - | TypeScript | 471 | - |
| - | qdrant-js | Not yet tested | - | TypeScript | 465 | - |
| - | python-sdk | Not yet tested | - | Python | 453 | - |
| - | shopkeeper-agent | Not yet tested | - | Python | 435 | - |
| - | redis-vl-python | Not yet tested | - | Python | 431 | - |
| - | prompt-cache | Not yet tested | - | Go | 428 | - |
| - | memento-mcp | Not yet tested | - | TypeScript | 425 | - |
| - | vector-db-benchmark | Not yet tested | - | Python | 374 | - |
| - | vectorlite | Not yet tested | - | Rust | 361 | - |
| - | vicinity | Not yet tested | - | Python | 351 | - |
| - | shodh-memory | Not yet tested | - | Rust | 304 | - |
| - | memlayer | Not yet tested | - | Python | 301 | - |
| - | knowledge-rag | Not yet tested | - | Python | 291 | - |
| - | RaBitQ-Library | Not yet tested | - | C++ | 289 | - |
| - | chromadb-admin | Not yet tested | - | TypeScript | 283 | - |
| - | relevanceai | Not yet tested | - | Python | 283 | - |
| - | fast-plaid | Not yet tested | - | Python | 282 | - |
| - | radient | Not yet tested | - | Python | 281 | - |
| - | pinecone-ts-client | Not yet tested | - | TypeScript | 279 | - |
| - | cortexdb | Not yet tested | - | Go | 274 | - |
| - | uteke | Not yet tested | - | Rust | 269 | - |
| - | vectordbz | Not yet tested | - | TypeScript | 264 | - |
| - | vector-graph-rag | Not yet tested | - | Python | 254 | - |
| - | skills | Not yet tested | - | Python | 253 | - |
| - | vector-storage | Not yet tested | - | TypeScript | 248 | - |
| - | ahnlich | Not yet tested | - | Rust | 247 | - |
| - | satoriDB | Not yet tested | - | Rust | 246 | - |
| - | k8ssandra-operator | Not yet tested | - | Go | 243 | - |
| - | RelaMind | Not yet tested | - | Java | 210 | - |
| - | chroma-go | Not yet tested | - | Go | 208 | - |
| - | nano-vectordb | Not yet tested | - | Python | 208 | - |
| - | corpusos | Not yet tested | - | Python | 205 | - |
| - | Personal_External_Brain | Not yet tested | - | Python | 201 | - |
runs installed and started with its real dependencies. runs with mocks started after stand-ins replaced external services such as a database or a third-party API. could not verify neither the standard agent nor the stronger one got it running within the time limit; the log shows where it stopped.
How we tested
On this list as of the latest test: 25 projects ran as-is, 5 with mocks, 2 could not be verified, 52 still waiting. Languages tested: C++, Go, Java, JavaScript, PHP, Python, Rust, TypeScript. Every attempt used a clean single-use machine, the subject at a pinned version, and a 45-minute limit; the complete procedure is on the methodology page.
Frequently asked questions (FAQs)
How is this list ranked?
By measurement, not opinion: projects Argusic installed and launched on a fresh machine come first, then those that ran with mocks in place of external services, then those it could not verify. Ties go to the Argusic Score, then how popular it is on its own source.
Why are some projects unranked?
52 projects are still waiting for a test or for a finished attempt. They are listed without a rank until Argusic has measured them.
Where is the evidence?
Every row links to the project's Argusic page, where each run has a full log and a terminal recording stored with a sha256 fingerprint. The same pages exist for every one of the tested projects, on this list or not.
More lists in this category
- search engines (shares FlashRank, SeekStorm, arcadedb, cortexdb, infinispan, infinity, lancedb, relevanceai, txtai, xerj, zvec with this list)
- AI agent frameworks (shares PageIndex, anything-llm, cognee, txtai with this list)
- Open source databases, installed and queried (shares arcadedb, crate, helix-db, infinispan with this list)
- self-hosted AI apps (shares Personal_External_Brain, memory-os, orbit with this list)
- MCP servers (shares iai-personal-memory-engine, mcp-memory-service with this list)
- LLM gateways (shares orbit with this list)
- open source coding agents (shares claude-context with this list)
- speech tools (shares orbit with this list)
- web scraping tools (shares chatWeb with this list)
- wikis (shares Personal_External_Brain with this list)
- API clients
- API gateways
- CI/CD tools
- CMS platforms
- VPN tools
- backup tools
- browser automation tools
- code editors
- developer CLI tools
- e-commerce platforms
- ebook readers
- game engines
- open source games
- home automation tools
- low-code platforms
- map tools
- message queues
- Open source music servers, installed and played
- observability tools
- Open source office suites, installed and launched
- self-hosted password managers
- project management tools
- screen recorders
- self-hosted analytics
- static site generators
- uptime monitors
- open source video editors
- video players
- whiteboard tools
- workflow automation tools
- self-hosted Notion alternatives
- self-hosted dashboards
- self-hosted git servers