#vector-database
Open source repositories tagged with #vector-database, ranked by health score.
XERJ is the new way for AI to search data. Its autoindex capability activates agents to know your data without the token waste of grep and sed. One command indexes code, docs, logs and PDFs for search, RAG, security audits and agent memory, using 40x fewer tokens than grep. Elasticsearch compatible, so existing clients just work.
The Agentic Framework of the PHP ecosystem to build production-ready AI driven applications. Connect components (LLMs, Tools, vector DBs, memory) to agents that interact with your data and UI.
Cognee is the open-source AI memory platform for agents. Give your AI agents persistent long-term memory with small models for free
🧠 The Brain for Your AI — Local-first memory engine for AI agents. Store, recall, and search memories with semantic embeddings. Single Rust binary, zero config, fully offline.
NextPlaid, ColGREP: Multi-vector search, from database to coding agents.
One Postgres for your application data, full-text search, vector retrieval, and aggregations. Home of the pg_search extension.
Nornicdb is a distributed low-latency, Graph+Vector, Temporal MVCC with all sub-ms HNSW search, graph traversal, and writes. Using Neo4j Bolt/Cypher and qdrant's gRPC means you can switch with no changes while adding intelligent features like schemas, managed embeddings, reranking+llm, GPU accel, Auto-TLP, Policy-based Memory Decay, and MCP server.
The AI search platform
LangChain4j is an idiomatic, open-source Java library for building LLM-powered applications on the JVM. It offers a unified API over popular LLM providers and vector stores, and makes implementing tool calling (including MCP support), agents and RAG easy. It integrates seamlessly with enterprise Java frameworks like Quarkus and Spring Boot.
ArcadeDB Multi-Model Database, one DBMS that supports SQL, Cypher, Gremlin, HTTP/JSON, MongoDB and Redis. ArcadeDB is a conceptual fork of OrientDB, the first Multi-Model DBMS. ArcadeDB supports Vector Embeddings.
AI-native HTAP database with Git-for-Data and built-in vector search, serving as the data and memory backbone for intelligent agents and applications.
Embedded retrieval library built on Parquet. Fast, efficient, and scalable.
An open-source, local-first CV & resume management platform featuring privacy-focused semantic search to instantly match your skills, projects, and experiences
A living benchmark of how quickly frontier coding agents adapt to real API changes — verified, timestamped API-evolution cases for evaluating coding models and agents. Direction reset Aug 2026: the original API-doc index is now measurement infrastructure.