A lightning-fast search engine API bringing AI-powered hybrid search to your sites and applications.
LLMRepos·Vector DB / Index & Search LLM projects
Updated dailyBrowse 131 open-source vector db / index & search projects in Infrastructure. Compare GitHub stars, recent growth, languages, licenses, and repository activity.
Vector DB / Index & Search
Browse 131 open-source vector db / index & search projects in Infrastructure. Compare GitHub stars, recent growth, languages, licenses, and repository activity.
Top Vector DB / Index & Search repositories
Ranked by current GitHub stars from the latest LLMRepos snapshot.
Showing 40 of 131
Milvus is a high-performance, cloud-native vector database built for scalable vector ANN search
Qdrant - High-performance, massive-scale Vector Database and Vector Search Engine for the next generation of AI. Also available in the cloud https://cloud.qdrant.io/
Weaviate is an open-source vector database that stores both objects and vectors, allowing for the combination of vector search with structured filtering with the fault tolerance and scalability of a cloud-native database.
Developer-friendly OSS embedded retrieval library for multimodal AI. Search More; Manage Less.
🌌 A complete search engine and RAG pipeline in your browser, server or edge network with support for full-text, vector, and hybrid search in less than 2kb.
The Fastest Distributed Database for Transactional, Analytical, and AI Workloads.
Data Agent Ready Warehouse : One for Analytics, Search, AI, Python Sandbox. — rebuilt from scratch. Unified architecture on your S3.
One Postgres for your application data, full-text search, vector retrieval, and aggregations. Home of the pg_search extension.
MariaDB server is a community developed fork of MySQL server. Started by core members of the original MySQL team, MariaDB actively works with outside developers to deliver the most featureful, stable, and sanely licensed open SQL server in the industry.
Open Lakehouse Format for Multimodal AI. Convert from Parquet in 2 lines of code for 100x faster random access, vector index, and data versioning. Compatible with Pandas, DuckDB, Polars, Pyarrow, and PyTorch with more integrations coming..
A query and indexing engine for Redis, providing secondary indexing, full-text search, vector similarity search and aggregations.
A suite of tools to develop RAG, semantic search, and other AI applications more easily with PostgreSQL
crate/crate
CrateDB is a distributed and scalable SQL database for storing and analyzing massive amounts of data in near real-time, even with complex queries. It is PostgreSQL-compatible, and based on Lucene.
The Best GUI for Milvus
Jupyter Notebooks to help you get hands-on with Pinecone vector databases
The universal tool suite for vector database management. Manage Pinecone, Chroma, Qdrant, Weaviate and more vector databases with ease.
Scalable, Low-latency and Hybrid-enabled Vector Search in Postgres. Revolutionize Vector Search, not Database.
The Agentic Framework of the PHP ecosystem to build production-ready AI driven applications. Connect components (LLMs, Tools, vector DBs, memory) to agents that interact with your data and UI.
SeekStorm: vector & lexical search - in-process library & multi-tenancy server, in Rust.
Nomic Developer API SDK
Practical course about Large Language Models.
Scalable, fast, and disk-friendly vector search in Postgres, the successor of pgvecto.rs.
LLPhant - A comprehensive PHP Generative AI Framework using OpenAI GPT 4. Inspired by Langchain
A multi-modal vector database that supports upserts and vector queries using unified SQL (MySQL-Compatible) on structured and unstructured data, while meeting the requirements of high concurrency and ultra-low latency.
Meet Ava, the WhatsApp Agent
XERJ is the new way for AI to search data. Its autoindex capability activates agents to know your data without the token waste of grep and sed. One command indexes code, docs, logs and PDFs for search, RAG, security audits and agent memory, using 40x fewer tokens than grep. Elasticsearch compatible, so existing clients just work.
A simple, fast and versatile Datalog database
Python SDK for Milvus Vector Database
NGT-labs/NGT
Nearest Neighbor Search with Neighborhood Graph and Tree for High-dimensional Data
Python client for Qdrant vector search engine
Infinispan is an open source data grid platform and highly scalable NoSQL cloud data store.
Endee.io – A high-performance vector database, designed to handle up to 1B vectors on a single node, delivering significant performance gains through optimized indexing and execution. Also available in cloud https://endee.io/
local-first semantic code search engine
Benchmark for vector databases.
Open Source Semantic Search for your AI Agent
Open-source, self-hosted enterprise & site search server built on OpenSearch. Crawls web / file / DB / cloud sources, 20+ languages, REST API, and AI/RAG & semantic search. Apache-2.0.
ArcadeDB Multi-Model Database, one DBMS that supports SQL, Cypher, Gremlin, HTTP/JSON, MongoDB and Redis. ArcadeDB is a conceptual fork of OrientDB, the first Multi-Model DBMS. ArcadeDB supports Vector Embeddings.
Lakehouse native graph engine with git-style workflows
A FastAPI service for semantic text search using precomputed embeddings and advanced similarity measures, with built-in support for various file types through textract.