Skip to content
73

Awesome Vector Search

Collections of vector search related libraries, service and research papers

1.6k stars134 forks79 entriesLast push Jul 6, 2026 (2 months ago)License MIT

This page lists names, links and short descriptions. The original list on GitHub is the source and belongs to its authors.

Awesome Vector Search Engine >Standalone Service

Apache Cassandra 5.0 – Vector search (cep-30), Strict Serialisable ACID (cep-15), horizontally scaling database

is an open source NoSQL distributed database trusted by thousands of companies for scalability and high availability without compromising performance. Cassandra provides linear scalability and proven fault-tolerance on commodity hardware or cloud infrastructure make it the perfect platform for…

In 4 listsDetails

Qdrant - Vector Similarity Search Engine with extended filtering support

Vector Search Engine and Database for the next generation of AI applications. Also available in the cloud

In 8 listsDetails

Vald - A Highly Scalable Distributed Vector Search Engine

Highly scalable distributed vector search engine. Cloud-native architecture with automatic indexing, horizontal scaling, and multiple ANN algorithm support. Apache 2.0 licensed.

In 3 lists

Milvus - A cloud-native vector database with high-performance and high scalability.

Milvus is a cloud-native, open-source vector database built to manage embedding vectors generated by machine learning models and neural networks.

In 14 listsDetails

Weaviate - A cloud-native, real-time vector search engine

Weaviate is an open source vector search engine that stores both objects and vectors, allowing for combining vector search with structured filtering with the fault-tolerance and scalability of a cloud-native database, all accessible through GraphQL, REST, and various language clients.

In 2 lists

Omnigraph - Typed graph database where agents branch and merge like Git. S3-native, Rust, traversal + vector + BM25 in…

Typed graph database where agents branch and merge like Git. S3-native, Rust, traversal + vector + BM25 in one runtime.

In 5 listsDetails

OpenDistro Elasticsearch KNN - A machine learning plugin which supports an approximate k-NN search algorithm for Open…

Elastiknn - Elasticsearch plugin for nearest neighbor search

Epsilla - A High Performance Vector Database Management System, Hippocampus For AI

Epsilla is a high performance Vector Database Management System. Try out hosted Epsilla at https://cloud.epsilla.com/

In 3 lists

Vearch - A scalable distributed system for efficient similarity search of deep learning vectors

Cloud-native distributed vector database for AI-native applications. Efficient similarity search of embedding vectors with horizontal scaling and real-time indexing. Apache 2.0 licensed.

In 3 lists

pgANN - Fast Approximate Nearest Neighbor (ANN) searches with a PostgreSQL database

Jina - Jina allows you to build deep learning-powered search-as-a-service.

Build multimodal AI services via cloud native technologies · Model Serving · Generative AI · Neural Search · Cloud Native

In 3 lists

Infinity - The AI-native database built for LLM applications, providing incredibly fast vector and full-text search

AI-native database built for LLM applications with incredibly fast hybrid search of dense vector, sparse vector, tensor (multi-vector), and full-text. Powers RAGFlow's document engine. Apache 2.0 licensed.

In 9 listsDetails

Aquila DB - Distribution focused k-NN search algorithm

An easy to use Neural Search Engine. Index latent vectors along with JSON metadata and do efficient k-NN search.

In 2 lists

Redis HNSW - A redis module for similarity search based on HNSW

Solr - Apache Solr

has a Dense Vector Search feature as of Solr 9.0

In 3 lists

Marqo - A semantic search engine which supports tensor search (sequence of vectors)

Multimodal vector search for text, image, and structured data. End-to-end indexing and search with built-in embedding models. Apache 2.0 licensed.

In 5 listsDetails

txtai - Build semantic search applications and workflows

All-in-one AI framework for semantic search, LLM orchestration and language model workflows. Embeddings database with customizable pipelines.

In 9 listsDetails

Semantra - A multipurpose tool for semantically searching documents.

SuperDuperDB - Bring AI to your favorite database

(label: good first issue) 🔮SuperDuperDB: Bring AI to your favourite database! Integrate, train and manage any AI models and APIs directly with your database and your data

In 2 lists

TensorDB - High Performance Vector Database Supporting Heterogeneous Computing

JVector - a pure Java, zero dependency, embedded vector search engine, used by DataStax Astra DB and Apache Cassandra.

VQLite - Simple and Lightweight Vector Search Engine

Vexvault - 100% browser based, open source, scalable, simple, zero-cost vector search

Vespa.ai - Text search engine and ... fast approximate vector search (ANN)

Vespa's large-scale ANN search using HNSW-IF indexes is described here

Awesome Vector Search Engine >Library

LangStream - LangStream is an open-source project that combines the best of event-based architectures with the latest…

CassIO - CassIO is the ultimate solution for seamlessly integrating Apache Cassandra® with generative artificial…

JVector - a pure Java, zero dependency, embedded vector search engine, used by DataStax Astra DB and Apache Cassandra.

Faiss - A library for efficient similarity search and clustering of dense vectors

is a library for efficient similarity search and clustering of dense vectors. It contains algorithms that search in sets of vectors of any size, up to ones that possibly do not fit in RAM. It also contains supporting code for evaluation and parameter tuning. Faiss is written in C++ with complete…

In 8 listsDetails

Distributed Faiss - Work with FAISS indexes which don't fit into a single server memory

Autofaiss - Automatically create Faiss knn indices

SimSIMD - Hardware-accelerated mixed-precision numerics library for dense and sparse vector math and search

适用于 x86 AVX2、AVX-512、Arm NEON 和 SVE 的矢量距离函数。

In 2 lists

ScaNN - A library efficient vector similarity search at scale.

Open-source library for distributed matrix factorization using Alternating Least Squares, more info in ALX: Large Scale Matrix Factorization on TPUs.

In 4 lists

NMSLIB - Non-Metric Space Library, an efficient similarity search library for generic non-metric spaces

Non-Metric Space Library for efficient similarity search in generic non-metric spaces. Comprehensive toolkit for evaluating k-NN methods with support for exotic distance functions. Apache 2.0 licensed.

In 4 listsDetails

Annoy - C++ library with Python bindings to search for points

is a C++ library with Python bindings to search for points in space that are close to a given query point. It also creates large read-only file-based data structures that are mmapped into memory so that many processes may share the same data.

In 6 listsDetails

FLANN - Library written in C++ and contains bindings for the following languages: C, MATLAB, Python, and Ruby

Fast Library for Approximate Nearest Neighbors.

In 4 lists

LLM App - Open-source Python library for a real-time data KNN (K-Nearest Neighbors) indexing

LLM App is a Python library that helps you build real-time LLM-enabled data pipelines with few lines of code.

In 5 listsDetails

MRPT - Fast nearest neighbor search with random projection

RPForest - Python library for approximate nearest neighbours search

pgvector - Open-source vector similarity search extension for Postgres

Vector search inside Postgres. Try this before adding a dedicated vector database.

In 6 listsDetails

PASE - Ultra-High-Dimensional approximate nearest neighbor search extension for Postgres

Pyserini - Toolkit for reproducible information retrieval research with sparse and dense representations

NGT - Provides commands and a library for performing high-speed approximate nearest neighbor

NearPy - Approximate search using different locality-sensitive hashing methods

TOROS N2 - lightweight approximate Nearest Neighbor library

PUFFINN - Parameterless and Universal Fast FInding of Nearest Neighbors

SPTAG - A distributed approximate nearest neighborhood search (ANN) library

A distributed approximate nearest neighborhood search (ANN) library which provides a high quality vector index build, search and distributed online serving toolkits for large scale vector search scenario.

In 3 lists

PyNNDescent - A python nearest neighbor descent for approximate k nearest neighbors

TarsosLSH - A Java library implementing practical nearest neighbour search algorithm for multidimensional vectors

TorchPQ - Efficient implementations of Product Quantization and its variants using Pytorch and CUDA

Granne - Graph-based retrieval of approximate nearest neighbors witten in rust

Embeddinghub - A database built for machine learning embeddings

Hora - Efficient approximate nearest neighbor search algorithm collections library written in Rust

Voy - A WASM vector similarity search engine written in Rust

In 2 lists

altor-vec - In-browser HNSW vector search in JavaScript. 54KB WASM, sub-millisecond queries, no server required. npm…

Chroma - The open-source embedding database for building LLM apps in Python or JavaScript with memory

An open-source embedding database for building AI applications with embeddings and semantic search.

In 8 listsDetails

USearch - Smaller & Faster Vector Search Engine for C++, Python, JavaScript, Rust, Java, GoLang, Wolfram

Fast single-file similarity search & clustering engine for vectors. Smaller and faster than FAISS with 20+ language bindings (C++, Python, JavaScript, Rust, Java, Go, etc.) and support for custom metrics. Apache 2.0 licensed.

In 6 listsDetails

Golang vector stores collection - Chroma, PGVector interfaces

Scalable Vector Search (SVS) - A performance library for vector similarity search

chromem-go - Embeddable vector database for Go with Chroma-like interface and zero third-party dependencies. In-memory…

star:832 Embeddable vector database for Go with Chroma-like interface and zero third-party dependencies. In-memory with optional persistence.

In 3 lists

CocoIndex - An open-source ETL framework with realtime incremental processing to keep index fresh

ETL framework to build fresh context for AI agents, with incremental processing

In 6 listsDetails

Moss - Sub-10ms semantic search engine for Voice & Conversational AI, built in Rust/WebAssembly for on-device /…

Awesome Vector Search Engine >Cloud Service

DataStax Astra Vector - Multi-cloud, serverless vector DBaaS

Epsilla Cloud - The fully managed serverless vector database with 10X faster, cheaper and better.

MongoDB Atlas Vector Search - Multi-cloud vector search with filtering, horizontal scaling, etc.

MyScale - A managed vector database based on ClickHouse

Pinecone - Managed vector search with filtering, live index updates, horizontal scaling, and a lot more

The Pinecone vector database makes it easy to build high-performance vector search applications. Developer-friendly, fully managed, and easily scalable without infrastructure hassles.

In 8 listsDetails

Redis Cloud - Managed vector database in Redis

SemaDB Cloud - Easy-to-use managed vector database with a RESTful API

Relevance AI - Vector Platform From Experimentation To Deployment

Zilliz Cloud - Cloud-native service for Milvus

Rivestack - Managed PostgreSQL with pgvector for AI workloads. Built-in SQL editor lets you query your database with…

Managed PostgreSQL with pgvector for AI workloads. Built-in SQL editor lets you query your database with natural language (auto-converted to vector embeddings). Free tier includes 2GB storage.

In 3 lists

Awesome Vector Search Engine >Research Papers

SPANN: Highly-efficient Billion-scale Approximate Nearest Neighborhood Search - NEURIPS 2021

Revisiting the Inverted Indices for Billion-Scale Approximate Nearest Neighbors - ECCV 2018

Accelerating Large-Scale Inference with Anisotropic Vector Quantization

Billion-scale similarity search with GPUs

Efficient and robust approximate nearest neighbor search using Hierarchical Navigable Small World graphs

Optimization of Indexing Based on k-Nearest Neighbor Graph for Proximity Search in High-dimensional Data

On Approximately Searching for Similar Word Embeddings - ACL 2016

See category
94

Table of Contents

hesreallyhim/awesome-claude-code

A hand-picked collection of the finest of resources for the most awesome of agents, Claude Code, the undisputed champion of coding companions, from the unstoppable team…

Fresh★ 55k202 entriesPushed today
94

Awesome Agent Skills

VoltAgent/awesome-agent-skills

A curated collection of 1000+ agent skills from official dev teams and the community, compatible with Claude Code, Codex, Gemini CLI, Cursor, and more.

Fresh★ 35k839 entriesPushed today
93

Awesome Machine Learning

josephmisiti/awesome-machine-learning

A curated list of awesome Machine Learning frameworks, libraries and software.

Fresh★ 74k1188 entriesPushed 7 days ago
92

Awesome Production Machine Learning

EthicalML/awesome-production-machine-learning

A curated list of awesome open source libraries to deploy, monitor, version and scale your machine learning

Fresh★ 21k519 entriesPushed 3 days ago
92

AWESOME DATA SCIENCE

academic/awesome-datascience

:memo: An awesome Data Science repository to learn and apply for real world problems.

Fresh★ 30k881 entriesPushed today
91

Static Analysis

analysis-tools-dev/static-analysis

⚙️ A curated list of static analysis (SAST) tools and linters for all programming languages, config files, build tools, and more. The focus is on tools which improve…

Fresh★ 15k528 entriesPushed 8 days ago