Awesome Big Data
Section: Vector Databases · open-source, in-process vector database for dense, sparse, and hybrid similarity search.
Entry
Appears in 7 awesome lists
Lightweight, lightning-fast, in-process vector database from Alibaba. Built on Proxima (Alibaba's battle-tested vector search engine) for production-grade, low-latency similarity search. Apache 2.0 licensed.
Section: Vector Databases · open-source, in-process vector database for dense, sparse, and hybrid similarity search.
Section: Database · A lightweight, lightning-fast, in-process vector database. [Apache2] website
Section: Databases · An embedded vector database for on-device RAG and edge AI, the SQLite of vector databases.
Section: 5. Retrieval-Augmented Generation (RAG) & Knowledge · Lightweight, lightning-fast, in-process vector database from Alibaba. Built on Proxima (Alibaba's battle-tested vector search engine) for production-grade, low-latency similarity search. Apache 2.0 licensed.
Section: Industry Strength Information Retrieval · Zvec is an open-source, in-process vector database for low-latency similarity search.
Section: Database · A lightweight, in-process vector database that embeds directly into applications.
Section: Other · A lightweight, lightning-fast, in-process vector database
Milvus is a cloud-native, open-source vector database built to manage embedding vectors generated by machine learning models and neural networks.
TiDB is built for agentic workloads that grow unpredictably, with ACID guarantees and native support for transactions, analytics, and vector search. No data silos. No noisy neighbors. No infrastructure ceiling.
is a search engine based on the Lucene library. It provides a distributed, multitenant-capable full-text search engine with an HTTP web interface and schema-free JSON documents. Elasticsearch is developed in Java.
General purpose, document-based, distributed database built for modern applications.
AI-native database built for LLM applications with incredibly fast hybrid search of dense vector, sparse vector, tensor (multi-vector), and full-text. Powers RAGFlow's document engine. Apache 2.0 licensed.
The lightweight, fault-tolerant database built on SQLite. Designed to keep your data highly available with minimal effort.
is a powerful, open source object-relational database system with over 30 years of active development that has earned it a strong reputation for reliability, feature robustness, and performance.