Milvus
by Community
Milvus is a high-performance, cloud-native vector database built for scalable vector ANN search
OSS
Milvus
Added 1 June 2026
Overview
Milvus is an open-source vector database written in Go that performs approximate nearest neighbor (ANN) search at scale. It handles high-dimensional vector indexing and retrieval for applications like semantic search, recommendation systems, and similarity matching. Designed for cloud-native deployment, it supports distributed architectures and multiple index types.
Best for
Best for
Teams building search or recommendation features who need to manage vector data at scale and prefer open-source control over managed services.
Use cases
- Semantic search over embeddings from LLMs
- Recommendation engines based on vector similarity
- Image or document retrieval by learned representations
Notes
Milvus is an open-source vector database written in Go that performs approximate nearest neighbor (ANN) search at scale. It handles high-dimensional vector indexing and retrieval for applications like semantic search, recommendation systems, and similarity matching. Designed for cloud-native deployment, it supports distributed architectures and multiple index types.
44,579 stars on GitHub. Last updated 2026-06-01. Licensed Apache-2.0.
Use cases
- Semantic search over embeddings from LLMs
- Recommendation engines based on vector similarity
- Image or document retrieval by learned representations
Pros
- High throughput ANN search with tunable accuracy-speed tradeoffs
- Cloud-native design with horizontal scaling support
- Active open-source community with 44k+ GitHub stars
Cons
- Requires operational overhead to deploy and maintain in production
- Learning curve for index tuning and configuration optimization
- Separate system to integrate alongside existing data infrastructure
Indexed from awesome-llmops and enriched against its public facts.
Pros
- High throughput ANN search with tunable accuracy-speed tradeoffs
- Cloud-native design with horizontal scaling support
- Active open-source community with 44k+ GitHub stars
Cons
- Requires operational overhead to deploy and maintain in production
- Learning curve for index tuning and configuration optimization
- Separate system to integrate alongside existing data infrastructure
Open-source & AI alternatives
Swap-in tools that solve the same job. Weigh the trade-offs before you commit.
Qdrant
Community
Qdrant - High-performance, massive-scale Vector Database and Vector Search Engine for the next generation of AI. Also available in the cloud https://cloud.qdrant.io/
Chroma
Community
Search infrastructure for AI
pgvector
Community
Open-source vector similarity search for Postgres
Weaviate
Community
Weaviate is an open-source vector database that stores both objects and vectors, allowing for the combination of vector search with structured filtering with the fault tolerance an
ArcadeData/arcadedb
Various
ArcadeDB Multi-Model Database, one DBMS that supports SQL, Cypher, Gremlin, HTTP/JSON, MongoDB and Redis. ArcadeDB is a conceptual fork of OrientDB, the first Multi-Model DBMS. Arc
AquilaDB
Community
An easy to use Neural Search Engine. Index latent vectors along with JSON metadata and do efficient k-NN search.
Awadb
Community
AI Native database for embedding vectors
Chroma
Community
Search infrastructure for AI
deeplake
Community
Deeplake is AI Data Runtime for Agents. It provides serverless postgres with a multimodal datalake, enabling scalable retrieval and training.
Lancedb
Community
Developer-friendly OSS embedded retrieval library for multimodal AI. Search More; Manage Less.
Marqo
Community
Ecommerce Search and Discovery - marqo.ai
pgvector
Community
Open-source vector similarity search for Postgres
Pinecone
Community
Search through billions of items for similar matches to any object, in milliseconds. It’s the next generation of search, an API call away.
Qdrant
Community
Qdrant - High-performance, massive-scale Vector Database and Vector Search Engine for the next generation of AI. Also available in the cloud https://cloud.qdrant.io/
Rivestack
Community
Managed pgvector on dedicated PostgreSQL with NVMe storage. 2,000 QPS at sub-4ms p50, from $35/month, migration help from Supabase, Neon, Pinecone, and self-hosted.
Statewave
Community
Open-source memory runtime for AI agents — reproducible, provenance-tagged context bundles instead of query-time retrieval. Apache-2.0, self-hosted on Postgres + pgvector, Python +
Vald
Community
Vald. A Highly Scalable Distributed Vector Search Engine
Vearch
Community
Distributed vector search for AI-native applications
VectorDB
Community
A Python vector database you just need - no more, no less.
Weaviate
Community
Weaviate is an open-source vector database that stores both objects and vectors, allowing for the combination of vector search with structured filtering with the fault tolerance an
Pairs with
Other entries in the index that connect to this one. Click through to see the chain.
topoteretes/cognee
Various
Memory platform for AI Agents in 6 lines of code
zilliztech/mcp-server-milvus
Various
Model Context Protocol Servers for Milvus
AI Getting Started
Community
A Javascript AI getting started stack for weekend projects, including image/text models, vector stores, auth, and deployment configs
Dify
Community
Production-ready platform for agentic workflow development.
Embedbase
Community
A dead-simple API to build LLM-powered apps
Embedchain
Community
Universal memory layer for AI Agents
Epsilla
Community
An all-in-one LLM Agent platform with your private data and knowledge, delivers your production-ready AI Agents on Day 1.
GPTCache
Community
Semantic cache for LLMs. Fully integrated with LangChain and llamaindex.
LangChain
Community
The agent engineering platform.
LlamaIndex
LlamaIndex
The data framework for LLM apps. RAG, ingestion, structured extraction, agents over your data.
Llmware
Community
Unified framework for building enterprise RAG pipelines with small, specialized models
mem0
mem0
Memory layer for AI apps. Personalisation, continuity, and recall as a service.
R2R
Community
SoTA production-ready AI retrieval system. Agentic Retrieval-Augmented Generation (RAG) with a RESTful API.
Semantic Cache Router
Community
Distributed semantic cache and stateful routing system that cuts LLM API costs by returning cached responses for semantically similar queries. Uses ANN vector search (cosine ≥ 0.8)
semantic-coverage
Community
Automated detection of knowledge gaps and blind spots in RAG vector stores.
Agentset
Various
The open-source platform to build AI apps that deliver reliable answers. Production-grade RAG in minutes, no expertise needed.
quivr
Various
Opiniated RAG for integrating GenAI in your apps 🧠 Focus on your product rather than the RAG. Easy integration in existing products with customisation! Any LLM: GPT4, Groq, Llama.
flowiseai
Community
Open source generative AI development platform for building AI agents, LLM orchestration, and more
AIMLPM/markcrawl
Various
Fast Python web crawler for RAG and AI ingestion. Extracts clean Markdown from any site for LLMs and vector stores.
VoxellInc/forge-mcp
Various
[](https://glama.ai/mcp/servers/VoxellInc/forge-mcp) 🎖️ 📇 ☁️ - Official MCP server for Forge, Voxell's hosted text-embedding API. Generate vector embeddings (turbo 1024d, pro 256
Clip-as-a-service
Community
🏄 Scalable embedding, reasoning, ranking for images and sentences with CLIP
currentslab/awesome-vector-search
Community
Collections of vector search related libraries, service and research papers
Haystack
Community
Create agentic, context engineered AI systems using Haystack’s modular and customizable building blocks, built for real-world, production-ready applications.
Infinity
Community
Infinity is a high-throughput, low-latency serving engine for text-embeddings, reranking models, clip, clap and colpali
Langflow
Community
Langflow is a powerful tool for building and deploying AI-powered agents and workflows.
LLMApp
Community
Ready-to-run cloud templates for RAG, AI pipelines, and enterprise search with live data. 🐳Docker-friendly.⚡Always in sync with Sharepoint, Google Drive, S3, Kafka, PostgreSQL, re
Phidata
Community
Build, run, and manage agent platforms.
Quiver
Community
Opiniated RAG for integrating GenAI in your apps 🧠 Focus on your product rather than the RAG. Easy integration in existing products with customisation! Any LLM: GPT4, Groq, Llama.
RagTune
Community
EXPLAIN ANALYZE for RAG retrieval — inspect, debug, benchmark, and tune your retrieval layer
Text-Embeddings-Inference
Community
A blazing fast inference solution for text embeddings models
VDP
Community
🔮 Instill Core is a full-stack AI infrastructure tool for data, model and pipeline orchestration, designed to streamline every aspect of building versatile AI-first applications
Awesome RAG Production
Various
A curated list of battle-tested tools, frameworks, and best practices for building scalable, production-grade Retrieval-Augmented Generation (RAG) systems.
Get the free Developer’s Field Guide
A 27-page field guide to the AI coding workflow with Claude. Claude Code, MCP servers, the prompt patterns that work, and what to delegate. Free.
Enter your work email. We send it straight over, plus a few short notes worth knowing. Unsubscribe any time.
