awesome-repositories.com
Blog
MCP
awesome-repositories.com

Descoperă cele mai bune repository-uri open source cu căutare AI.

ExploreazăCăutări recomandateAlternative open-sourceSoftware self-hostedBlogHartă site
ProiectDespreCum realizăm clasamentulPresăServer MCP
LegalConfidențialitateTermeni
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
milvus-io avatar

milvus-io/milvus

0
View on GitHub↗
44,804 stele·4,068 fork-uri·Go·Apache-2.0·15 vizualizărimilvus.io↗

Milvus

Milvus is a specialized vector database engine designed for the indexing, management, and high-speed similarity retrieval of high-dimensional vector embeddings. It functions as a similarity search engine capable of identifying nearest neighbors within large-scale vector spaces, supporting the storage and retrieval of billions of data points while maintaining consistent performance.

The system utilizes a distributed architecture that decouples storage, query, and coordination into independent services, allowing for horizontal scaling across clusters. It employs a global indexing mechanism that builds specialized data structures across immutable, independently indexed segments. This design, combined with a shared-storage decoupled model, enables compute and storage resources to scale independently in cloud environments, while a log-based persistence layer ensures data durability and state recovery.

The platform supports a wide range of data retrieval patterns, including retrieval-augmented generation, hybrid search, and multimodal data retrieval for text, images, and graphs. Deployment options range from lightweight local instances for rapid prototyping to robust standalone setups and fully managed distributed clusters. Documentation includes sizing tools to assist in estimating hardware requirements based on specific data volumes and operational patterns.

Features

  • Similarity Search Engines - Provides high-performance similarity search capabilities for high-dimensional vector data.
  • Vector Databases - Functions as a specialized storage engine optimized for indexing and high-speed similarity retrieval of vector embeddings.
  • Vector Search Engines - Provides high-speed similarity search capabilities across massive high-dimensional datasets.
  • Distributed Database Clusters - Supports distributed architecture to handle horizontal scaling across clusters for large-scale production needs.
  • Retrieval-Augmented Generation Frameworks - Provides the foundational infrastructure for building retrieval-augmented generation applications.
  • Indexing Engines - Implements specialized indexing structures to enable high-performance similarity searches across massive vector datasets.
  • Distributed Data Architectures - Implements a distributed architecture that supports horizontal scaling and high availability across clusters.
  • Distributed Data Infrastructure - Manages and scales complex data storage systems across multiple server nodes for production environments.
  • Retrieval-Augmented Generation - Enhances AI models by providing contextually relevant data retrieved from vector-based knowledge bases.
  • Vector Databases - Scalable open-source vector database for similarity search.
  • Gestionarea datelor - Vector similarity search engine for embeddings.
  • Data Storage Systems - Manages embedding vectors for ML and neural networks.
  • Database Engines - Vector database for embedding management and search.
  • Database Systems - Vector database for scalable similarity search and AI tasks.
  • Database Tools - Vector database.
  • Databases and RAG - Cloud-native vector database for AI.
  • Vector Databases - Cloud-native vector database for scalable similarity search.
  • Infrastructure and Serving - Vector database for similarity search.
  • Backend and Infrastructure - Vector database for AI and machine learning.
  • Storage Decoupling - Separates compute and storage nodes to allow independent scaling of processing power and data capacity.
  • Hybrid Search Systems - Implements hybrid search capabilities to combine vector similarity with other retrieval methods.
  • Multimodal Databases - Acts as a unified storage environment for organizing and retrieving complex data types like text and images.
  • Microservice Architectures - Decouples storage, query, and coordination into independent services to enable horizontal scaling.
  • Multimodal Retrieval Systems - Enables searching across diverse media types by utilizing shared vector representations.
  • Multimodal Search Engines - Supports multimodal search patterns to query across diverse data types.
  • Data Partitioning - Partitions data into immutable segments to optimize memory usage and parallel search performance.
  • Standalone Databases - Provides a standalone configuration for single-machine environments.
  • Write-Ahead Logs - Ensures data durability and consistent state recovery by recording all mutations in a distributed message log.

Istoric stele

Graficul istoricului de stele pentru milvus-io/milvusGraficul istoricului de stele pentru milvus-io/milvus

Căutare AI

Explorează mai multe repository-uri excelente

Descrie ce ai nevoie în limbaj simplu — AI-ul sortează mii de proiecte open source selectate în funcție de relevanță.

Start searching with AI

Alternative open-source pentru Milvus

Proiecte open-source similare, clasificate după numărul de funcționalități comune cu Milvus.
  • qdrant/qdrantAvatar qdrant

    qdrant/qdrant

    32,372Vezi pe GitHub↗

    Qdrant is a high-performance vector similarity database designed to store, index, and search high-dimensional vectors alongside structured metadata. It functions as a distributed search engine that manages large-scale data clusters, providing low-latency retrieval and complex filtering capabilities. The system is built to serve as a specialized middleware layer, connecting machine learning pipelines and AI agents to persistent storage for intelligent information retrieval and recommendation tasks. The platform distinguishes itself through advanced retrieval techniques, including support for h

    Rustai-searchai-search-engineembeddings-similarity
    Vezi pe GitHub↗32,372
  • chroma-core/chromaAvatar chroma-core

    chroma-core/chroma

    26,198Vezi pe GitHub↗

    Chroma is a specialized vector database designed to index and retrieve high-dimensional data representations for semantic similarity search. It functions as a comprehensive platform for information retrieval, enabling the storage and management of unstructured documents alongside structured metadata. By mapping data into numerical representations, the system facilitates rapid similarity lookups across large datasets. The platform distinguishes itself through a hybrid search infrastructure that combines dense vector embeddings with sparse keyword and regular expression matching to balance sema

    Rustaidatabasedocument-retrieval
    Vezi pe GitHub↗26,198
  • weaviate/weaviateAvatar weaviate

    weaviate/weaviate

    15,620Vezi pe GitHub↗

    Weaviate is an AI-native vector database designed to store and index high-dimensional vector embeddings alongside traditional data objects. It serves as a backend infrastructure for retrieval-augmented generation, enabling applications to ground language model responses in private, context-aware data. The platform distinguishes itself by combining vector similarity search with traditional keyword filtering through a hybrid storage architecture. It integrates directly with external machine learning models to automate the generation of embeddings and perform complex inference tasks during inges

    Goapproximate-nearest-neighbor-searchgenerative-searchgrpc
    Vezi pe GitHub↗15,620
  • lancedb/lancedbAvatar lancedb

    lancedb/lancedb

    9,031Vezi pe GitHub↗

    LanceDB is a vector database and columnar data store designed to function as a versioned dataset manager and vector search engine. It serves as a high-performance backend for indexing and retrieving high-dimensional embeddings, providing the foundation for machine learning data pipelines. The system distinguishes itself through a combination of cloud-native object storage and immutable version tracking, allowing for data time-travel and reproducible AI experiments. It integrates hybrid search capabilities, merging dense vector similarity with BM25 full-text search and SQL-like scalar filters

    HTMLapproximate-nearest-neighbor-searchimage-searchnearest-neighbor-search
    Vezi pe GitHub↗9,031
Vezi toate cele 30 alternative pentru Milvus→

Întrebări frecvente

Ce face milvus-io/milvus?

Milvus is a specialized vector database engine designed for the indexing, management, and high-speed similarity retrieval of high-dimensional vector embeddings. It functions as a similarity search engine capable of identifying nearest neighbors within large-scale vector spaces, supporting the storage and retrieval of billions of data points while maintaining consistent performance.

Care sunt principalele funcționalități ale milvus-io/milvus?

Principalele funcționalități ale milvus-io/milvus sunt: Similarity Search Engines, Vector Databases, Vector Search Engines, Distributed Database Clusters, Retrieval-Augmented Generation Frameworks, Indexing Engines, Distributed Data Architectures, Distributed Data Infrastructure.

Care sunt câteva alternative open-source pentru milvus-io/milvus?

Alternativele open-source pentru milvus-io/milvus includ: qdrant/qdrant — Qdrant is a high-performance vector similarity database designed to store, index, and search high-dimensional vectors… chroma-core/chroma — Chroma is a specialized vector database designed to index and retrieve high-dimensional data representations for… weaviate/weaviate — Weaviate is an AI-native vector database designed to store and index high-dimensional vector embeddings alongside… lancedb/lancedb — LanceDB is a vector database and columnar data store designed to function as a versioned dataset manager and vector… infiniflow/infinity — Infinity is a distributed vector database and multimodal vector store designed to manage large-scale datasets for… activeloopai/deeplake — DeepLake is AI data infrastructure consisting of a multimodal data lake, a hybrid search engine, and a serverless…