42 مستودعات
Tools for configuring, querying, and maintaining search indices.
Distinguishing note: Focuses on index-level management rather than general database querying.
Explore 42 awesome GitHub repositories matching data & databases · Search Index Management. Refine with filters or upvote what's useful.
This project is an AI agent orchestration platform that provides a visual environment for building, testing, and deploying complex automation workflows. It functions as a low-code development interface where users can chain discrete functional blocks into dependency-aware pipelines to integrate artificial intelligence with external data and services. The platform supports the creation of intelligent conversational agents, automated business processes, and multi-service API orchestrations within a unified workspace. The platform distinguishes itself through its event-driven integration engine,
Workflow Platform searches and manages indices by performing operations such as querying data, updating records, configuring settings, and handling maintenance tasks through a unified interface.
This project is a feature-rich Go client library designed for interacting with Redis. It serves as a comprehensive interface for managing remote data stores, enabling developers to execute standard database commands, handle complex data structures, and perform asynchronous operations within Go applications. The library distinguishes itself through its support for advanced Redis capabilities, including connection pooling, pipelining, and transactional integrity. It provides specialized primitives for managing distributed clusters, including automated topology updates and request routing to sha
Enables the creation and management of topic-specific indexes for data isolation.
Sonic is a high-performance, lightweight search backend designed to provide real-time full-text search and autocomplete capabilities for applications. It functions as a persistent indexing server that maps text terms to object identifiers, allowing developers to integrate rapid search functionality without storing raw document content directly within the search engine. The system distinguishes itself through a specialized graph-based index that enables real-time word prediction and typo correction. Communication is handled via a custom, low-latency binary protocol over raw TCP sockets, which
Maintains search indices through background maintenance, data removal, and configuration tasks.
Temporal is a distributed workflow orchestration engine designed to manage fault-tolerant, stateful, and long-running background processes. It functions as a platform for coordinating complex cross-service operations, ensuring consistency and reliability in distributed environments by decoupling workflow orchestration from task execution. The platform distinguishes itself through a deterministic, event-sourced execution model that reconstructs workflow state by re-executing code from an immutable event log. This approach isolates non-deterministic side effects into managed activities, allowin
Defines custom fields for indexing workflows, enabling advanced filtering and searching of execution history.
Zincsearch is a high-performance, self-hosted full-text search engine and database written in Go. It provides a lightweight infrastructure for indexing and searching unstructured text data, specializing in log and event analysis through a schemaless indexing model. The system is designed as a resource-efficient alternative to heavier search infrastructure, featuring an API surface compatible with Elasticsearch for indexing and querying documents. It distinguishes itself by packaging the entire server and its built-in web search interface into a single statically linked binary. The engine cov
Provides administrative control to create, retrieve, and delete indices while defining field mappings.
Zinc is a high-performance full-text search engine written in Go. It provides a schema-less document index that organizes arbitrary datasets into searchable structures without requiring a predefined data format. The engine features an API compatible with Elasticsearch for indexing and querying data, which facilitates the ingestion of single and bulk records. It is designed as an in-process search engine that embeds indexing and retrieval logic within a single binary to operate with minimal system resource overhead. The system includes a built-in web-based management interface for executing s
Organizes arbitrary datasets into searchable indexes for efficient full-text search with minimal overhead.
Grav is a flat-file content management system that eliminates the need for a traditional database by storing site content and configuration in human-readable Markdown and YAML files. Built as a modular PHP web framework, it uses a hierarchical page routing system where the physical directory structure directly determines the site's URL paths. The platform is distinguished by its event-driven plugin architecture and a command-line interface that prioritizes system administration, deployment, and maintenance tasks. It utilizes a blueprint-driven system to generate administrative forms from stru
Provides automated and manual tools for maintaining, optimizing, and managing search indices.
Annoy is a C++ library designed for approximate nearest neighbor search in high-dimensional vector spaces. It functions as a vector similarity search engine that constructs static, disk-based data structures to facilitate fast lookups. By mapping identifiers to vector data and persisting these structures to disk, the library enables efficient, memory-mapped access to large datasets. The project distinguishes itself through the use of random projection trees and distance-metric-based partitioning, which organize data into hierarchical binary trees to balance search precision against computatio
Enables saving and loading of static search indices to disk for distribution and reuse across system environments.
Redis is a high-performance in-memory key-value store that functions as a distributed cache, message broker, and NoSQL database. It provides sub-millisecond read and write access to data stored in RAM and can operate as a vector database for indexing high-dimensional embeddings. The system supports a wide range of data storage and synchronization primitives, including the management of strings, hashes, lists, sets, and JSON documents. It enables real-time data operations through atomic transactions, hybrid persistence using snapshots and append-only logs, and high-availability configurations
Provides tools for configuring and maintaining named search indices to isolate datasets and manage index lifecycles.
RedisInsight is a graphical user interface and management tool for browsing, analyzing, and administering Redis databases. It provides a visual environment for exploring key-value data structures, managing database instances, and performing data analysis across different operating systems and deployments. The tool distinguishes itself by providing dedicated visual managers for complex operations, including a vector database manager for configuring embeddings and similarity searches, a query workbench for executing raw commands and Lua scripts, and a performance monitoring dashboard for tracki
Provides tools to modify existing index schemas, add new fields, or remove indexes entirely.
Dejavu is a containerized administration panel and web interface for managing data within Elasticsearch and OpenSearch clusters. It serves as a search index management tool for browsing, editing, and deleting records through a visual explorer rather than raw API queries. The project distinguishes itself by providing a search interface prototyping tool. This allows users to visually design search screens to test result relevancy and export the final layout configuration as usable code. The tool covers broad data management capabilities, including structured data import from CSV or JSON files
Provides a comprehensive visual interface for configuring, querying, and maintaining search indices in Elasticsearch and OpenSearch clusters.
The mongo-go-driver is a Go library for building applications that integrate with a MongoDB document store. It enables the storage and retrieval of flexible document data by providing a bridge between Go backends and the database. The driver implements specialized capabilities for semantic vector search, allowing the handling and execution of high-dimensional vector data for similarity-based retrieval. It also supports full-text search via linguistic analysis and programmatic search index management. The project covers a broad range of database operations, including document-based CRUD, bulk
Offers a programmatic interface to manage search indexes and configure queryable encryption.
The google-indexing-script is a Google Indexing API Manager designed to automate page discovery and indexing requests to accelerate search engine visibility. It includes an SEO Indexing Monitor to track page status and an automated reporting engine for analyzing indexing trends and performance. The project features a keyword intent clustering tool to group pages by topic and funnel stage and a search console analytics dashboard that unifies search data with web analytics. It provides specialized utilities for detecting keyword cannibalization and identifying striking distance keywords to prio
Automates requests to search engine APIs to ensure new or updated pages are discovered and indexed quickly.
This project is a Go client library and API wrapper for interacting with Elasticsearch clusters. It serves as a programmatic interface for managing documents, indices, and cluster health, allowing Go applications to perform search and indexing operations via the REST API. The library functions as a distributed search orchestrator, providing specialized tools for high-throughput data ingestion and cluster administration. It features a buffered bulk processor with exponential backoff retries for optimizing write performance and supports automated index lifecycle transitions and historical data
Offers comprehensive tools for configuring, querying, and maintaining search indices.
Lance is a versioned columnar data format and storage engine designed as a multimodal AI lakehouse. It serves as a vector database storage engine and a cloud object store dataset manager, organizing images, video, audio, and embeddings into a unified format optimized for machine learning workflows. The project distinguishes itself by combining a columnar layout for structured data with a specialized blob store for large multimodal tensors. It implements a hybrid search engine that integrates vector similarity search, full-text search, and SQL analytics on a single dataset, supported by a stor
Treats search indices as versioned table objects decoupled from file encoding to support independent evolution.
The official Go client for Elasticsearch
Creates and deletes search indexes to organize data into logical containers.
Hound is a self-hosted code search engine that indexes source code repositories and provides fast regular expression search results using a trigram-based index. It is designed to be deployed on your own infrastructure, enabling you to search across multiple public and private code repositories simultaneously. The engine builds its search index by decomposing source code into three-character trigrams, which allows for fast substring matching with regular expressions. It supports searching across multiple repositories in parallel, returning results from the pre-built trigram index. Hound can in
Configures the frequency of repository re-indexing to balance freshness against system load.
Lettuce is a Redis client library for Java that provides synchronous, asynchronous, and reactive programming models for interacting with Redis databases. It supports standalone, cluster, sentinel, pub/sub, and search operations through a single thread-safe connection model that handles command execution without blocking the calling thread. The library distinguishes itself through its reactive streams integration with Project Reactor, enabling non-blocking, backpressure-aware data processing with Mono and Flux types. It offers cluster slot routing that transparently handles MOVED and ASK redir
Creates, updates, and removes aliases to switch between index versions without changing application queries.
هذا المشروع عبارة عن مجموعة أدوات تطوير برمجيات (SDK) وأداة لإدارة العناقيد (clusters) مخصصة لـ PHP. يعمل كـ SDK للبحث بالنص الكامل وواجهة للبحث المتجهي (vector search)، مما يتيح للتطبيقات إجراء عمليات بحث معجمية، وتقريبية، ودلالية على البيانات المفهرسة. تطبق المكتبة عميل HTTP متوافق مع معيار PSR 7 لضمان التوافق عبر بيئات مختلفة من خلال واجهات مراسلة موحدة. كما توفر واجهة متخصصة لاسترجاع التضمينات (embeddings) وتنفيذ مهام الاسترجاع الدلالي باستخدام البيانات المتجهية. تغطي قدرات المشروع مجموعة واسعة من المهام الإدارية والتشغيلية، بما في ذلك إدارة فهارس البحث، ومراقبة صحة العناقيد، وعمليات دورة حياة المستندات. ويدعم المشروع طرق استعلام متنوعة مثل SQL وEQL وES|QL، إلى جانب تجميع البيانات والتحليل الجغرافي المكاني. بالإضافة إلى ذلك، يوفر أدوات لتنسيق تعلم الآلة، واكتشاف الشذوذ، وإدارة الهوية والوصول.
Provides tools for configuring, querying, and maintaining search indices and mappings.
Claude-context is a retrieval-augmented generation pipeline and semantic code search tool. It functions as an LLM codebase indexer and RAG context provider, designed to index local directories and retrieve relevant code files to provide context for large language models. The system operates as a hybrid search engine that combines keyword matching with dense vector search. This allows for the retrieval of code snippets and logic using natural language queries based on meaning rather than exact text matches. The project covers codebase indexing and search index management, utilizing asynchrono
Tracks indexing progress, manages codebase index clearing, and handles incremental file updates.