awesome-repositories.com
المدونة
MCP
awesome-repositories.com

اكتشف أفضل مستودعات المصادر المفتوحة باستخدام بحث مدعوم بالذكاء الاصطناعي.

استكشفعمليات بحث منسقةبدائل مفتوحة المصدربرمجيات ذاتية الاستضافةالمدونةخريطة الموقع
المشروعخادم MCPحولكيفية ترتيب النتائجالصحافة
قانونيالخصوصيةالشروط
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
opensemanticsearch avatar

opensemanticsearch/open-semantic-search

0
View on GitHub↗
1,181 نجوم·199 تفرعات·Shell·GPL-3.0·9 مشاهداتopensemanticsearch.org↗

Open Semantic Search

Open Semantic Search هي منصة اكتشاف مؤسسية مفتوحة المصدر مصممة لفهرسة، وتحليل، واستكشاف مجموعات المستندات الكبيرة والمتنوعة. تعمل كمحرك بحث شامل ومجموعة تحليلات تحول البيانات غير المهيكلة إلى معلومات مهيكلة من خلال خطوط أنابيب معالجة مؤتمتة.

تتميز المنصة بدمج الاستكشاف الدلالي مع طرق الاسترجاع التقليدية. وتستخدم ربط كيانات الرسم البياني المعرفي وتوسيع الاستعلام القائم على القاموس لربط المفاهيم ذات الصلة، مما يسمح للمستخدمين بالتنقل في مجموعات البيانات بما يتجاوز مطابقة الكلمات الرئيسية البسيطة. ويكتمل ذلك بواجهة مستندة إلى الويب توفر تصفية متعددة الأوجه وتصوراً تفاعلياً للبيانات، مما يمكن المستخدمين من تحديد الأنماط والعلاقات داخل مستودعات مستنداتهم.

يغطي النظام نطاقاً واسعاً من القدرات، بما في ذلك تنقيب النصوص المؤتمت، والتعرف الضوئي على الحروف، والتعليق التوضيحي التعاوني للمستندات. ويدعم استيعاب البيانات المستمر من مصادر مختلفة، مع الحفاظ على فهارس محدثة من خلال المراقبة المؤتمتة وتنسيق مهام الخلفية. تعتمد المعمارية على خدمات مصغرة في حاويات لإدارة مهام الفهرسة والتحليل هذه بكفاءة.

Features

  • Semantic Search Engines - Provides a comprehensive enterprise search engine that uses semantic retrieval and knowledge graph linking to explore large document collections.
  • Enterprise Search - Provides a centralized search platform to index, organize, and retrieve information from large, diverse document repositories.
  • Semantic Search - Expands search queries using thesauri and linguistic heuristics to identify synonyms and related concepts.
  • Enterprise Discovery Platforms - Ingests diverse file formats and metadata to enable full-text search, knowledge graph visualization, and collaborative document annotation.
  • Text Analytics - Automates the extraction of structured insights from unstructured documents using natural language processing and optical character recognition.
  • Full Text Search - Executes keyword-based full-text search across diverse document collections and file formats.
  • Full-Text Inverted Indexes - Powers full-text retrieval by mapping document terms to their locations, enabling rapid keyword lookups and complex boolean queries.
  • Faceted Navigation - Provides interactive faceted navigation to filter and refine large search result sets by metadata attributes.
  • Automated Text Analysis - Applies natural language processing to extract entities, topics, and sentiment from unstructured documents to enrich the searchable index.
  • Entity Linking - Connects extracted entities and metadata into a structured network to support semantic navigation and relationship discovery.
  • Optical Character Recognition - Performs optical character recognition on images and scanned documents to convert graphical content into searchable text.
  • Data Pipelines and ETL - Transforms raw unstructured documents into structured data through sequential stages of extraction, normalization, and semantic enrichment.
  • Collaborative Document Annotations - Enables collaborative document annotation, allowing teams to tag, categorize, and add notes to shared content.
  • Knowledge and Information Management - Organizes internal document collections through collaborative tagging, metadata management, and faceted navigation to improve information accessibility.
  • Data Ingestion Sources - Collects information from local files, websites, feeds, and databases to consolidate disparate content into a single searchable index.
  • Query Expansion - Enhances search precision by automatically augmenting user queries with synonyms and related concepts from controlled vocabularies.
  • Faceted Search Engines - Calculates real-time counts of document attributes to provide interactive filtering and drill-down navigation across large datasets.
  • Faceted Search Implementation - Provides a web-based interface that allows users to navigate large datasets through interactive filters, semantic query expansion, and relationship mapping.
  • Advanced Query Types - Supports complex search syntax including boolean logic, wildcards, and fuzzy matching for precise information retrieval.
  • Automated Indexing - Triggers indexing updates through file system monitoring or notifications to ensure search results reflect content changes in real time.
  • Distributed Text Analytics - Extracts structured data, named entities, and semantic relationships from unstructured documents to uncover patterns and insights automatically.
  • Background Task Runners - Coordinates parallel indexing and analysis tasks using a background task queue to maintain high system throughput.
  • Data Trend Visualizations - Generates interactive charts and relationship graphs from search results to visualize patterns and entity connections.
  • Distributed Task Queues - Coordinates parallel document processing and indexing workflows by distributing heavy analysis tasks across multiple background worker nodes.
  • Filesystem Change Monitors - Monitors file systems and data sources in real time to trigger automated indexing updates when content changes.
  • Search-Based Navigation Interfaces - Offers a web-based interface for full-text, faceted, and exploratory search across document repositories.
  • Data Explorers - Navigates complex datasets using conceptual relationships and thesauri to find relevant information beyond simple keyword matching.

سجل النجوم

مخطط تاريخ النجوم لـ opensemanticsearch/open-semantic-searchمخطط تاريخ النجوم لـ opensemanticsearch/open-semantic-search

بحث بالذكاء الاصطناعي

استكشف المزيد من المستودعات الرائعة

صف ما تحتاجه بلغة بسيطة — وسيقوم الذكاء الاصطناعي بترتيب آلاف المشاريع مفتوحة المصدر المنسقة حسب الصلة.

Start searching with AI

مجموعات مختارة تضم Open Semantic Search

مجموعات منسقة بعناية يظهر فيها Open Semantic Search.
  • أداة تلقائية لفهرسة الملفات على الخوادم
  • أداة لتنظيم الأوراق البحثية الأكاديمية
  • منصة مفتوحة المصدر لفهارس البيانات

الأسئلة الشائعة

ما هي وظيفة opensemanticsearch/open-semantic-search؟

Open Semantic Search هي منصة اكتشاف مؤسسية مفتوحة المصدر مصممة لفهرسة، وتحليل، واستكشاف مجموعات المستندات الكبيرة والمتنوعة. تعمل كمحرك بحث شامل ومجموعة تحليلات تحول البيانات غير المهيكلة إلى معلومات مهيكلة من خلال خطوط أنابيب معالجة مؤتمتة.

ما هي الميزات الرئيسية لـ opensemanticsearch/open-semantic-search؟

الميزات الرئيسية لـ opensemanticsearch/open-semantic-search هي: Semantic Search Engines, Enterprise Search, Semantic Search, Enterprise Discovery Platforms, Text Analytics, Full Text Search, Full-Text Inverted Indexes, Faceted Navigation.

ما هي البدائل مفتوحة المصدر لـ opensemanticsearch/open-semantic-search؟

تشمل البدائل مفتوحة المصدر لـ opensemanticsearch/open-semantic-search: ravendb/ravendb — RavenDB is a multi-model NoSQL document database designed for high-performance, ACID-compliant data storage. It… apache/lucene-solr — This project is a full text search engine and enterprise search infrastructure designed for indexing and retrieving… marqo-ai/marqo — Marqo is an ecommerce product discovery platform, multimodal vector database, and AI search merchandising tool. It… paradedb/paradedb — ParadeDB is a database extension that integrates full-text search, vector database capabilities, and real-time… vendurehq/vendure — Vendure is a Node.js e-commerce engine and headless commerce framework built with NestJS and TypeScript. It serves as… llmware-ai/llmware — llmware is a Python framework for AI agent orchestration and model management, designed to coordinate multi-model…

بدائل مفتوحة المصدر لـ Open Semantic Search

مشاريع مفتوحة المصدر مشابهة، مرتبة حسب عدد الميزات المشتركة مع Open Semantic Search.
  • ravendb/ravendbالصورة الرمزية لـ ravendb

    ravendb/ravendb

    3,961عرض على GitHub↗

    RavenDB is a multi-model NoSQL document database designed for high-performance, ACID-compliant data storage. It persists structured information as schema-flexible JSON documents and utilizes a unit-of-work session pattern to track entity changes and batch modifications into atomic transactions. The platform is built on a distributed architecture that supports horizontal scaling through sharding and ensures high availability via multi-node, master-to-master cluster replication. The database distinguishes itself through a self-optimizing query engine that automatically creates and maintains ind

    C#csharpdatabasedocument-database
    عرض على GitHub↗3,961
  • apache/lucene-solrالصورة الرمزية لـ apache

    apache/lucene-solr

    4,357عرض على GitHub↗

    This project is a full text search engine and enterprise search infrastructure designed for indexing and retrieving large sets of documents. It provides a comprehensive framework for information discovery using ranked results and linguistic analysis. The system integrates high-dimensional vector similarity search for semantic retrieval alongside traditional full-text capabilities. It distinguishes itself through support for geospatial data retrieval, multilingual text processing, and a search suggestion workflow that includes typo-tolerant query completion and spellchecking. The platform cov

    backendinformation-retrievaljava
    عرض على GitHub↗4,357
  • marqo-ai/marqoالصورة الرمزية لـ marqo-ai

    marqo-ai/marqo

    5,022عرض على GitHub↗

    Marqo is an ecommerce product discovery platform, multimodal vector database, and AI search merchandising tool. It provides infrastructure for implementing semantic search and recommendations, allowing shoppers to find products using natural language and images. The platform distinguishes itself through a hybrid ranking pipeline that combines neural semantic scores with business-defined boosting and pinning rules. It features a conversational commerce engine that uses large language models to process user intent and provides a search performance analytics suite for measuring conversion uplift

    Python
    عرض على GitHub↗5,022
  • paradedb/paradedbالصورة الرمزية لـ paradedb

    paradedb/paradedb

    8,370عرض على GitHub↗

    ParadeDB is a database extension that integrates full-text search, vector database capabilities, and real-time analytics directly into a relational engine. It functions as a plugin that adds new storage and query execution capabilities to an existing database architecture. The project distinguishes itself by supporting hybrid search workflows that combine lexical keyword matching with dense and sparse vector similarity in a single query. It utilizes reciprocal rank fusion to merge these ranked result sets and employs logical replication to synchronize data from external instances, removing th

    Rustaggregationsanalyticsbm25
    عرض على GitHub↗8,370
عرض جميع البدائل الـ 30 لـ Open Semantic Search→