awesome-repositories.com
المدونة
MCP
awesome-repositories.com

اكتشف أفضل مستودعات المصادر المفتوحة باستخدام بحث مدعوم بالذكاء الاصطناعي.

استكشفعمليات بحث منسقةبدائل مفتوحة المصدربرمجيات ذاتية الاستضافةالمدونةخريطة الموقع
المشروعخادم MCPحولكيفية ترتيب النتائجالصحافة
قانونيالخصوصيةالشروط
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

5 مستودعات

Awesome GitHub RepositoriesDatabase Metadata Ingestion

Automated extraction of technical and operational metadata from database instances.

Distinct from Database Metadata Discovery: Distinct from general discovery: focuses on the ingestion and synchronization of metadata from specific database instances.

Explore 5 awesome GitHub repositories matching data & databases · Database Metadata Ingestion. Refine with filters or upvote what's useful.

Awesome Database Metadata Ingestion GitHub Repositories

اعثر على أفضل المستودعات باستخدام الذكاء الاصطناعي.سنبحث عن أفضل المستودعات المطابقة باستخدام الذكاء الاصطناعي.
  • open-metadata/openmetadataالصورة الرمزية لـ open-metadata

    open-metadata/OpenMetadata

    14,213عرض على GitHub↗

    OpenMetadata is an enterprise data catalog, metadata platform, and governance suite that functions as a knowledge graph for data assets. It serves as an AI-ready metadata layer, providing governed context and organizational memory to large language model agents via the Model Context Protocol. The platform distinguishes itself by capturing institutional knowledge, linking conversations, decisions, and remediation notes directly to data assets to preserve tribal knowledge. It integrates AI agents to automate metadata governance, such as suggesting descriptions and identifying sensitive data thr

    Automatically extracts and synchronizes metadata from diverse sources using SDKs, APIs, and webhooks.

    TypeScriptcontextcontext-layerdata-catalog
    عرض على GitHub↗14,213
  • linkedin/datahubالصورة الرمزية لـ linkedin

    linkedin/datahub

    12,106عرض على GitHub↗

    DataHub is a metadata management system and data catalog platform designed to provide a centralized directory for discovering, managing, and documenting datasets across a diverse data stack. It serves as a comprehensive framework for metadata management, incorporating a data governance framework to classify sensitive information and assign ownership for organizational accountability. The platform distinguishes itself through AI-enabled data discovery, which connects large language models to a metadata graph to allow for natural language search and exploration of data assets. It also provides

    Automates the extraction of technical and operational metadata from external warehouses and BI tools.

    Python
    عرض على GitHub↗12,106
  • datahub-project/datahubالصورة الرمزية لـ datahub-project

    datahub-project/datahub

    12,141عرض على GitHub↗

    DataHub is a metadata management platform designed to unify technical, operational, and business context across diverse data ecosystems. By utilizing a graph-based metadata model and an event-driven ingestion architecture, it creates a centralized source of truth that maps complex data relationships, lineage, and ownership. This foundational framework enables organizations to maintain a synchronized view of their data landscape, supporting both human-led discovery and automated data operations. The platform distinguishes itself through its focus on grounding artificial intelligence and autono

    Ingests and centralizes metadata from diverse databases and warehouses to enable cross-domain discovery.

    Pythondata-catalogdata-discoverydata-governance
    عرض على GitHub↗12,141
  • amundsen-io/amundsenالصورة الرمزية لـ amundsen-io

    amundsen-io/amundsen

    4,737عرض على GitHub↗

    Amundsen is a data catalog and discovery platform that provides a centralized directory for indexing tables and dashboards. It functions as a metadata management system and search engine, allowing users to locate and understand available data assets across diverse distributed sources. The platform includes capabilities for data lineage tracking to map the origin and movement of datasets between systems. It also serves as a data profiling tool, calculating distribution and quality statistics for individual table columns to provide automated insights into the nature of the data. The system man

    Automates the extraction and synchronization of technical metadata from various database instances.

    Pythonamundsendata-catalogdata-discovery
    عرض على GitHub↗4,737
  • apache/incubator-devlakeالصورة الرمزية لـ apache

    apache/incubator-devlake

    2,940عرض على GitHub↗

    DevLake is a DevOps data platform and analytics tool designed to orchestrate data pipelines that ingest, transform, and sync metadata from external development tools into a unified database. It functions as a system for collecting and normalizing data from source control, CI/CD pipelines, and issue trackers into a standardized schema to enable consistent software delivery analytics. The platform distinguishes itself by transforming tool-specific data into a common domain model, allowing for the calculation of engineering metrics via SQL. It provides specialized frameworks for measuring DORA m

    Uses a modular plugin architecture to ingest metadata from diverse external development tools.

    Godashboard-friendlydatadata-analysis
    عرض على GitHub↗2,940
  1. Home
  2. Data & Databases
  3. Database Metadata Discovery
  4. Database Metadata Ingestion

استكشف الوسوم الفرعية

  • Plugin-Based Ingestion1 وسم فرعيModular interfaces for pulling metadata from diverse external databases and orchestration tools. **Distinct from Database Metadata Ingestion:** Focuses on the modular plugin architecture for ingestion rather than just the extraction process
  • Push-Pull Ingestion MechanismsSupports both active push and passive pull workflows for comprehensive metadata collection. **Distinct from Database Metadata Ingestion:** Focuses on the dual-mode ingestion strategy, distinct from standard database-specific ingestion.