awesome-repositories.com
ब्लॉग
MCP
awesome-repositories.com

AI-संचालित खोज के साथ बेहतरीन ओपन-सोर्स रिपॉजिटरी खोजें।

एक्सप्लोर करेंक्यूरेटेड खोजेंओपन-सोर्स विकल्पसेल्फ-होस्टेड सॉफ्टवेयरब्लॉगसाइटमैप
प्रोजेक्टMCP सर्वरहमारे बारे मेंहम रैंकिंग कैसे करते हैंप्रेस
कानूनीगोपनीयताशर्तें
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

5 रिपॉजिटरी

Awesome GitHub RepositoriesDatabase Metadata Ingestion

Automated extraction of technical and operational metadata from database instances.

Distinct from Database Metadata Discovery: Distinct from general discovery: focuses on the ingestion and synchronization of metadata from specific database instances.

Explore 5 awesome GitHub repositories matching data & databases · Database Metadata Ingestion. Refine with filters or upvote what's useful.

Awesome Database Metadata Ingestion GitHub Repositories

AI के साथ बेहतरीन रिपॉजिटरी खोजें।हम AI का उपयोग करके सबसे सटीक रिपॉजिटरी खोजेंगे।
  • open-metadata/openmetadataopen-metadata का अवतार

    open-metadata/OpenMetadata

    14,213GitHub पर देखें↗

    OpenMetadata is an enterprise data catalog, metadata platform, and governance suite that functions as a knowledge graph for data assets. It serves as an AI-ready metadata layer, providing governed context and organizational memory to large language model agents via the Model Context Protocol. The platform distinguishes itself by capturing institutional knowledge, linking conversations, decisions, and remediation notes directly to data assets to preserve tribal knowledge. It integrates AI agents to automate metadata governance, such as suggesting descriptions and identifying sensitive data thr

    Automatically extracts and synchronizes metadata from diverse sources using SDKs, APIs, and webhooks.

    TypeScriptcontextcontext-layerdata-catalog
    GitHub पर देखें↗14,213
  • linkedin/datahublinkedin का अवतार

    linkedin/datahub

    12,106GitHub पर देखें↗

    DataHub is a metadata management system and data catalog platform designed to provide a centralized directory for discovering, managing, and documenting datasets across a diverse data stack. It serves as a comprehensive framework for metadata management, incorporating a data governance framework to classify sensitive information and assign ownership for organizational accountability. The platform distinguishes itself through AI-enabled data discovery, which connects large language models to a metadata graph to allow for natural language search and exploration of data assets. It also provides

    Automates the extraction of technical and operational metadata from external warehouses and BI tools.

    Python
    GitHub पर देखें↗12,106
  • datahub-project/datahubdatahub-project का अवतार

    datahub-project/datahub

    12,141GitHub पर देखें↗

    DataHub is a metadata management platform designed to unify technical, operational, and business context across diverse data ecosystems. By utilizing a graph-based metadata model and an event-driven ingestion architecture, it creates a centralized source of truth that maps complex data relationships, lineage, and ownership. This foundational framework enables organizations to maintain a synchronized view of their data landscape, supporting both human-led discovery and automated data operations. The platform distinguishes itself through its focus on grounding artificial intelligence and autono

    Ingests and centralizes metadata from diverse databases and warehouses to enable cross-domain discovery.

    Pythondata-catalogdata-discoverydata-governance
    GitHub पर देखें↗12,141
  • amundsen-io/amundsenamundsen-io का अवतार

    amundsen-io/amundsen

    4,737GitHub पर देखें↗

    Amundsen is a data catalog and discovery platform that provides a centralized directory for indexing tables and dashboards. It functions as a metadata management system and search engine, allowing users to locate and understand available data assets across diverse distributed sources. The platform includes capabilities for data lineage tracking to map the origin and movement of datasets between systems. It also serves as a data profiling tool, calculating distribution and quality statistics for individual table columns to provide automated insights into the nature of the data. The system man

    Automates the extraction and synchronization of technical metadata from various database instances.

    Pythonamundsendata-catalogdata-discovery
    GitHub पर देखें↗4,737
  • apache/incubator-devlakeapache का अवतार

    apache/incubator-devlake

    2,940GitHub पर देखें↗

    DevLake is a DevOps data platform and analytics tool designed to orchestrate data pipelines that ingest, transform, and sync metadata from external development tools into a unified database. It functions as a system for collecting and normalizing data from source control, CI/CD pipelines, and issue trackers into a standardized schema to enable consistent software delivery analytics. The platform distinguishes itself by transforming tool-specific data into a common domain model, allowing for the calculation of engineering metrics via SQL. It provides specialized frameworks for measuring DORA m

    Uses a modular plugin architecture to ingest metadata from diverse external development tools.

    Godashboard-friendlydatadata-analysis
    GitHub पर देखें↗2,940
  1. Home
  2. Data & Databases
  3. Database Metadata Discovery
  4. Database Metadata Ingestion

सब-टैग एक्सप्लोर करें

  • Plugin-Based Ingestion1 सब-टैगModular interfaces for pulling metadata from diverse external databases and orchestration tools. **Distinct from Database Metadata Ingestion:** Focuses on the modular plugin architecture for ingestion rather than just the extraction process
  • Push-Pull Ingestion MechanismsSupports both active push and passive pull workflows for comprehensive metadata collection. **Distinct from Database Metadata Ingestion:** Focuses on the dual-mode ingestion strategy, distinct from standard database-specific ingestion.