awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
MarquezProject avatar

MarquezProject/marquez

0
View on GitHub↗
2,215 stars·401 forks·Java·Apache-2.0·8 viewsmarquezproject.ai↗

Marquez

Collect, aggregate, and visualize a data ecosystem's metadata

Features

  • Data Catalogs - Open-source metadata service for data lineage and lifecycle management.
  • Data Management - Metadata collection and visualization for data ecosystems.
  • Metadata Management - Aggregates and visualizes metadata across data ecosystems.

Star history

Star history chart for marquezproject/marquezStar history chart for marquezproject/marquez

How this analysis was created: This summary and feature list were written by an AI model that read the project's README and public documentation pages. Each feature links to the documentation it came from; stars, license and language come straight from the GitHub API. The model does not read the source code, and the analysis is refreshed when the project is re-analysed. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Open-source alternatives to Marquez

Similar open-source projects, ranked by how many features they share with Marquez.
  • apache/atlasapache avatar

    apache/atlas

    2,110View on GitHub↗

    Apache Atlas - Open Metadata Management and Governance capabilities across the Hadoop platform and beyond

    Java
    View on GitHub↗2,110
  • datahub-project/datahubdatahub-project avatar

    datahub-project/datahub

    12,141View on GitHub↗

    DataHub is a metadata management platform designed to unify technical, operational, and business context across diverse data ecosystems. By utilizing a graph-based metadata model and an event-driven ingestion architecture, it creates a centralized source of truth that maps complex data relationships, lineage, and ownership. This foundational framework enables organizations to maintain a synchronized view of their data landscape, supporting both human-led discovery and automated data operations. The platform distinguishes itself through its focus on grounding artificial intelligence and autono

    Pythondata-catalogdata-discoverydata-governance
    View on GitHub↗12,141
  • amundsen-io/amundsenamundsen-io avatar

    amundsen-io/amundsen

    4,737View on GitHub↗

    Amundsen is a data catalog and discovery platform that provides a centralized directory for indexing tables and dashboards. It functions as a metadata management system and search engine, allowing users to locate and understand available data assets across diverse distributed sources. The platform includes capabilities for data lineage tracking to map the origin and movement of datasets between systems. It also serves as a data profiling tool, calculating distribution and quality statistics for individual table columns to provide automated insights into the nature of the data. The system man

    Pythonamundsendata-catalogdata-discovery
    View on GitHub↗4,737
  • open-metadata/openmetadataopen-metadata avatar

    open-metadata/OpenMetadata

    14,213View on GitHub↗

    OpenMetadata is an enterprise data catalog, metadata platform, and governance suite that functions as a knowledge graph for data assets. It serves as an AI-ready metadata layer, providing governed context and organizational memory to large language model agents via the Model Context Protocol. The platform distinguishes itself by capturing institutional knowledge, linking conversations, decisions, and remediation notes directly to data assets to preserve tribal knowledge. It integrates AI agents to automate metadata governance, such as suggesting descriptions and identifying sensitive data thr

    TypeScriptcontextcontext-layerdata-catalog
    View on GitHub↗14,213
See all 30 alternatives to Marquez→

Frequently asked questions

What does marquezproject/marquez do?

Collect, aggregate, and visualize a data ecosystem's metadata

What are the main features of marquezproject/marquez?

The main features of marquezproject/marquez are: Data Catalogs, Data Management, Metadata Management.

What are some open-source alternatives to marquezproject/marquez?

Open-source alternatives to marquezproject/marquez include: datahub-project/datahub — DataHub is a metadata management platform designed to unify technical, operational, and business context across… open-metadata/openmetadata — OpenMetadata is an enterprise data catalog, metadata platform, and governance suite that functions as a knowledge… amundsen-io/amundsen — Amundsen is a data catalog and discovery platform that provides a centralized directory for indexing tables and… apache/atlas — Apache Atlas - Open Metadata Management and Governance capabilities across the Hadoop platform and beyond. linkedin/datahub — DataHub is a metadata management system and data catalog platform designed to provide a centralized directory for… algolia/algoliasearch-rails — AlgoliaSearch integration to your favorite ORM.