How this analysis was created: This summary and feature list were written by an AI model that read the project's README and public documentation pages. Each feature links to the documentation it came from; stars, license and language come straight from the GitHub API. The model does not read the source code, and the analysis is refreshed when the project is re-analysed. Learn more on our About page.
Apache Hamilton — portable & expressive data transformation DAGs
The main features of apache/hamilton are: Data Catalogs.
Open-source alternatives to apache/hamilton include: linkedin/datahub — DataHub is a metadata management system and data catalog platform designed to provide a centralized directory for… apache/gravitino — Gravitino is a federated metadata lake and unified data catalog designed to manage tables, files, and AI models across… apache/incubator-gravitino. ckan/ckan — CKAN is an open-source data management platform that provides the foundation for building data portals. It supports… datahub-project/datahub — DataHub is a metadata management platform designed to unify technical, operational, and business context across… apache/atlas — Apache Atlas - Open Metadata Management and Governance capabilities across the Hadoop platform and beyond.
DataHub is a metadata management system and data catalog platform designed to provide a centralized directory for discovering, managing, and documenting datasets across a diverse data stack. It serves as a comprehensive framework for metadata management, incorporating a data governance framework to classify sensitive information and assign ownership for organizational accountability. The platform distinguishes itself through AI-enabled data discovery, which connects large language models to a metadata graph to allow for natural language search and exploration of data assets. It also provides
Gravitino is a federated metadata lake and unified data catalog designed to manage tables, files, and AI models across diverse data sources and cloud storage. It serves as a centralized interface for governing schemas, access controls, and tagging across relational databases, messaging queues, and object stores. The project distinguishes itself by unifying the management of AI assets, such as machine learning models and their version lineages, alongside traditional tabular data. It also implements the Iceberg REST specification to provide a standardized metadata server and proxy for lakehouse
Apache Atlas - Open Metadata Management and Governance capabilities across the Hadoop platform and beyond