awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
Back to grai-io/grai-core

Projects sharing features with Grai Core

30 open-source projects similar to grai-io/grai-core, ranked by shared indexed features. Tags may describe platforms or build tools rather than the same primary purpose. Check each project’s use case, license, and deployment requirements before treating it as a replacement.

  • sodadata/soda-coresodadata avatar

    sodadata/soda-core

    2,292View on GitHub↗
    Pythondata-contractsdata-engineeringdata-governance
    View on GitHub↗2,292
  • elementary-data/elementaryelementary-data avatar

    elementary-data/elementary

    2,373View on GitHub↗

    Elementary OSS: dbt-native data observability

    HTML
    View on GitHub↗2,373
  • odpi/egeriaodpi avatar

    odpi/egeria

    916View on GitHub↗

    Egeria provides the Apache-2.0 licensed open metadata and governance type system, frameworks, APIs, event payloads and interchange protocols to enable tools, engines and platforms to exchange metadata in order to get the best value from data, whilst ensuring it is properly governed.

    Java
    View on GitHub↗916
  • apache/gravitinoapache avatar

    apache/gravitino

    2,866View on GitHub↗

    Gravitino is a federated metadata lake and unified data catalog designed to manage tables, files, and AI models across diverse data sources and cloud storage. It serves as a centralized interface for governing schemas, access controls, and tagging across relational databases, messaging queues, and object stores. The project distinguishes itself by unifying the management of AI assets, such as machine learning models and their version lineages, alongside traditional tabular data. It also implements the Iceberg REST specification to provide a standardized metadata server and proxy for lakehouse

    Javaai-catalogdata-catalogdatalake
    View on GitHub↗2,866

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Find more with AI search
  • datahub-project/datahubdatahub-project avatar

    datahub-project/datahub

    12,141View on GitHub↗

    DataHub is a metadata management platform designed to unify technical, operational, and business context across diverse data ecosystems. By utilizing a graph-based metadata model and an event-driven ingestion architecture, it creates a centralized source of truth that maps complex data relationships, lineage, and ownership. This foundational framework enables organizations to maintain a synchronized view of their data landscape, supporting both human-led discovery and automated data operations. The platform distinguishes itself through its focus on grounding artificial intelligence and autono

    Pythondata-catalogdata-discoverydata-governance
    View on GitHub↗12,141
  • dagworks-inc/hamiltondagworks-inc avatar

    dagworks-inc/hamilton

    2,528View on GitHub↗

    Apache Hamilton helps data scientists and engineers define testable, modular, self-documenting dataflows, that encode lineage/tracing and metadata. Runs and scales everywhere python does.

    Jupyter Notebook
    View on GitHub↗2,528
  • opendatadiscovery/odd-platformopendatadiscovery avatar

    opendatadiscovery/odd-platform

    1,410View on GitHub↗

    Next-Gen Data Discovery and Data Observability Platform

    Java
    View on GitHub↗1,410
  • open-metadata/openmetadataopen-metadata avatar

    open-metadata/OpenMetadata

    14,213View on GitHub↗

    OpenMetadata is an enterprise data catalog, metadata platform, and governance suite that functions as a knowledge graph for data assets. It serves as an AI-ready metadata layer, providing governed context and organizational memory to large language model agents via the Model Context Protocol. The platform distinguishes itself by capturing institutional knowledge, linking conversations, decisions, and remediation notes directly to data assets to preserve tribal knowledge. It integrates AI agents to automate metadata governance, such as suggesting descriptions and identifying sensitive data thr

    TypeScriptcontextcontext-layerdata-catalog
    View on GitHub↗14,213
  • linkedin/datahublinkedin avatar

    linkedin/datahub

    12,106View on GitHub↗

    DataHub is a metadata management system and data catalog platform designed to provide a centralized directory for discovering, managing, and documenting datasets across a diverse data stack. It serves as a comprehensive framework for metadata management, incorporating a data governance framework to classify sensitive information and assign ownership for organizational accountability. The platform distinguishes itself through AI-enabled data discovery, which connects large language models to a metadata graph to allow for natural language search and exploration of data assets. It also provides

    Python
    View on GitHub↗12,106
  • openaddresses/openaddressesopenaddresses avatar

    openaddresses/openaddresses

    3,113View on GitHub↗

    OpenAddresses is an open-source geospatial data aggregator and directory that collects public domain and open-license address, parcel, and building datasets from governments and organizations worldwide. It functions as a global index and data warehouse for locating and distributing free geospatial records. The project operates a normalization pipeline that cleans and standardizes diverse source formats into a consistent global coordinate and attribute schema. This process includes a crowdsourced curation pipeline and programmatic quality validation to verify the spatial accuracy and formattin

    JavaScriptaddressesgeocodinghacktoberfest
    View on GitHub↗3,113
  • cachethq/cachetcachethq avatar

    cachethq/cachet

    14,932View on GitHub↗

    Cachet is a self-hosted, open-source status page system designed to communicate service uptime, incident history, and infrastructure performance to end users. It provides a centralized dashboard for managing the operational lifecycle of system components, tracking service disruptions, and scheduling maintenance windows. The platform distinguishes itself through a comprehensive RESTful API that enables programmatic status page management and automated incident reporting. It supports deep integration with external monitoring tools, allowing for the synchronization of performance metrics and the

    PHPcachetlaravelphp
    View on GitHub↗14,932
  • bloomberg/goldpingerbloomberg avatar

    bloomberg/goldpinger

    2,706View on GitHub↗

    Debugging tool for Kubernetes which tests and displays connectivity between nodes in the cluster.

    JavaScriptkuberneteskubernetes-monitoringprometheus
    View on GitHub↗2,706
  • deep-on/dockprobedeep-on avatar

    deep-on/dockprobe

    15View on GitHub↗

    Lightweight Docker monitoring dashboard with anomaly detection & Telegram alerts. One-liner install, zero config.

    Python
    View on GitHub↗15
  • cilium/tetragoncilium avatar

    cilium/tetragon

    4,753View on GitHub↗

    Tetragon is an eBPF-based runtime security and observability toolset designed for Linux and Kubernetes environments. It functions as a security policy manager, observability agent, and enforcement engine that hooks into kernel functions and tracepoints to detect privilege escalation, container escapes, and unauthorized system activity. The project distinguishes itself through its ability to perform real-time, in-kernel enforcement, allowing it to synchronously terminate malicious processes or modify function return values before a system call completes. It provides deep Kubernetes integration

    C
    View on GitHub↗4,753
  • avito-tech/bioyinoavito-tech avatar

    avito-tech/bioyino

    236View on GitHub↗

    High performance and high-precision multithreaded StatsD server

    Rust
    View on GitHub↗236
  • apache/atlasapache avatar

    apache/atlas

    2,110View on GitHub↗

    Apache Atlas - Open Metadata Management and Governance capabilities across the Hadoop platform and beyond

    Java
    View on GitHub↗2,110
  • amerkurev/dokuamerkurev avatar

    amerkurev/doku

    419View on GitHub↗

    💽 Doku - Docker disk usage dashboard

    Python
    View on GitHub↗419
  • arachnys/cabotarachnys avatar

    arachnys/cabot

    5,668View on GitHub↗

    Cabot is a self-hosted incident management platform and monitoring alert manager. It serves as a centralized system for receiving external monitoring data, managing on-call responder rotations, and dispatching incident notifications to administrators to ensure prompt resolution of system failures. The platform provides a containerized monitoring interface and a REST API for programmatically creating and modifying monitoring data and incident reports. It coordinates the routing of critical system alerts and manages the scheduling of personnel for on-call rotations. The system includes capabil

    JavaScript
    View on GitHub↗5,668
  • codeswhat/drydockCodesWhat avatar

    CodesWhat/drydock

    204View on GitHub↗

    Open source container update monitoring — 23 registries, 20 notification triggers, audit log, OIDC auth, Prometheus metrics, and a modern dashboard.

    TypeScript
    View on GitHub↗204
  • codeofmario/wiremapcodeofmario avatar

    codeofmario/wiremap

    4View on GitHub↗

    Visual Docker network topology explorer — single binary, zero dependencies.

    TypeScript
    View on GitHub↗4
  • collectd/collectdcollectd avatar

    collectd/collectd

    3,358View on GitHub↗

    The system statistics collection daemon. Please send Pull Requests here!

    C
    View on GitHub↗3,358
  • containership/konstellateC

    containership/konstellate

    0View on GitHub↗
    View on GitHub↗0
  • cortexproject/cortexcortexproject avatar

    cortexproject/cortex

    5,751View on GitHub↗

    Cortex is an open-source, horizontally scalable metrics platform that ingests, stores, and queries Prometheus-compatible time-series data with multi-tenant isolation. It accepts metrics via Prometheus remote write and OpenTelemetry, executes PromQL queries against both recent and historical data, and provides a Prometheus-compatible alerting and recording rule engine with an integrated Alertmanager. The system is built as a set of independently scalable microservices that use hash-ring-based sharding, gossip-based cluster membership, and tenant-aware object storage to distribute workloads acro

    Gocncfhacktoberfestkubernetes
    View on GitHub↗5,751
  • appdynamics/docker-monitoring-extensionAppdynamics avatar

    Appdynamics/docker-monitoring-extension

    5View on GitHub↗

    Docker Monitoring Extension

    Java
    View on GitHub↗5
  • databand-ai/dbnddataband-ai avatar

    databand-ai/dbnd

    267View on GitHub↗

    DBND an open source framework for building and tracking data pipelines. DBND is used for processes ranging from data ingestion, preparation, machine learning model training and production.

    Python
    View on GitHub↗267
  • amundsen-io/amundsenamundsen-io avatar

    amundsen-io/amundsen

    4,737View on GitHub↗

    Amundsen is a data catalog and discovery platform that provides a centralized directory for indexing tables and dashboards. It functions as a metadata management system and search engine, allowing users to locate and understand available data assets across diverse distributed sources. The platform includes capabilities for data lineage tracking to map the origin and movement of datasets between systems. It also serves as a data profiling tool, calculating distribution and quality statistics for individual table columns to provide automated insights into the nature of the data. The system man

    Pythonamundsendata-catalogdata-discovery
    View on GitHub↗4,737
  • cloudprober/cloudprobercloudprober avatar

    cloudprober/cloudprober

    687View on GitHub↗

    An active monitoring software to detect failures before your customers do.

    Go
    View on GitHub↗687
  • deepflowys/deepflowD

    deepflowys/deepflow

    0View on GitHub↗
    View on GitHub↗0
  • deviantony/docker-elkdeviantony avatar

    deviantony/docker-elk

    18,375View on GitHub↗

    This project is a containerized orchestration layer for the Elastic Stack, providing a pre-configured set of Docker Compose files to deploy Elasticsearch, Logstash, and Kibana as a unified data analysis stack. It functions as a centralized log management system for ingesting, indexing, and searching log data using a cluster of interconnected services. The deployment pattern includes an Elasticsearch cluster manager that enables scaling data nodes through replica scaling and internal discovery. It provides a web-based administration interface for monitoring cluster health and status. The syst

    Shelldockerdocker-composeelasticsearch
    View on GitHub↗18,375
  • cloud-foundations/tricorderCloud-Foundations avatar

    Cloud-Foundations/tricorder

    1View on GitHub↗
    Go
    View on GitHub↗1