awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
DataKitchen avatar

DataKitchen/data-observability-installer

0
View on GitHub↗
138 stars·12 forks·Python·Apache-2.0·7 viewsdatakitchen.io↗

Data Observability Installer

Data breaks. Servers break. Your toolchain breaks. Ensure your data team is the first to know and the first to solve with visibility across and down your data estate. Save time with simple, fast data quality test generation and execution. Trust your data, tools, and systems from end to end.

Features

  • Data Quality - Observability and alerting across data stacks.

Star history

Star history chart for datakitchen/data-observability-installerStar history chart for datakitchen/data-observability-installer

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Open-source alternatives to Data Observability Installer

Similar open-source projects, ranked by how many features they share with Data Observability Installer.
  • unionai-oss/panderaunionai-oss avatar

    unionai-oss/pandera

    4,382View on GitHub↗

    Pandera is a data pipeline validation framework and statistical type validation tool. It functions as a library for defining and enforcing schemas on datasets to ensure data quality and consistency, specifically providing validation capabilities for Pandas dataframes. The project includes a schema inference tool that automates setup by analyzing existing dataset samples to generate validation schemas. It also serves as a synthetic data generator, creating artificial datasets based on predefined schemas to verify data-producing functions. The framework covers data engineering quality assuranc

    Pythonassertionsdata-assertionsdata-check
    View on GitHub↗4,382
  • ydataai/pandas-profilingydataai avatar

    ydataai/pandas-profiling

    13,610View on GitHub↗

    This project is an exploratory data analysis framework and profiling tool designed to generate comprehensive statistical reports from Pandas and Spark DataFrames. It functions as a data quality profiler that identifies missing values, duplicates, and high correlations within tabular datasets. The tool distinguishes itself through specialized capabilities for time-series analysis, extracting temporal statistics, seasonality, and auto-correlation plots. It also includes a dataset comparison utility to identify structural or content changes between different versions of a dataset. The analysis

    Python
    View on GitHub↗13,610
  • pimcore/pimcorepimcore avatar

    pimcore/pimcore

    3,784View on GitHub↗

    Pimcore is an open-source data experience platform that serves as a unified framework for managing product information, digital assets, and customer data. It functions as an enterprise content management system and a master data management platform, providing a centralized source of truth for complex business information. The system is designed to support omnichannel delivery, enabling organizations to publish content and manage digital experiences across diverse platforms through both traditional and headless architectures. The platform distinguishes itself through a metadata-driven object m

    PHP
    View on GitHub↗3,784
  • turboway/bigdata_analyseTurboWay avatar

    TurboWay/bigdata_analyse

    5,238View on GitHub↗

    This project is a collection of big data frameworks and pipelines, including an Apache Hive analysis framework, a behavioral data analytics platform, a predictive analytics engine, and real-time data pipelines. It provides the infrastructure for building Extract, Transform, Load (ETL) workflows to process large datasets for distributed storage and SQL-based analysis. The system supports diverse analytical implementations, such as a predictive engine using linear regression for value forecasting and a real-time architecture that moves data through message brokers for immediate reporting. It in

    Pythonhqlpythonsql
    View on GitHub↗5,238
See all 8 alternatives to Data Observability Installer→

Frequently asked questions

What does datakitchen/data-observability-installer do?

Data breaks. Servers break. Your toolchain breaks. Ensure your data team is the first to know and the first to solve with visibility across and down your data estate. Save time with simple, fast data quality test generation and execution. Trust your data, tools, and systems from end to end.

What are the main features of datakitchen/data-observability-installer?

The main features of datakitchen/data-observability-installer are: Data Quality.

What are some open-source alternatives to datakitchen/data-observability-installer?

Open-source alternatives to datakitchen/data-observability-installer include: unionai-oss/pandera — Pandera is a data pipeline validation framework and statistical type validation tool. It functions as a library for… ydataai/pandas-profiling — This project is an exploratory data analysis framework and profiling tool designed to generate comprehensive… pimcore/pimcore — Pimcore is an open-source data experience platform that serves as a unified framework for managing product… turboway/bigdata_analyse — This project is a collection of big data frameworks and pipelines, including an Apache Hive analysis framework, a… rbmuller/scherlok — A detective for your data. Zero-config data quality monitoring — works with dbt, Postgres, BigQuery, Snowflake. No YAML. elementary-data/elementary — Elementary OSS: dbt-native data observability.