awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
bruin-data avatar

bruin-data/bruin

0
View on GitHub↗
1,620 stars·81 forks·Go·Apache-2.0·9 viewsgetbruin.com/docs/bruin↗

Bruin

Build data pipelines with SQL and Python, ingest data from different sources, add quality checks, and build end-to-end flows.

Features

  • Data Ingestion - End-to-end pipeline tool for ingestion, transformation, and quality.
  • Data Ingestion Pipelines - End-to-end pipeline tool for ingestion and quality checks.
  • Data Integration Providers - Layer for data ingestion and transformation across multiple sources.
  • Data Orchestration - Data pipeline framework supporting SQL and Python DAGs.
  • Data Pipelines - End-to-end pipeline tool for ingestion and transformation.
  • Workflow Orchestration - CLI tool for end-to-end pipeline management and data quality.
  • CI/CD and Orchestration - Run and schedule SQL transformations without Airflow.

Star history

Star history chart for bruin-data/bruinStar history chart for bruin-data/bruin

How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Projects sharing features with Bruin

These projects share indexed features with Bruin. Shared tags can include platform or build tooling; verify the primary use case before treating a result as a replacement.
  • rudderlabs/rudder-serverrudderlabs avatar

    rudderlabs/rudder-server

    4,437View on GitHub↗

    Rudder Server is a customer data platform and event routing pipeline designed to collect, transform, and route customer event data from various sources to data warehouses and business tools. It functions as a customer identity resolver, linking identifiers from multiple sources to build a unified identity graph and comprehensive behavioral customer profiles. The system differentiates itself through reverse ETL capabilities, which push processed customer segments and audiences from data warehouses back into operational third-party applications. It also provides a containerized data plane for K

    Gobigquerycdpcustomer-data
    View on GitHub↗4,437
  • bruin-data/ingestrbruin-data avatar

    bruin-data/ingestr

    3,714View on GitHub↗

    ingestr is a command-line tool for copying and syncing data between different database engines and third-party platforms without writing custom code. It functions as an ETL pipeline utility that extracts data from diverse sources and loads it into destinations. The tool features a schema-agnostic data loader that maps source fields to destination columns dynamically, removing the need for predefined static table definitions. It also operates as an incremental data synchronizer, updating destination tables by appending new records or merging changes to maintain current datasets. The system pr

    Go
    View on GitHub↗3,714
  • netflix/suroNetflix avatar

    Netflix/suro

    796View on GitHub↗

    Netflix's distributed Data Pipeline

    Java
    View on GitHub↗796
  • gazette/coregazette avatar

    gazette/core

    793View on GitHub↗

    Build platforms that flexibly mix SQL, batch, and stream processing paradigms

    Go
    View on GitHub↗793
Compare all 30 related projects→

Frequently asked questions

What does bruin-data/bruin do?

Build data pipelines with SQL and Python, ingest data from different sources, add quality checks, and build end-to-end flows.

What are the main features of bruin-data/bruin?

The main features of bruin-data/bruin are: Data Ingestion, Data Ingestion Pipelines, Data Integration Providers, Data Orchestration, Data Pipelines, Workflow Orchestration, CI/CD and Orchestration.

Which projects share features with bruin-data/bruin?

Projects with overlapping indexed features include: rudderlabs/rudder-server — Rudder Server is a customer data platform and event routing pipeline designed to collect, transform, and route… bruin-data/ingestr — ingestr is a command-line tool for copying and syncing data between different database engines and third-party… gazette/core — Build platforms that flexibly mix SQL, batch, and stream processing paradigms. netflix/suro — Netflix's distributed Data Pipeline. apache/airflow — Airflow is a platform for programmatically authoring, scheduling, and monitoring complex data pipelines. It functions… apache/pulsar — Apache Pulsar is a cloud-native distributed pub-sub messaging system designed for high-performance data ingestion. It…