awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
Netflix avatar

Netflix/suroArchived

0
View on GitHub↗
796 stars·170 forks·Java·Apache-2.0·22 views

Suro

Netflix's distributed Data Pipeline

Features

  • Data Ingestion - Log aggregation system based on Chukwa.
  • Data Ingestion and Integration - Distributed data pipeline for ingestion.
  • Data Ingestion Pipelines - Log aggregator based on Chukwa architecture.
  • Data Pipelines - Data pipeline service for application event collection.
  • Data Processing and Analytics - Data pipeline service for collecting and dispatching events.

Star history

Star history chart for netflix/suroStar history chart for netflix/suro

How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Frequently asked questions

What does netflix/suro do?

Netflix's distributed Data Pipeline

What are the main features of netflix/suro?

The main features of netflix/suro are: Data Ingestion, Data Ingestion and Integration, Data Ingestion Pipelines, Data Pipelines, Data Processing and Analytics.

Which projects share features with netflix/suro?

Projects with overlapping indexed features include: gazette/core — Build platforms that flexibly mix SQL, batch, and stream processing paradigms. rudderlabs/rudder-server — Rudder Server is a customer data platform and event routing pipeline designed to collect, transform, and route… bruin-data/ingestr — ingestr is a command-line tool for copying and syncing data between different database engines and third-party… bruin-data/bruin — Build data pipelines with SQL and Python, ingest data from different sources, add quality checks, and build end-to-end… linkedin/gobblin — A distributed data integration framework that simplifies common aspects of big data integration such as data… aklivity/zilla — 🦎 A multi-protocol edge & service proxy. Seamlessly interface web apps, IoT clients, & microservices to Apache Kafka®…

Projects sharing features with Suro

These projects share indexed features with Suro. Shared tags can include platform or build tooling; verify the primary use case before treating a result as a replacement.
  • gazette/coregazette avatar

    gazette/core

    793View on GitHub↗

    Build platforms that flexibly mix SQL, batch, and stream processing paradigms

    Go
    View on GitHub↗793
  • bruin-data/bruinbruin-data avatar

    bruin-data/bruin

    1,620View on GitHub↗

    Build data pipelines with SQL and Python, ingest data from different sources, add quality checks, and build end-to-end flows.

    Goanalyticsbigquerydata-analysis
    View on GitHub↗1,620
  • bruin-data/ingestrbruin-data avatar

    bruin-data/ingestr

    3,714View on GitHub↗

    ingestr is a command-line tool for copying and syncing data between different database engines and third-party platforms without writing custom code. It functions as an ETL pipeline utility that extracts data from diverse sources and loads it into destinations. The tool features a schema-agnostic data loader that maps source fields to destination columns dynamically, removing the need for predefined static table definitions. It also operates as an incremental data synchronizer, updating destination tables by appending new records or merging changes to maintain current datasets. The system pr

    Go
    View on GitHub↗3,714
  • linkedin/gobblinlinkedin avatar

    linkedin/gobblin

    2,267View on GitHub↗

    A distributed data integration framework that simplifies common aspects of big data integration such as data ingestion, replication, organization and lifecycle management for both streaming and batch data ecosystems.

    Java
    View on GitHub↗2,267
Compare all 30 related projects→