awesome-repositories.com
المدونة
MCP
awesome-repositories.com

اكتشف أفضل مستودعات المصادر المفتوحة باستخدام بحث مدعوم بالذكاء الاصطناعي.

استكشفعمليات بحث منسقةبدائل مفتوحة المصدربرمجيات ذاتية الاستضافةالمدونةخريطة الموقع
المشروعخادم MCPحولكيفية ترتيب النتائجالصحافة
قانونيالخصوصيةالشروط
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
S

streamsets/datacollector

0
View on GitHub↗
0 نجوم·0 تفرعات·9 مشاهدات

Datacollector

Features

  • Data Ingestion - Infrastructure for continuous big data ingestion.
  • Data Ingestion Pipelines - Infrastructure for continuous big data ingestion.

سجل النجوم

مخطط تاريخ النجوم لـ streamsets/datacollectorمخطط تاريخ النجوم لـ streamsets/datacollector

بحث بالذكاء الاصطناعي

استكشف المزيد من المستودعات الرائعة

صف ما تحتاجه بلغة بسيطة — وسيقوم الذكاء الاصطناعي بترتيب آلاف المشاريع مفتوحة المصدر المنسقة حسب الصلة.

Start searching with AI

بدائل مفتوحة المصدر لـ Datacollector

مشاريع مفتوحة المصدر مشابهة، مرتبة حسب عدد الميزات المشتركة مع Datacollector.
  • apache/pulsarالصورة الرمزية لـ apache

    apache/pulsar

    15,276عرض على GitHub↗

    Apache Pulsar is a cloud-native distributed pub-sub messaging system designed for high-performance data ingestion. It functions as a geo-replicated data streamer and a multi-tenant event streaming platform, providing a serverless stream processing engine and a tiered storage messaging broker. The system distinguishes itself by separating serving layers from storage layers to allow independent scaling of compute and data retention. It features native geo-replication to synchronize messages across different geographical regions and employs a multi-layered tenant isolation model using authentica

    Java
    عرض على GitHub↗15,276
  • bruin-data/bruinالصورة الرمزية لـ bruin-data

    bruin-data/bruin

    1,620عرض على GitHub↗

    Build data pipelines with SQL and Python, ingest data from different sources, add quality checks, and build end-to-end flows.

    Goanalyticsbigquerydata-analysis
    عرض على GitHub↗1,620
  • aklivity/zillaالصورة الرمزية لـ aklivity

    aklivity/zilla

    690عرض على GitHub↗

    🦎 A multi-protocol edge & service proxy. Seamlessly interface web apps, IoT clients, & microservices to Apache Kafka® via declaratively defined, stateless APIs.

    Java
    عرض على GitHub↗690
  • bruin-data/ingestrالصورة الرمزية لـ bruin-data

    bruin-data/ingestr

    3,714عرض على GitHub↗

    ingestr is a command-line tool for copying and syncing data between different database engines and third-party platforms without writing custom code. It functions as an ETL pipeline utility that extracts data from diverse sources and loads it into destinations. The tool features a schema-agnostic data loader that maps source fields to destination columns dynamically, removing the need for predefined static table definitions. It also operates as an incremental data synchronizer, updating destination tables by appending new records or merging changes to maintain current datasets. The system pr

    Go
    عرض على GitHub↗3,714
عرض جميع البدائل الـ 30 لـ Datacollector→

الأسئلة الشائعة

ما هي الميزات الرئيسية لـ streamsets/datacollector؟

الميزات الرئيسية لـ streamsets/datacollector هي: Data Ingestion, Data Ingestion Pipelines.

ما هي البدائل مفتوحة المصدر لـ streamsets/datacollector؟

تشمل البدائل مفتوحة المصدر لـ streamsets/datacollector: bruin-data/bruin — Build data pipelines with SQL and Python, ingest data from different sources, add quality checks, and build end-to-end… facebookarchive/scribe — Scribe is a distributed log aggregation system designed to collect and route real-time log data from numerous servers… aklivity/zilla — 🦎 A multi-protocol edge & service proxy. Seamlessly interface web apps, IoT clients, & microservices to Apache Kafka®… apache/pulsar — Apache Pulsar is a cloud-native distributed pub-sub messaging system designed for high-performance data ingestion. It… bruin-data/ingestr — ingestr is a command-line tool for copying and syncing data between different database engines and third-party… gazette/core — Build platforms that flexibly mix SQL, batch, and stream processing paradigms.