2 Repos
Specialized interfaces for moving data between databases and distributed processing frameworks.
Distinct from Big Data Processing: Focuses on the movement and integration pipeline rather than the processing logic itself
Explore 2 awesome GitHub repositories matching data & databases · Data Pipeline Connectors. Refine with filters or upvote what's useful.
Nebula is a distributed graph database designed for storing and querying massive volumes of interconnected vertices and edges across a horizontally scalable cluster. It functions as a Kubernetes-native database and a distributed graph analytics engine, utilizing a Raft-based distributed store to ensure strong consistency and high availability. The system features an OpenCypher query engine for performing complex graph traversals and pattern matching. It distinguishes itself with a decoupled compute-storage architecture and a shared-nothing distributed design, allowing query processing and dat
Enables high-volume data exchange between the graph database and distributed frameworks like Apache Spark and Flink.
Hazelcast is a distributed data platform that combines an in-memory data grid with a stream processing engine to support real-time analytics and event-driven applications. It functions as a partitioned, distributed key-value store that replicates data across cluster nodes to provide low-latency access and high availability. The platform also serves as a distributed SQL query engine, allowing users to execute standard SQL statements against both in-memory datasets and external data sources. What distinguishes Hazelcast is its use of a distributed consensus subsystem to maintain strongly consis
Moves data between internal pipelines and external systems like messaging queues and databases via built-in connectors.