11 مستودعات
Libraries and APIs for defining and executing distributed data workflows.
Explore 11 awesome GitHub repositories matching part of an awesome list · Distributed Programming. Refine with filters or upvote what's useful.
Ray is a distributed computing framework designed to scale Python and Java applications across clusters by abstracting task scheduling and resource management. It functions as a resource-aware execution engine that manages task dependencies, placement, and fault tolerance across networked compute nodes. At its core, the system provides a stateful actor model, allowing developers to define classes that run in dedicated processes to maintain and mutate internal state across remote method calls. The framework distinguishes itself through a robust cross-language interoperability layer, enabling f
Fast framework for building and running distributed applications.
Apache Heron (Incubating) is a realtime, distributed, fault-tolerant stream processing engine from Twitter
Real-time, distributed, fault-tolerant stream processing engine.
A Scala API for Cascading
Scala library for MapReduce jobs built on Cascading.
Streaming MapReduce with Scalding and Storm
Streaming MapReduce library integrating Scalding and Storm.
Map-Reduce for Clojure
MapReduce implementation for Clojure compiling to Apache Pig.
High performance distributed data processing engine
High-performance distributed data processing for Node.js.
Hadoop MapReduce in idiomatic Clojure.
MapReduce library for the Clojure language.
Big Data Science Swiss Army Knife - http://www.tuktu.io --
Platform for batch and streaming computation using Akka.
Tuple MapReduce for Hadoop: Hadoop API made easy
Alternative paradigm for implementing MapReduce jobs.
Develop streaming applications for IBM Streams in Python, Java & Scala.
Libraries for building streaming applications in Java, Python, or Scala.