awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
onyx-platform avatar

onyx-platform/onyxArchived

0
View on GitHub↗
2,050 stars·200 forks·Clojure·EPL-1.0·7 viewswww.onyxplatform.org↗

Onyx

Distributed, masterless, high performance, fault tolerant data processing

Features

  • Science and Data Analysis - Distributed data processing platform.
  • Data Processing and Analysis - Distributed, fault-tolerant data processing platform for Clojure.
  • Streaming Engines - Distributed, masterless, high-performance data processing.

Star history

Star history chart for onyx-platform/onyxStar history chart for onyx-platform/onyx

How this analysis was created: This summary and feature list were written by an AI model that read the project's README and public documentation pages. Each feature links to the documentation it came from; stars, license and language come straight from the GitHub API. The model does not read the source code, and the analysis is refreshed when the project is re-analysed. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Open-source alternatives to Onyx

Similar open-source projects, ranked by how many features they share with Onyx.
  • apache/sparkapache avatar

    apache/spark

    43,467View on GitHub↗

    Apache Spark is a unified distributed data processing engine designed for large-scale data analysis and computation graphs. It functions as a distributed machine learning framework, a graph processing system, a real-time stream processor, and a SQL analytics engine. The system enables the execution of distributed SQL querying, large-scale graph analysis, and real-time stream analytics across clusters of machines. It also provides a scalable environment for implementing machine learning algorithms and predictive model development on massive datasets. The engine incorporates relational query e

    Scalabig-datajavajdbc
    View on GitHub↗43,467
  • microsoft/c9-python-getting-startedmicrosoft avatar

    microsoft/c9-python-getting-started

    8,012View on GitHub↗

    This project is a Python education repository and programming tutorial designed to teach language fundamentals, from basic syntax and variables to advanced concepts. It serves as a data science starter kit and a guide for REST API integration. The repository provides instructional scripts and sample code covering object-oriented programming patterns and asynchronous programming. It includes practical demonstrations for fetching and processing JSON data from external web services using HTTP requests. The materials cover a broad capability surface including data analysis workflows with interac

    Jupyter Notebook
    View on GitHub↗8,012
  • e2b-dev/code-interpretere2b-dev avatar

    e2b-dev/code-interpreter

    2,348View on GitHub↗

    This project is an infrastructure platform designed to provide secure, isolated, and ephemeral cloud-based Linux environments for AI agents and automated code execution. It functions as an orchestrator that provisions on-demand virtual machines, allowing developers to run arbitrary code generated by large language models within hardware-level security boundaries. The platform distinguishes itself through its ability to manage stateful, long-lived sessions that persist across multiple execution calls, enabling complex, multi-step workflows. It supports high-concurrency scaling, allowing for th

    Pythonaiai-data-analysisanthropic
    View on GitHub↗2,348
  • alteryx/featuretoolsalteryx avatar

    alteryx/featuretools

    7,658View on GitHub↗

    Featuretools is an automated feature engineering library and data transformation framework written in Python. It automatically generates machine learning feature vectors from multi-table datasets by applying synthesis patterns to relational and timestamped data. The system functions as a distributed feature synthesis engine, allowing the process of creating feature vectors to scale across multiple cores or clusters to handle large-scale datasets. The library supports the synthesis of multi-table datasets, time series feature generation, and the creation of custom machine learning primitives

    Python
    View on GitHub↗7,658
See all 30 alternatives to Onyx→

Frequently asked questions

What does onyx-platform/onyx do?

Distributed, masterless, high performance, fault tolerant data processing

What are the main features of onyx-platform/onyx?

The main features of onyx-platform/onyx are: Science and Data Analysis, Data Processing and Analysis, Streaming Engines.

What are some open-source alternatives to onyx-platform/onyx?

Open-source alternatives to onyx-platform/onyx include: apache/spark — Apache Spark is a unified distributed data processing engine designed for large-scale data analysis and computation… microsoft/c9-python-getting-started — This project is a Python education repository and programming tutorial designed to teach language fundamentals, from… e2b-dev/code-interpreter — This project is an infrastructure platform designed to provide secure, isolated, and ephemeral cloud-based Linux… alteryx/featuretools — Featuretools is an automated feature engineering library and data transformation framework written in Python. It… apache/arrow-ballista — Apache DataFusion Ballista Distributed Query Engine. apache/apex-core — Mirror of Apache Apex core.