awesome-repositories.com
博客
awesome-repositories.com

通过 AI 驱动的搜索,发现最优秀的开源仓库。

探索精选搜索开源替代品自托管软件博客网站地图
项目关于排名机制媒体报道MCP 服务器
法律隐私政策服务条款
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
SPLWare avatar

SPLWare/esProc

0
View on GitHub↗
4,685 星标·363 分支·Java·Apache-2.0·3 次浏览doc.esproc.com/esproc↗

EsProc

esProc 是一个分布式 ETL 框架和嵌入式数据计算引擎。它为 Java 虚拟机提供了一种结构化数据语言,专为关系查询、复杂数据计算和结构化数据分析而设计。

该系统具有自然语言数据查询界面,利用大语言模型将请求转换为针对结构化数据集的可执行查询。它采用具有简洁语法的领域特定查询语言来建立表关系并检索信息。

该平台涵盖跨不同关系型和 NoSQL 源的数据集成,并管理 ETL 工作流以在文件和数据库之间移动数据。其他功能包括结构化数据报告生成、用于分步执行可视化的实时网格界面,以及集成自定义外部共享库的能力。

Features

  • ETL Workflows - Provides a distributed framework for extracting, transforming, and loading data across cluster servers and disparate sources.
  • Cross-Source Data Integration - Enables joining and merging datasets from diverse external relational and NoSQL sources into a single result set.
  • Hybrid Data Integration - Orchestrates data movement and transformation across disparate relational and NoSQL sources without a centralized warehouse.
  • Embedded Analytics Engines - Provides a lightweight, embeddable database component for native query and analysis capabilities within applications.
  • Large-Scale Data Computation - Provides a JVM-based environment for executing complex data analysis and computation graphs across distributed clusters.
  • Query Domain Specific Languages - Employs a domain-specific query language with concise syntax to define table relationships and retrieve structured information.
  • Embedded SQL Query Engines - Integrates a JVM-based engine that executes queries directly against connected data sources.
  • JVM-Based Runtime Executions - Executes data computations and scripts within a JVM to ensure cross-platform compatibility and memory efficiency.
  • Structured Data Languages - Implements a specialized structured data language for the JVM designed for complex relational queries and analysis.
  • Core Engine Embedding - Implements an architectural pattern for integrating the core data engine into other software environments for localized processing.
  • Distributed Cluster Coordination - Provides mechanisms for synchronizing state and scheduling across multiple compute nodes to handle large-scale processing.
  • Natural Language Command Translation - Integrates large language models to translate natural language requests into executable data query commands.
  • Natural Language Querying - Allows users to retrieve and analyze structured datasets using natural language queries powered by large language models.
  • Natural Language Querying Interfaces - Includes an interface that translates natural language prompts into structured queries for data retrieval.
  • Complex Data Structure Transformation - Performs complex data analysis and reshaping using set operations and relational queries.
  • Execution Visualizers - Provides a real-time grid interface for step-by-step execution visualization to debug and verify data flows.
  • Execution Visualizers - Ships a real-time grid interface to visualize and debug data computation steps.

Star 历史

splware/esproc 的 Star 历史图表splware/esproc 的 Star 历史图表

AI 搜索

探索更多 awesome 仓库

用简单的语言描述您的需求 —— AI 将根据相关性为您从数千个精选开源项目中进行排序。

Start searching with AI

EsProc 的开源替代方案

相似的开源项目,按与 EsProc 的功能重合度排序。
  • pentaho/pentaho-kettlepentaho 的头像

    pentaho/pentaho-kettle

    8,353在 GitHub 上查看↗

    Pentaho Kettle is an enterprise ETL data integration platform designed to extract, transform, and load data between disparate sources and target databases. It functions as a metadata-driven orchestrator that utilizes a visual workflow designer to create and manage complex sequences of data tasks and transformation pipelines. The system is distinguished by its distributed data processing engine, which executes workloads across clusters of server nodes to increase throughput. It employs a plugin-based architecture, allowing the platform to be extended via external JAR files to provide connectiv

    Java
    在 GitHub 上查看↗8,353
  • alibaba/otteralibaba 的头像

    alibaba/otter

    8,127在 GitHub 上查看↗

    Otter is a distributed database synchronization system and change data capture tool designed to replicate data between databases across multiple geographic regions. It functions as a synchronization orchestrator and ETL data pipeline that mirrors records and associated files in real time. The system employs incremental log parsing to capture database changes and utilizes a consistency-based convergence algorithm and loop-avoidance logic to manage bi-directional replication. It processes data through a pipeline of selection, extraction, transformation, and loading to handle joins and format co

    Java
    在 GitHub 上查看↗8,127
  • clickhouse/clickhouseClickHouse 的头像

    ClickHouse/ClickHouse

    48,229在 GitHub 上查看↗

    ClickHouse is a high-performance, columnar analytical database designed for real-time query execution and large-scale data aggregation. It functions as a distributed data warehouse capable of processing petabytes of information, while also providing an embedded engine that integrates directly into applications for native query capabilities without external dependencies. The system is built to handle high-throughput ingestion and complex analytical workloads, delivering millisecond-level latency for interactive dashboards and operational monitoring. The platform distinguishes itself through ad

    C++aianalyticsbig-data
    在 GitHub 上查看↗48,229
  • evidence-dev/evidenceevidence-dev 的头像

    evidence-dev/evidence

    5,919在 GitHub 上查看↗
    JavaScriptanalyticsbusiness-intelligencedashboard
    在 GitHub 上查看↗5,919
查看 EsProc 的所有 30 个替代方案→

常见问题解答

splware/esproc 是做什么的?

esProc 是一个分布式 ETL 框架和嵌入式数据计算引擎。它为 Java 虚拟机提供了一种结构化数据语言,专为关系查询、复杂数据计算和结构化数据分析而设计。

splware/esproc 的主要功能有哪些?

splware/esproc 的主要功能包括:ETL Workflows, Cross-Source Data Integration, Hybrid Data Integration, Embedded Analytics Engines, Large-Scale Data Computation, Query Domain Specific Languages, Embedded SQL Query Engines, JVM-Based Runtime Executions。

splware/esproc 有哪些开源替代品?

splware/esproc 的开源替代品包括: pentaho/pentaho-kettle — Pentaho Kettle is an enterprise ETL data integration platform designed to extract, transform, and load data between… alibaba/otter — Otter is a distributed database synchronization system and change data capture tool designed to replicate data between… clickhouse/clickhouse — ClickHouse is a high-performance, columnar analytical database designed for real-time query execution and large-scale… evidence-dev/evidence. bruin-data/ingestr — ingestr is a command-line tool for copying and syncing data between different database engines and third-party… apache/datafusion — Apache DataFusion is an extensible, columnar SQL query engine that runs embedded within a host application without…