awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
PaddlePaddle avatar

PaddlePaddle/ERNIE

0
View on GitHub↗
7,717 stars·1,441 forks·Python·Apache-2.0·21 viewsernie.baidu.com↗

ERNIE

ERNIE is a development toolkit for training, fine-tuning, and deploying large language models built on the PaddlePaddle deep learning platform. It provides a comprehensive suite of core components, including an inference server for vision and language models, a training and fine-tuning toolkit, and a framework for building retrieval-augmented generation systems using private knowledge bases.

The project features multimodal AI models capable of reasoning across text, images, and video to perform complex visual understanding and information extraction. It distinguishes itself through specialized training methodologies for function calling and the use of mixture-of-experts architectures to enhance cross-modal reasoning.

The system covers a broad range of capabilities including industrial natural language processing deployment, visual mathematical reasoning, and document information extraction. Performance is addressed through quantization, hybrid-parallelism training, and disaggregated inference serving to optimize memory usage and throughput.

A web-based user interface is provided for supervising training processes and conducting interactive conversations.

Features

  • Model Inference Servers - Ships a production-ready server application that hosts vision and language models with quantization and hardware acceleration.
  • Large Language Model Training Frameworks - Provides a development toolkit for training, fine-tuning, and deploying large language models built on PaddlePaddle.
  • Distributed Training - Provides tools for configuring data and model parallelism to train large neural networks across multiple devices.
  • Inference Acceleration - Reduces memory footprints and latency by applying quantization, multi-expert collaboration, and disaggregation techniques.
  • Knowledge Retrieval Systems - Provides a framework for building question-answering systems that surface information from private, domain-specific datasets.
  • LLM Development Toolkits - Provides a development toolkit for training, fine-tuning, and deploying large language models built on the PaddlePaddle platform.
  • LLM Fine-Tuning Toolsets - Provides a comprehensive toolkit for supervised fine-tuning and parameter-efficient updates like LoRA for language models.
  • Multimodal Perception Models - Ships models designed to interpret and analyze visual data, charts, and cross-modal inputs alongside text.
  • Multimodal Processing - Reasons across text, images, and video using heterogeneous structures to perform complex visual understanding tasks.
  • Multimodal Training - Processes textual and visual data simultaneously using mixture-of-experts to improve cross-modal reasoning and generation.
  • Industrial NLP Pipelines - Provides a specialized set of tools to build and deploy industrial-grade natural language processing applications.
  • Parameter Efficient Fine-Tuning - Provides memory-efficient adaptation techniques like LoRA to update a small subset of model parameters.
  • Visual Content Analysis - Extracts visual knowledge from images, documents, and charts to maintain high perception accuracy across datasets.
  • Large Language Model Deployments - Executes resource-efficient training and inference workflows for large-scale models on private industrial hardware.
  • Multimodal AI - Builds models that bridge text, images, and video to perform complex visual understanding and reasoning.
  • Multimodal Models - Implements a model architecture capable of reasoning across text, images, and video for visual information extraction.
  • Inference Deployment - Provides infrastructure and tools for hosting and serving language and vision models on multi-hardware setups.
  • Chat Bot Frameworks - Provides tools for creating automated agents that interact with users through chat, web search, and function calling.
  • Visual Mathematical Reasoning - Solves complex multimodal reasoning puzzles and mathematical visual tasks by combining visual perception with a thinking mode.
  • Conversational Bot Development - Enables the creation of interactive chat interfaces and bots that integrate real-time web search for dynamic information delivery.
  • RAG Document Retrieval - Retrieves relevant snippets from local and private documents to provide grounded context for model responses.
  • Function Calling Interfaces - Trains models to recognize and execute external tool calls through specialized function call training methodologies.
  • Information Extraction - Extracts key data and performs deep information extraction from contracts using text recognition and language modeling.
  • Instruction-Following Models - Processes detailed user prompts and world knowledge to generate precise responses based on specific constraints.
  • Mixture of Experts - Supports routing and recording expert paths using a mixture-of-experts architecture to improve reasoning efficiency.
  • Preference Alignment - Optimizes task accuracy and aligns model outputs with human preferences using supervised fine-tuning methods.
  • Multimodal Fine-Tuning - Enables customizing pre-trained models for general language or visual tasks using supervised fine-tuning and preference optimization.
  • Model Deployment - Transforms trained language models into production-ready services for real-world industrial environments.
  • Model Management Interfaces - Ships a graphical user interface for supervising model training processes and conducting interactive conversations.
  • Disaggregated Inference - Implements architectures that separate prefill and decode stages across distinct hardware nodes to reduce latency.
  • Model Performance Optimization - Enhances model speed and accuracy through quantization and hardware acceleration within the PaddlePaddle framework.
  • Quantization-Aware Training - Integrates low-precision arithmetic into the training loop to reduce model size while maintaining high accuracy.
  • Reasoning Models - Executes mathematical and knowledge-intensive tasks using specialized model architectures to achieve high-accuracy results.
  • Training Memory Management - Eliminates padding and reduces GPU memory consumption by packing multiple data samples into a single sequence.
  • Training Throughput Optimization - Increases pre-training speed by using hybrid parallelism, mixed-precision formats, and hierarchical load balancing.
  • Language Models - Enhanced representation through knowledge integration for Chinese tasks.
  • Natural Language Processing - Enhances semantic representation through knowledge-integrated pre-training.
  • Pretrained Models and Embeddings - Chinese-specific language representation models.

Star history

Star history chart for paddlepaddle/ernieStar history chart for paddlepaddle/ernie

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Open-source alternatives to ERNIE

Similar open-source projects, ranked by how many features they share with ERNIE.
  • nvidia/isaac-gr00tNVIDIA avatar

    NVIDIA/Isaac-GR00T

    6,222View on GitHub↗
    Jupyter Notebook
    View on GitHub↗6,222
  • paddlepaddle/larkPaddlePaddle avatar

    PaddlePaddle/LARK

    7,717View on GitHub↗

    LARK is a development toolkit for training, fine-tuning, and deploying large language models and multimodal models based on PaddlePaddle. It functions as a comprehensive framework that includes an LLM training orchestrator, an inference server, and a multimodal model framework for processing text, image, and video inputs. The project features a retrieval-augmented generation system for building conversational applications that integrate web search and private knowledge bases. It provides specific capabilities for multimodal reasoning and complex logic, enabling the extraction of structured da

    Python
    View on GitHub↗7,717
  • dusty-nv/jetson-inferencedusty-nv avatar

    dusty-nv/jetson-inference

    8,734View on GitHub↗

    jetson-inference is a set of libraries and tools for executing optimized deep learning models on embedded GPU hardware. Its primary purpose is to enable real-time computer vision and AI inference at the edge with low latency and high throughput. The project distinguishes itself through high-performance streaming analytics and the ability to execute concurrent AI pipelines on auto-grade silicon. It provides specialized support for multi-sensor stream processing, utilizing zero-copy data transport to load camera frames directly into GPU memory. The codebase covers a broad surface of capabiliti

    C++caffecomputer-visiondeep-learning
    View on GitHub↗8,734
  • sgl-project/sglangsgl-project avatar

    sgl-project/sglang

    29,079View on GitHub↗

    Sglang is a high-performance inference engine and serving system designed for large language and multimodal models. It provides a programmable interface for orchestrating complex generation workflows, enabling developers to coordinate multi-turn dialogues, tool invocations, and reasoning chains through a domain-specific language. The platform is built to support production-scale deployments, offering an OpenAI-compatible API that allows for integration with existing application ecosystems. The system distinguishes itself through a disaggregated architecture that separates compute-intensive pr

    Pythonattentionblackwellcuda
    View on GitHub↗29,079
See all 30 alternatives to ERNIE→

Frequently asked questions

What does paddlepaddle/ernie do?

ERNIE is a development toolkit for training, fine-tuning, and deploying large language models built on the PaddlePaddle deep learning platform. It provides a comprehensive suite of core components, including an inference server for vision and language models, a training and fine-tuning toolkit, and a framework for building retrieval-augmented generation systems using private knowledge bases.

What are the main features of paddlepaddle/ernie?

The main features of paddlepaddle/ernie are: Model Inference Servers, Large Language Model Training Frameworks, Distributed Training, Inference Acceleration, Knowledge Retrieval Systems, LLM Development Toolkits, LLM Fine-Tuning Toolsets, Multimodal Perception Models.

What are some open-source alternatives to paddlepaddle/ernie?

Open-source alternatives to paddlepaddle/ernie include: nvidia/isaac-gr00t. paddlepaddle/lark — LARK is a development toolkit for training, fine-tuning, and deploying large language models and multimodal models… dusty-nv/jetson-inference — jetson-inference is a set of libraries and tools for executing optimized deep learning models on embedded GPU… sgl-project/sglang — Sglang is a high-performance inference engine and serving system designed for large language and multimodal models. It… zhaochenyang20/awesome-ml-sys-tutorial — This project provides a comprehensive technical guide and framework for engineering large-scale machine learning… internlm/xtuner — xtuner is a comprehensive training engine for large language models, offering a toolkit for pre-training, supervised…