awesome-repositories.com
المدونة
MCP
awesome-repositories.com

اكتشف أفضل مستودعات المصادر المفتوحة باستخدام بحث مدعوم بالذكاء الاصطناعي.

استكشفعمليات بحث منسقةبدائل مفتوحة المصدربرمجيات ذاتية الاستضافةالمدونةخريطة الموقع
المشروعخادم MCPحولكيفية ترتيب النتائجالصحافة
قانونيالخصوصيةالشروط
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
Back to wenge-research/yayi

Open-source alternatives to YAYI

30 open-source projects similar to wenge-research/yayi, ranked by how many features they have in common. Compare stars, activity and what each one does to find the best YAYI alternative.

  • thudm/chatglm2-6bالصورة الرمزية لـ THUDM

    THUDM/ChatGLM2-6B

    15,565عرض على GitHub↗

    ChatGLM2-6B is an open-weight large language model designed for natural language conversations and text generation in both English and Chinese. It functions as a bilingual chat model capable of processing and maintaining coherence across text sequences up to 32K tokens. The model is optimized for local deployment through precision quantization, which reduces memory requirements to allow execution on consumer-grade hardware. It supports distributing model weights across multiple graphics cards to handle parameters that exceed the memory of a single device. The project covers capabilities for

    Python
    عرض على GitHub↗15,565
  • thudm/chatglm-6bالصورة الرمزية لـ THUDM

    THUDM/ChatGLM-6B

    41,040عرض على GitHub↗

    ChatGLM-6B is an open-source bilingual large language model designed for natural dialogue and text generation in both English and Chinese. It is structured as a dialogue model capable of tasks such as role-playing and information extraction. The project provides implementations for quantized language models, using low-precision weights to reduce GPU memory requirements for local inference. It also supports parameter-efficient fine-tuning, allowing model behavior to be optimized for specific tasks without requiring full retraining. The model includes capabilities for local execution on GPUs a

    Python
    عرض على GitHub↗41,040
  • lc1332/luotuo-chinese-llmالصورة الرمزية لـ LC1332

    LC1332/Luotuo-Chinese-LLM

    3,600عرض على GitHub↗

    骆驼(Luotuo): Open Sourced Chinese Language Models. Developed by 陈启源 @ 华中师范大学 & 李鲁鲁 @ 商汤科技 & 冷子昂 @ 商汤科技

    Jupyter Notebook
    عرض على GitHub↗3,600
  • skyworkai/skyworkالصورة الرمزية لـ SkyworkAI

    SkyworkAI/Skywork

    1,495عرض على GitHub↗

    Skywork series models are pre-trained on 3.2TB of high-quality multilingual (mainly Chinese and English) and code data. We have open-sourced the model, training data, evaluation data, evaluation methods, etc.

    Pythonllm
    عرض على GitHub↗1,495

بحث بالذكاء الاصطناعي

استكشف المزيد من المستودعات الرائعة

صف ما تحتاجه بلغة بسيطة — وسيقوم الذكاء الاصطناعي بترتيب آلاف المشاريع مفتوحة المصدر المنسقة حسب الصلة.

Find more with AI search
  • thudm/chatglm3الصورة الرمزية لـ THUDM

    THUDM/ChatGLM3

    13,676عرض على GitHub↗

    ChatGLM3 is an open-weights large language model designed for bilingual conversational interactions in English and Chinese. It functions as a tool-augmented system capable of calling external functions and executing internal code to resolve complex tasks. The model utilizes four-bit quantization to reduce memory requirements, enabling inference on consumer hardware and diverse processing units including GPUs and CPUs. It features an expanded context window for processing and summarizing long documents and includes a supervised fine-tuning pipeline for adapting the model to specialized domains

    Python
    عرض على GitHub↗13,676
  • vivo-ai-lab/bluelmالصورة الرمزية لـ vivo-ai-lab

    vivo-ai-lab/BlueLM

    941عرض على GitHub↗

    BlueLM(蓝心大模型): Open large language models developed by vivo AI Lab

    Python
    عرض على GitHub↗941
  • hit-scir/huoziالصورة الرمزية لـ HIT-SCIR

    HIT-SCIR/huozi

    395عرض على GitHub↗

    活字通用大模型

    Pythonfine-tuninglarge-language-modelsllm
    عرض على GitHub↗395
  • internlm/internlmالصورة الرمزية لـ InternLM

    InternLM/InternLM

    7,224عرض على GitHub↗

    InternLM is a large language model and a comprehensive suite of weights designed for text generation and complex reasoning. It functions as an inference engine for serving responses, a fine-tuning framework for adjusting model weights, and a platform for building autonomous AI agents. The system is capable of processing long-context input sequences up to one million tokens for document analysis. It employs chain-of-thought reasoning to solve knowledge-intensive tasks by generating intermediate logic steps before producing a final answer. The project covers model weight optimization through s

    Pythonchatbotchinesefine-tuning-llm
    عرض على GitHub↗7,224
  • openlmlab/mossالصورة الرمزية لـ OpenLMLab

    OpenLMLab/MOSS

    12,140عرض على GitHub↗

    MOSS is a conversational AI platform, fine-tuning toolkit, and quantized model runtime. It provides a framework for deploying large language models capable of multi-turn dialogue, general-purpose response generation, and following complex instructions. The system functions as a tool-augmented framework that extends model knowledge through external plugins and tool-call loops. This allows the model to execute tasks via search engines and calculators to augment responses with external data. The project covers model training through supervised conversational fine-tuning and optimizes deployment

    Python
    عرض على GitHub↗12,140
  • openbmb/minicpmالصورة الرمزية لـ OpenBMB

    OpenBMB/MiniCPM

    9,464عرض على GitHub↗

    MiniCPM is a collection of small language models designed for local, on-device deployment in resource-constrained environments. The project focuses on running dense Transformer models on consumer hardware, including GPUs, CPUs, and Apple Silicon, without requiring custom code forks. The project distinguishes itself through heavy optimization for edge hardware, utilizing quantized weight compression in GGUF and MLX formats to reduce memory overhead. It implements advanced inference techniques such as speculative sampling and radix-tree prefix caching to accelerate generation speed and throughp

    Jupyter Notebook
    عرض على GitHub↗9,464
  • xverse-ai/xverse-13bالصورة الرمزية لـ xverse-ai

    xverse-ai/XVERSE-13B

    642عرض على GitHub↗

    XVERSE-13B: A multilingual large language model developed by XVERSE Technology Inc.

    Python
    عرض على GitHub↗642
  • wenge-research/yayi2الصورة الرمزية لـ wenge-research

    wenge-research/YAYI2

    2,799عرض على GitHub↗

    YAYI 2 是中科闻歌研发的新一代开源大语言模型,采用了超过 2 万亿 Tokens 的高质量、多语言语料进行预训练。(Repo for YaYi 2 Chinese LLMs)

    Pythonartificial-intelligencechatchinese
    عرض على GitHub↗2,799
  • xverse-ai/xverse-7bالصورة الرمزية لـ xverse-ai

    xverse-ai/XVERSE-7B

    51عرض على GitHub↗

    XVERSE-7B: A multilingual large language model developed by XVERSE Technology Inc.

    Python
    عرض على GitHub↗51
  • xverse-ai/xverse-65bالصورة الرمزية لـ xverse-ai

    xverse-ai/XVERSE-65B

    139عرض على GitHub↗

    XVERSE-65B: A multilingual large language model developed by XVERSE Technology Inc.

    Python
    عرض على GitHub↗139
  • 01-ai/yiالصورة الرمزية لـ 01-ai

    01-ai/Yi

    7,822عرض على GitHub↗

    Yi is a bilingual language model and foundation model designed for natural language processing, reasoning, and reading comprehension in both English and Chinese. It is built as a transformer-based architecture capable of general purpose text generation and conversational tasks. The model is distinguished by its ability to function as a long context system, processing and analyzing extended input sequences up to 200k tokens. It also supports quantized versions that use low-bit precision to reduce memory footprints, enabling execution on consumer-grade hardware. The project covers a broad rang

    Jupyter Notebooklarge-language-models
    عرض على GitHub↗7,822
  • flagai-open/flagaiالصورة الرمزية لـ FlagAI-Open

    FlagAI-Open/FlagAI

    3,870عرض على GitHub↗

    FlagAI is a distributed deep learning framework and platform designed for the end-to-end lifecycle of large-scale foundation models. It provides a toolkit for training, fine-tuning, and deploying large language models and multi-modal systems across multi-node computing clusters. The project features hardware-agnostic compute abstractions to ensure consistent execution across different accelerators. It includes a dedicated library for parameter-efficient fine-tuning, allowing large neural networks to be adapted to specific tasks with minimal parameter updates and reduced computational overhead

    Python
    عرض على GitHub↗3,870
  • baichuan-inc/baichuan-13bالصورة الرمزية لـ baichuan-inc

    baichuan-inc/Baichuan-13B

    2,931عرض على GitHub↗

    A 13B large language model developed by Baichuan Intelligent Technology

    Pythonartificial-intelligencebenchmarkceval
    عرض على GitHub↗2,931
  • ieit-yuan/yuan-2.0الصورة الرمزية لـ IEIT-Yuan

    IEIT-Yuan/Yuan-2.0

    688عرض على GitHub↗

    Yuan 2.0 Large Language Model

    Python
    عرض على GitHub↗688
  • baichuan-inc/baichuan-7bالصورة الرمزية لـ baichuan-inc

    baichuan-inc/Baichuan-7B

    5,654عرض على GitHub↗

    Baichuan-7B is an open-source 7 billion parameter bilingual Transformer model designed for text generation and few-shot learning across Chinese and English. It is built on a large Transformer architecture trained on a bilingual corpus, enabling it to produce coherent text in both languages from a single model. The model incorporates several optimization techniques that distinguish it from standard large language models. It uses rotary position embeddings that can extrapolate to longer sequences than seen during training, allowing context extension beyond the original 4096-token training lengt

    Pythonartificial-intelligencecevalchatgpt
    عرض على GitHub↗5,654
  • lianjiatech/belleالصورة الرمزية لـ LianjiaTech

    LianjiaTech/BELLE

    8,273عرض على GitHub↗

    BELLE is a specialized implementation of Chinese conversational large language models, encompassing a full instruction tuning framework. It provides a pipeline for training, evaluating, and deploying models optimized for natural language understanding and dialogue tasks in the Chinese language. The project is distinguished by its integrated approach to model refinement, combining the curation of multi-million entry instruction datasets with a distributed training pipeline. This pipeline supports both full fine-tuning and low-rank adaptation to optimize conversational performance. The system

    HTMLbloomchinese-nlpgpt-evaluation
    عرض على GitHub↗8,273
  • baichuan-inc/baichuan2الصورة الرمزية لـ baichuan-inc

    baichuan-inc/Baichuan2

    4,098عرض على GitHub↗

    Baichuan2 is a collection of pre-trained large language models, including base and chat variants, designed for natural language generation and multi-turn conversational AI. It provides an inference engine and a fine-tuning framework to adapt these models to custom datasets and specialized domains. The project features a quantization toolkit and an inference engine that enable model execution across diverse hardware, including graphics processors, central processors, and specialized accelerators. These tools support low-bit weight quantization to reduce memory usage and increase inference spee

    Pythonartificial-intelligencebenchmarkceval
    عرض على GitHub↗4,098
  • openbmb/cpm-beeالصورة الرمزية لـ OpenBMB

    OpenBMB/CPM-Bee

    2,405عرض على GitHub↗

    百亿参数的中英文双语基座大模型

    Python
    عرض على GitHub↗2,405
  • openai/gpt-2الصورة الرمزية لـ openai

    openai/gpt-2

    24,967عرض على GitHub↗

    This project is a transformer-based language model and autoregressive text generator designed to predict the next token in a sequence to produce human-like prose and synthetic text. It functions as a large language model that utilizes a transformer architecture to learn linguistic patterns from large datasets for unsupervised multitask learning. The repository provides a distribution of pre-trained weights, enabling natural language processing tasks without requiring additional training. This allows the model to perform zero-shot task generalization by applying learned patterns to new tasks.

    Python
    عرض على GitHub↗24,967
  • amzxyz/rime_wanxiangالصورة الرمزية لـ amzxyz

    amzxyz/rime_wanxiang

    2,863عرض على GitHub↗

    This project is a CJK input method framework and configuration set designed for the Rime input engine. It provides a comprehensive system of schemas and dictionary packs to optimize Chinese character entry through pinyin and double-pinyin workflows. The framework is distinguished by its use of Lua-powered extensions that add dynamic utilities, such as inline mathematical calculators, automated timestamps, and text formatting, directly to the input interface. It also features refined word libraries and language models specifically tuned to improve prediction accuracy and first-choice hit rates

    Luadictsrimerime-config
    عرض على GitHub↗2,863
  • camenduru/text-to-video-synthesis-colabالصورة الرمزية لـ camenduru

    camenduru/text-to-video-synthesis-colab

    1,515عرض على GitHub↗

    Text To Video Synthesis Colab

    Jupyter Notebookcolabcolab-notebookcolaboratory
    عرض على GitHub↗1,515
  • deepseek-ai/deepseek-v2الصورة الرمزية لـ deepseek-ai

    deepseek-ai/DeepSeek-V2

    5,014عرض على GitHub↗

    DeepSeek-V2 is a large language model designed for natural language processing and the analysis of long text sequences. It utilizes a mixture-of-experts architecture to balance high performance with inference efficiency. The model employs a sparse routing mechanism and shared expert neurons to capture common knowledge while maintaining specialization. It further reduces memory overhead and increases throughput through multi-head latent attention, group-query attention, and low-rank tensor compression. These capabilities enable the processing and retrieval of information from extensive token

    عرض على GitHub↗5,014
  • deepseek-ai/deepseek-llmالصورة الرمزية لـ deepseek-ai

    deepseek-ai/deepseek-LLM

    7,100عرض على GitHub↗

    DeepSeek-LLM is a large language model and causal language model designed for natural language generation. It functions as a multi-lingual system capable of predicting the next token in a sequence to perform text completion and conversational generation. The model is specialized for logical reasoning, specifically as a code and math LLM. This enables it to perform complex problem solving, which includes generating executable code and solving mathematical equations through step-by-step analysis. The system's broader capabilities cover conversational AI, including the generation of chat comple

    Makefile
    عرض على GitHub↗7,100
  • bowang-lab/scgptالصورة الرمزية لـ bowang-lab

    bowang-lab/scGPT

    1,585عرض على GitHub↗

    This is the official codebase for scGPT: Towards Building a Foundation Model for Single-Cell Multi-omics Using Generative AI.

    Jupyter Notebookfoundation-modelgptsingle-cell
    عرض على GitHub↗1,585
  • duomo/transgptالصورة الرمزية لـ DUOMO

    DUOMO/TransGPT

    839عرض على GitHub↗

    🤗 TransGPT-7B • 🤗 TransGPT-MM-6B • 🤖 DUOMO • 💬 WeChat

    Python
    عرض على GitHub↗839
  • baaivision/emuالصورة الرمزية لـ baaivision

    baaivision/Emu

    1,775عرض على GitHub↗

    Emu Series: Generative Multimodal Models from BAAI

    Python
    عرض على GitHub↗1,775