awesome-repositories.com
Blog
MCP
awesome-repositories.com

Entdecke die besten Open-Source-Repositories mit KI-gestützter Suche.

EntdeckenKuratierte SuchenOpen-Source-AlternativenSelf-hosted SoftwareBlogSitemap
ProjektMCP-ServerÜber unsRanking-MethodikPresse
RechtlichesDatenschutzAGB
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
Back to ruc-gsai/yulan-mini

Open-source alternatives to YuLan Mini

30 open-source projects similar to ruc-gsai/yulan-mini, ranked by how many features they have in common. Compare stars, activity and what each one does to find the best YuLan Mini alternative.

  • openlmlab/mossAvatar von OpenLMLab

    OpenLMLab/MOSS

    12,140Auf GitHub ansehen↗

    MOSS is a conversational AI platform, fine-tuning toolkit, and quantized model runtime. It provides a framework for deploying large language models capable of multi-turn dialogue, general-purpose response generation, and following complex instructions. The system functions as a tool-augmented framework that extends model knowledge through external plugins and tool-call loops. This allows the model to execute tasks via search engines and calculators to augment responses with external data. The project covers model training through supervised conversational fine-tuning and optimizes deployment

    Python
    Auf GitHub ansehen↗12,140
  • thudm/chatglm2-6bAvatar von THUDM

    THUDM/ChatGLM2-6B

    15,565Auf GitHub ansehen↗

    ChatGLM2-6B is an open-weight large language model designed for natural language conversations and text generation in both English and Chinese. It functions as a bilingual chat model capable of processing and maintaining coherence across text sequences up to 32K tokens. The model is optimized for local deployment through precision quantization, which reduces memory requirements to allow execution on consumer-grade hardware. It supports distributing model weights across multiple graphics cards to handle parameters that exceed the memory of a single device. The project covers capabilities for

    Python
    Auf GitHub ansehen↗15,565
  • thudm/chatglm-6bAvatar von THUDM

    THUDM/ChatGLM-6B

    41,040Auf GitHub ansehen↗

    ChatGLM-6B is an open-source bilingual large language model designed for natural dialogue and text generation in both English and Chinese. It is structured as a dialogue model capable of tasks such as role-playing and information extraction. The project provides implementations for quantized language models, using low-precision weights to reduce GPU memory requirements for local inference. It also supports parameter-efficient fine-tuning, allowing model behavior to be optimized for specific tasks without requiring full retraining. The model includes capabilities for local execution on GPUs a

    Python
    Auf GitHub ansehen↗41,040

KI-Suche

Entdecke weitere awesome Repositories

Beschreibe in einfachen Worten, was du brauchst — die KI bewertet tausende kuratierte Open-Source-Projekte nach Relevanz.

Find more with AI search
  • thudm/chatglm3Avatar von THUDM

    THUDM/ChatGLM3

    13,676Auf GitHub ansehen↗

    ChatGLM3 is an open-weights large language model designed for bilingual conversational interactions in English and Chinese. It functions as a tool-augmented system capable of calling external functions and executing internal code to resolve complex tasks. The model utilizes four-bit quantization to reduce memory requirements, enabling inference on consumer hardware and diverse processing units including GPUs and CPUs. It features an expanded context window for processing and summarizing long documents and includes a supervised fine-tuning pipeline for adapting the model to specialized domains

    Python
    Auf GitHub ansehen↗13,676
  • skyworkai/skyworkAvatar von SkyworkAI

    SkyworkAI/Skywork

    1,495Auf GitHub ansehen↗

    Skywork series models are pre-trained on 3.2TB of high-quality multilingual (mainly Chinese and English) and code data. We have open-sourced the model, training data, evaluation data, evaluation methods, etc.

    Pythonllm
    Auf GitHub ansehen↗1,495
  • openai/gpt-2Avatar von openai

    openai/gpt-2

    24,967Auf GitHub ansehen↗

    This project is a transformer-based language model and autoregressive text generator designed to predict the next token in a sequence to produce human-like prose and synthetic text. It functions as a large language model that utilizes a transformer architecture to learn linguistic patterns from large datasets for unsupervised multitask learning. The repository provides a distribution of pre-trained weights, enabling natural language processing tasks without requiring additional training. This allows the model to perform zero-shot task generalization by applying learned patterns to new tasks.

    Python
    Auf GitHub ansehen↗24,967
  • karpathy/ng-video-lectureAvatar von karpathy

    karpathy/ng-video-lecture

    4,798Auf GitHub ansehen↗

    This project is an educational implementation of a small-scale generative pre-trained transformer designed to teach the fundamentals of neural network architecture and training. It serves as a reference implementation and tutorial for constructing a text-generating neural network from scratch. The codebase demonstrates the mechanics of tokenization, self-attention, and the construction of a lightweight language model. It focuses on the step-by-step process of building a generative model to illustrate how large language models are constructed. The implementation covers transformer-based archi

    Python
    Auf GitHub ansehen↗4,798
  • baichuan-inc/baichuan-13bAvatar von baichuan-inc

    baichuan-inc/Baichuan-13B

    2,931Auf GitHub ansehen↗

    A 13B large language model developed by Baichuan Intelligent Technology

    Pythonartificial-intelligencebenchmarkceval
    Auf GitHub ansehen↗2,931
  • atla-ai/selene-miniAvatar von atla-ai

    atla-ai/selene-mini

    30Auf GitHub ansehen↗

    🛝 Playground | 📄 Technical report | 💻 GitHub | 👀 Sign up for the API

    Jupyter Notebook
    Auf GitHub ansehen↗30
  • abacaj/mpt-30b-inferenceAvatar von abacaj

    abacaj/mpt-30B-inference

    574Auf GitHub ansehen↗

    Run inference on the latest MPT-30B model using your CPU. This inference code uses a ggml quantized model. To run the model we'll use a library called ctransformers that has bindings to ggml in python.

    Python
    Auf GitHub ansehen↗574
  • charent/chatlm-mini-chineseAvatar von charent

    charent/ChatLM-mini-Chinese

    1,711Auf GitHub ansehen↗

    中文对话0.2B小模型(ChatLM-Chinese-0.2B),开源所有数据集来源、数据清洗、tokenizer训练、模型预训练、SFT指令微调、RLHF优化等流程的全部代码。支持下游任务sft微调,给出三元组信息抽取微调示例。

    Pythonchatbotlanguage-modelt5-model
    Auf GitHub ansehen↗1,711
  • charent/phi2-mini-chineseC

    charent/Phi2-mini-Chinese

    0Auf GitHub ansehen↗
    Auf GitHub ansehen↗0
  • chinese-tiny-llm/chinese-tiny-llmC

    Chinese-Tiny-LLM/Chinese-Tiny-LLM

    0Auf GitHub ansehen↗
    Auf GitHub ansehen↗0
  • clue-ai/chatyuanAvatar von clue-ai

    clue-ai/ChatYuan

    1,870Auf GitHub ansehen↗

    ChatYuan: Large Language Model for Dialogue in Chinese and English

    Python
    Auf GitHub ansehen↗1,870
  • blinkdl/rwkv-lmAvatar von BlinkDL

    BlinkDL/RWKV-LM

    14,568Auf GitHub ansehen↗

    RWKV-LM is a framework for training and deploying recurrent language models. It utilizes a linear-time recurrent architecture that enables text generation and sequence processing with constant memory and time complexity, avoiding the quadratic scaling of traditional attention caches. The project implements a parallelizable training mechanism that allows recurrent models to be trained using global operations while maintaining cache-free inference. It includes state-tuning capabilities to optimize the initial hidden state and utilizes adaptive probability-mass sampling to control token diversit

    Python
    Auf GitHub ansehen↗14,568
  • blinkdl/chatrwkvAvatar von BlinkDL

    BlinkDL/ChatRWKV

    9,492Auf GitHub ansehen↗

    ChatRWKV is an open-source frontend and GPU-accelerated inference engine designed for interacting with RWKV recurrent neural network language models. It provides a self-hosted web chat interface and a specialized client for generating human-like text using a linear-complexity architecture. The project utilizes a GPU-accelerated backend that employs custom CUDA kernels and dynamic model format conversion to increase processing speed and reduce memory overhead. It manages conversation history through state-based context management, updating a fixed-size hidden state to maintain a constant memor

    Pythonchatbotchatgptlanguage-model
    Auf GitHub ansehen↗9,492
  • allenai/olmoAvatar von allenai

    allenai/OLMo

    6,313Auf GitHub ansehen↗
    Python
    Auf GitHub ansehen↗6,313
  • databrickslabs/dollyAvatar von databrickslabs

    databrickslabs/dolly

    10,795Auf GitHub ansehen↗

    Dolly is an instruction-tuned large language model designed to follow complex natural language directions. It operates as a causal language model that predicts the next token in a sequence to generate coherent conversational responses and perform tasks such as brainstorming, classification, and question answering. The project focuses on the development of models using open datasets suitable for commercial application. It enables the creation of instruction-following models by utilizing curated collections of human-generated instruction-response pairs. The repository provides capabilities for

    Python
    Auf GitHub ansehen↗10,795
  • deepseek-ai/deepseek-llmAvatar von deepseek-ai

    deepseek-ai/deepseek-LLM

    7,100Auf GitHub ansehen↗

    DeepSeek-LLM is a large language model and causal language model designed for natural language generation. It functions as a multi-lingual system capable of predicting the next token in a sequence to perform text completion and conversational generation. The model is specialized for logical reasoning, specifically as a code and math LLM. This enables it to perform complex problem solving, which includes generating executable code and solving mathematical equations through step-by-step analysis. The system's broader capabilities cover conversational AI, including the generation of chat comple

    Makefile
    Auf GitHub ansehen↗7,100
  • deepseek-ai/deepseek-v2Avatar von deepseek-ai

    deepseek-ai/DeepSeek-V2

    5,014Auf GitHub ansehen↗

    DeepSeek-V2 is a large language model designed for natural language processing and the analysis of long text sequences. It utilizes a mixture-of-experts architecture to balance high performance with inference efficiency. The model employs a sparse routing mechanism and shared expert neurons to capture common knowledge while maintaining specialization. It further reduces memory overhead and increases throughput through multi-head latent attention, group-query attention, and low-rank tensor compression. These capabilities enable the processing and retrieval of information from extensive token

    Auf GitHub ansehen↗5,014
  • dllxw/baby-llama2-chineseAvatar von DLLXW

    DLLXW/baby-llama2-chinese

    2,891Auf GitHub ansehen↗

    This project is a training pipeline and framework for developing Chinese language models based on the Llama 2 architecture. It functions as a distributed GPU trainer and dataset preprocessing toolkit designed for both the initial pre-training of baseline models and subsequent supervised fine-tuning. The system distinguishes itself through a specialized workflow for Chinese text, incorporating a data curation pipeline that uses similarity hashing for deduplication and a tokenization process that converts raw text into memory-mapped binary files for efficient disk access. It implements a superv

    Python
    Auf GitHub ansehen↗2,891
  • ecnu-icalk/educhatAvatar von ECNU-ICALK

    ECNU-ICALK/EduChat

    940Auf GitHub ansehen↗

    An open-source educational chat model from ICALK, East China Normal University. 开源中英教育对话大模型。(通用基座模型,GPU部署,数据清理) 致敬: LLaMA, MOSS, BELLE, Ziya, vLLM

    Jupyter Notebookbellechinese-nlpdata-cleaning
    Auf GitHub ansehen↗940
  • eleutherai/pythiaAvatar von EleutherAI

    EleutherAI/pythia

    2,827Auf GitHub ansehen↗

    This repository is for EleutherAI's project Pythia which combines interpretability analysis and scaling laws to understand how knowledge develops and evolves during training in autoregressive transformers. For detailed info on the models, their training, and their properties, please see our…

    Jupyter Notebook
    Auf GitHub ansehen↗2,827
  • facebookresearch/llamaAvatar von facebookresearch

    facebookresearch/llama

    59,466Auf GitHub ansehen↗

    Llama is a large language model runtime and inference engine designed to load and execute autoregressive transformer models. It enables the generation of natural language text completions from prompts using pretrained weights. The system features multi-GPU model parallelism, which distributes model weights and workloads across multiple graphics processors to support larger parameter counts. It also incorporates a content safety filter that uses classifiers to intercept and block unsafe inputs or outputs during the inference process. The project covers broad capabilities in distributed model

    Python
    Auf GitHub ansehen↗59,466
  • flagai-open/aquila2Avatar von FlagAI-Open

    FlagAI-Open/Aquila2

    445Auf GitHub ansehen↗

    The official repo of Aquila2 series proposed by BAAI, including pretrained & chat large language models.

    Pythonllmllm-inferencellm-training
    Auf GitHub ansehen↗445
  • flagai-open/flagaiAvatar von FlagAI-Open

    FlagAI-Open/FlagAI

    3,870Auf GitHub ansehen↗

    FlagAI is a distributed deep learning framework and platform designed for the end-to-end lifecycle of large-scale foundation models. It provides a toolkit for training, fine-tuning, and deploying large language models and multi-modal systems across multi-node computing clusters. The project features hardware-agnostic compute abstractions to ensure consistent execution across different accelerators. It includes a dedicated library for parameter-efficient fine-tuning, allowing large neural networks to be adapted to specific tasks with minimal parameter updates and reduced computational overhead

    Python
    Auf GitHub ansehen↗3,870
  • google-research/google-researchAvatar von google-research

    google-research/google-research

    38,139Auf GitHub ansehen↗

    This repository serves as a comprehensive research platform and toolkit for advancing machine learning, quantum computing, and large-scale scientific data analysis. It provides foundational frameworks for developing complex algorithmic systems, offering the necessary infrastructure for distributed training, computational graph execution, and high-performance model development. The project distinguishes itself by integrating specialized research domains with robust, privacy-preserving methodologies. It supports diverse scientific discovery through tools for quantum simulation, physics-informed

    Jupyter Notebookaimachine-learningresearch
    Auf GitHub ansehen↗38,139
  • google-research/t5xAvatar von google-research

    google-research/t5x

    2,972Auf GitHub ansehen↗

    Go to T5X ReadTheDocs Documentation Page.

    Python
    Auf GitHub ansehen↗2,972
  • google-research/text-to-text-transfer-transformerAvatar von google-research

    google-research/text-to-text-transfer-transformer

    6,528Auf GitHub ansehen↗

    This is a machine learning framework for treating diverse natural language processing tasks as a unified text-to-text problem. It provides a toolkit for pre-training and fine-tuning large-scale transformer models, utilizing a system where both inputs and outputs are formatted as raw text sequences. The framework is distinguished by its distributed training system, which uses mesh-based strategies to scale model weights and training batches across multiple TPU cores. It supports multi-task learning by combining diverse datasets into a single training stream using configurable mixture rates, al

    Python
    Auf GitHub ansehen↗6,528
  • dandelionsllm/pandallmAvatar von dandelionsllm

    dandelionsllm/pandallm

    1,033Auf GitHub ansehen↗

    Panda项目是于2023年5月启动的开源海外中文大语言模型项目,致力于大模型时代探索整个技术栈,旨在推动中文自然语言处理领域的创新和合作。

    Python
    Auf GitHub ansehen↗1,033