awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
Back to xiaomimimo/mimo

Projects sharing features with MiMo

30 open-source projects similar to xiaomimimo/mimo, ranked by shared indexed features. Tags may describe platforms or build tools rather than the same primary purpose. Check each project’s use case, license, and deployment requirements before treating it as a replacement.

  • qwenlm/qwen3QwenLM avatar

    QwenLM/Qwen3

    27,324View on GitHub↗

    Qwen3 is a transformer-based large language model designed as a generative AI foundation for understanding, reasoning, and generating human language. It functions as a comprehensive ecosystem for model training, fine-tuning, and production-ready inference, providing the underlying architecture and weights necessary to build diverse artificial intelligence applications. The project distinguishes itself through extensive support for model quantization and distributed inference, enabling efficient execution across a wide range of hardware from consumer-grade devices to scalable cloud infrastruct

    Python
    View on GitHub↗27,324
  • deepseek-ai/deepseek-r1deepseek-ai avatar

    deepseek-ai/DeepSeek-R1

    91,996View on GitHub↗

    DeepSeek-R1 is an open-weights large language model focused on advanced reasoning. It uses chain-of-thought processing and internal monologues to solve complex mathematical and logical problems by breaking tasks into sequential, verifiable thought processes. The model is developed using reinforcement learning to optimize reasoning patterns and verify logical steps. It employs a distillation process to transfer these high-performance logic capabilities from a large teacher model into smaller, computationally efficient versions. The training framework incorporates group relative policy optimiz

    View on GitHub↗91,996
  • moonshotai/kimi-k2MoonshotAI avatar

    MoonshotAI/Kimi-K2

    10,401View on GitHub↗

    Kimi-K2 is a conversational AI engine and reasoning framework designed for text generation, advanced problem solving, and coding tasks. It functions as a tool-augmented language model capable of producing human-like chat responses through a compatible model interface. The system utilizes a reasoning-optimized architecture that separates standard conversational flow from deep logical processing. This allows the model to execute autonomous tasks by invoking external functions and calling APIs to retrieve real-time data. The project supports structured JSON output parsing for function-call inte

    View on GitHub↗10,401

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Find more with AI search
  • qwenlm/qwen2.5QwenLM avatar

    QwenLM/Qwen2.5

    27,307View on GitHub↗

    Qwen2.5 is a suite of large language model foundation models designed for natural language generation, code production, and complex mathematical reasoning. The project encompasses a multilingual language model capable of processing dozens of languages and a specialized code generation model for technical problem solving and debugging. The framework is distinguished by its long context capabilities, enabling the analysis of massive inputs ranging from 256K up to 1 million tokens. It further functions as an agentic framework, utilizing standardized templates and parsers to execute autonomous wo

    Python
    View on GitHub↗27,307
  • open-reasoner-zero/open-reasoner-zeroOpen-Reasoner-Zero avatar

    Open-Reasoner-Zero/Open-Reasoner-Zero

    2,095View on GitHub↗

    An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model

    Python
    View on GitHub↗2,095
  • skyworkai/skywork-r1vSkyworkAI avatar

    SkyworkAI/Skywork-R1V

    3,159View on GitHub↗

    Skywork-R1V4

    Python
    View on GitHub↗3,159
  • nvidia/megatron-lmNVIDIA avatar

    NVIDIA/Megatron-LM

    16,731View on GitHub↗

    Megatron-LM is a distributed transformer training library and large language model training framework designed to scale models across thousands of GPUs. It functions as a GPU-optimized deep learning toolkit and a scaling engine for mixture-of-experts architectures, enabling the training of models with hundreds of billions of parameters. The project implements multi-dimensional model parallelism, combining tensor, pipeline, data, expert, and context-based workload distribution. It specifically optimizes mixture-of-experts architectures through integrated memory and communication improvements t

    Python
    View on GitHub↗16,731
  • minimax-ai/minimax-m1MiniMax-AI avatar

    MiniMax-AI/MiniMax-M1

    3,159View on GitHub↗

    MiniMax-M1, the world's first open-weight, large-scale hybrid-attention reasoning model.

    Python
    View on GitHub↗3,159
  • stepfun-ai/step3stepfun-ai avatar

    stepfun-ai/Step3

    453View on GitHub↗
    View on GitHub↗453
  • zai-org/glm-4.5zai-org avatar

    zai-org/GLM-4.5

    4,210View on GitHub↗

    GLM-4.5 is a multimodal large language model and advanced reasoning system. It functions as an AI coding assistant, an autonomous AI agent, and a multimodal content generator capable of processing and generating text, images, audio, and video within a single unified system. The project is distinguished by its deep reasoning capabilities, utilizing chain-of-thought processing to solve complex mathematical, logical, and technical problems. It features an agentic architecture that allows for autonomous task execution, long-horizon goal planning, and the ability to interact with external tools an

    Pythonagentglmllm
    View on GitHub↗4,210
  • skyworkai/skywork-or1SkyworkAI avatar

    SkyworkAI/Skywork-OR1

    745View on GitHub↗

    ✊ Unleashing the Power of Reinforcement Learning for Math and Code Reasoners 🤖

    Python
    View on GitHub↗745
  • atfortes/awesome-llm-reasoningatfortes avatar

    atfortes/Awesome-LLM-Reasoning

    3,640View on GitHub↗

    From Chain-of-Thought prompting to OpenAI o1 and DeepSeek-R1 🍓

    awesomechain-of-thoughtchatgpt
    View on GitHub↗3,640
  • airobotlab/kochatgptA

    airobotlab/KoChatGPT

    0View on GitHub↗
    View on GitHub↗0
  • baibizhe/efficient-r1-vllmB

    baibizhe/Efficient-R1-VLLM

    0View on GitHub↗
    View on GitHub↗0
  • artidoro/qloraartidoro avatar

    artidoro/qlora

    10,929View on GitHub↗

    This project is a quantized fine-tuning framework for large language models. It implements a low-rank adaptation library and a four-bit quantizer to reduce the GPU memory requirements needed to train large models. The framework utilizes four-bit quantization and low-rank adapters to enable model training on consumer-grade hardware. It further reduces the memory footprint through double quantization and a paged optimizer that offloads states to system RAM. The system supports distributed training across multiple GPUs to handle larger parameter scales and includes utilities for custom dataset

    Jupyter Notebook
    View on GitHub↗10,929
  • appletea233/temporal-r1appletea233 avatar

    appletea233/Temporal-R1

    62View on GitHub↗

    Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency

    Python
    View on GitHub↗62
  • aigc-audio/audiogptAIGC-Audio avatar

    AIGC-Audio/AudioGPT

    10,174View on GitHub↗

    AudioGPT is an LLM-driven audio framework and processing suite that uses large language models to orchestrate neural audio pipelines. It functions as a multimodal audio generator and processing system, integrating a collection of pretrained models to handle speech synthesis, sound generation, and audio manipulation. The system is distinguished by its ability to generate audio from diverse inputs, including text and images, and its capacity to produce synchronized talking head videos. It also operates as a neural speech translator, converting spoken language between different tongues while pre

    Pythonaudiogptmusic
    View on GitHub↗10,174
  • agentica-project/deepscalerA

    agentica-project/deepscaler

    0View on GitHub↗
    View on GitHub↗0
  • bklieger-groq/g1B

    bklieger-groq/g1

    0View on GitHub↗
    View on GitHub↗0
  • antimatter15/alpaca.cppantimatter15 avatar

    antimatter15/alpaca.cpp

    10,138View on GitHub↗

    alpaca.cpp is a high-performance local inference engine implemented in C++ for executing instruction-tuned large language models. It serves as a quantized model runtime designed to load and run model tensors on local hardware with minimal dependencies, removing the requirement for a full Python environment. The project focuses on on-device text generation and the deployment of private AI chatbots. It utilizes model weight quantization to reduce memory requirements and increase inference speed on consumer-grade devices. The system covers hardware-optimized model execution through thread-pool

    C
    View on GitHub↗10,138
  • bigscience-workshop/petalsbigscience-workshop avatar

    bigscience-workshop/petals

    10,208View on GitHub↗

    Petals is a decentralized framework and inference engine for running large language models across a peer-to-peer network. It enables the execution of models that exceed the memory of any single machine by splitting computations and model layers across a collaborative swarm of GPUs. The system functions as a collaborative compute network where participants share local GPU resources and host model weights. It supports distributed prompt-tuning to adapt massive models to specific tasks and allows for the establishment of private compute swarms to process sensitive data within restricted, trusted

    Python
    View on GitHub↗10,208
  • brendanhogan/deepseekrl-extendedbrendanhogan avatar

    brendanhogan/DeepSeekRL-Extended

    252View on GitHub↗

    Exploring Applications of GRPO

    Python
    View on GitHub↗252
  • bytedance-seed/seed-ossByteDance-Seed avatar

    ByteDance-Seed/seed-oss

    886View on GitHub↗

    👋 Hi, everyone! We are ByteDance Seed Team.

    Python
    View on GitHub↗886
  • bytedance-seed/seed-thinking-v1.5B

    ByteDance-Seed/Seed-Thinking-v1.5

    0View on GitHub↗
    View on GitHub↗0
  • chroma-core/chromachroma-core avatar

    chroma-core/chroma

    26,198View on GitHub↗

    Chroma is a specialized vector database designed to index and retrieve high-dimensional data representations for semantic similarity search. It functions as a comprehensive platform for information retrieval, enabling the storage and management of unstructured documents alongside structured metadata. By mapping data into numerical representations, the system facilitates rapid similarity lookups across large datasets. The platform distinguishes itself through a hybrid search infrastructure that combines dense vector embeddings with sparse keyword and regular expression matching to balance sema

    Rustaidatabasedocument-retrieval
    View on GitHub↗26,198
  • codedotal/gpt-code-clippyCodedotAl avatar

    CodedotAl/gpt-code-clippy

    3,268View on GitHub↗

    Please refer to our new GitHub Wiki which documents our efforts in detail in creating the open source version of GitHub Copilot

    Python
    View on GitHub↗3,268
  • confident-ai/deepevalconfident-ai avatar

    confident-ai/deepeval

    13,733View on GitHub↗

    Deepeval is a framework for testing and evaluating large language model applications. It provides a suite of tools for executing automated regression tests, validating model output quality against defined standards, and tracing the execution of complex agent workflows. By integrating these capabilities into development pipelines, the platform ensures consistent performance and reliability throughout the software lifecycle. The platform distinguishes itself through its focus on programmatic validation and observability. It utilizes secondary language models to score output quality and employs

    Pythonevaluation-frameworkevaluation-metricsllm-evaluation
    View on GitHub↗13,733
  • cranot/chatbot-injections-exploitsCranot avatar

    Cranot/chatbot-injections-exploits

    407View on GitHub↗

    ChatBot Injection and Exploit Examples: A Curated List of Prompt Engineer Commands - ChatGPT

    View on GitHub↗407
  • csfufu/revisual-r1C

    CSfufu/Revisual-R1

    0View on GitHub↗
    View on GitHub↗0
  • bigcode-project/starcoder2bigcode-project avatar

    bigcode-project/starcoder2

    2,075View on GitHub↗

    StarCoder2 is a family of code generation models (3B, 7B, and 15B), trained on 600+ programming languages from The Stack v2 and some natural language text such as Wikipedia, Arxiv, and GitHub issues. The models use Grouped Query Attention, a context window of 16,384 tokens, with sliding window…

    Python
    View on GitHub↗2,075