awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
XiaomiMiMo avatar

XiaomiMiMo/MiMo

0
View on GitHub↗
2,257 stars·102 forks·Python·Apache-2.0·23 views

MiMo

MiMo: Unlocking the Reasoning Potential of Language Model – From Pretraining to Posttraining

Features

  • Frontier Reasoning Models - Reasoning model optimized from pretraining through post-training.
  • Large Language Models - Conversational model designed for balanced performance and efficiency.
  • Reasoning Models - Multimodal reasoning model.
  • Research Papers - Research on multimodal model architectures.

Star history

Star history chart for xiaomimimo/mimoStar history chart for xiaomimimo/mimo

How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Frequently asked questions

What does xiaomimimo/mimo do?

MiMo: Unlocking the Reasoning Potential of Language Model – From Pretraining to Posttraining

What are the main features of xiaomimimo/mimo?

The main features of xiaomimimo/mimo are: Frontier Reasoning Models, Large Language Models, Reasoning Models, Research Papers.

Which projects share features with xiaomimimo/mimo?

Projects with overlapping indexed features include: qwenlm/qwen3 — Qwen3 is a transformer-based large language model designed as a generative AI foundation for understanding, reasoning,… moonshotai/kimi-k2 — Kimi-K2 is a conversational AI engine and reasoning framework designed for text generation, advanced problem solving,… deepseek-ai/deepseek-r1 — DeepSeek-R1 is an open-weights large language model focused on advanced reasoning. It uses chain-of-thought processing… qwenlm/qwen2.5 — Qwen2.5 is a suite of large language model foundation models designed for natural language generation, code… nvidia/megatron-lm — Megatron-LM is a distributed transformer training library and large language model training framework designed to… open-reasoner-zero/open-reasoner-zero — An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model.

Projects sharing features with MiMo

These projects share indexed features with MiMo. Shared tags can include platform or build tooling; verify the primary use case before treating a result as a replacement.
  • qwenlm/qwen3QwenLM avatar

    QwenLM/Qwen3

    27,324View on GitHub↗

    Qwen3 is a transformer-based large language model designed as a generative AI foundation for understanding, reasoning, and generating human language. It functions as a comprehensive ecosystem for model training, fine-tuning, and production-ready inference, providing the underlying architecture and weights necessary to build diverse artificial intelligence applications. The project distinguishes itself through extensive support for model quantization and distributed inference, enabling efficient execution across a wide range of hardware from consumer-grade devices to scalable cloud infrastruct

    Python
    View on GitHub↗27,324
  • moonshotai/kimi-k2MoonshotAI avatar

    MoonshotAI/Kimi-K2

    10,401View on GitHub↗

    Kimi-K2 is a conversational AI engine and reasoning framework designed for text generation, advanced problem solving, and coding tasks. It functions as a tool-augmented language model capable of producing human-like chat responses through a compatible model interface. The system utilizes a reasoning-optimized architecture that separates standard conversational flow from deep logical processing. This allows the model to execute autonomous tasks by invoking external functions and calling APIs to retrieve real-time data. The project supports structured JSON output parsing for function-call inte

    View on GitHub↗10,401
  • deepseek-ai/deepseek-r1deepseek-ai avatar

    deepseek-ai/DeepSeek-R1

    91,996View on GitHub↗

    DeepSeek-R1 is an open-weights large language model focused on advanced reasoning. It uses chain-of-thought processing and internal monologues to solve complex mathematical and logical problems by breaking tasks into sequential, verifiable thought processes. The model is developed using reinforcement learning to optimize reasoning patterns and verify logical steps. It employs a distillation process to transfer these high-performance logic capabilities from a large teacher model into smaller, computationally efficient versions. The training framework incorporates group relative policy optimiz

    View on GitHub↗91,996
  • qwenlm/qwen2.5QwenLM avatar

    QwenLM/Qwen2.5

    27,307View on GitHub↗

    Qwen2.5 is a suite of large language model foundation models designed for natural language generation, code production, and complex mathematical reasoning. The project encompasses a multilingual language model capable of processing dozens of languages and a specialized code generation model for technical problem solving and debugging. The framework is distinguished by its long context capabilities, enabling the analysis of massive inputs ranging from 256K up to 1 million tokens. It further functions as an agentic framework, utilizing standardized templates and parsers to execute autonomous wo

    Python
    View on GitHub↗27,307
Compare all 30 related projects→