awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
Back to microsoft/tora

Projects sharing features with ToRA

27 open-source projects similar to microsoft/tora, ranked by shared indexed features. Tags may describe platforms or build tools rather than the same primary purpose. Check each project’s use case, license, and deployment requirements before treating it as a replacement.

  • thudm/glm-4THUDM avatar

    THUDM/GLM-4

    7,059View on GitHub↗

    GLM-4 is an open weights large language model designed as a multimodal chat system. It functions as a reasoning-focused and multilingual model capable of processing and generating responses across text and visual data types. The model is distinguished by its function-calling capabilities, allowing it to interface with external tools and APIs to execute tasks and retrieve real-time information. It is optimized for complex logical reasoning, mathematical problem solving, and deep research involving long-form content generation. Broad capabilities include multilingual text generation, the creat

    Python
    View on GitHub↗7,059
  • deepseek-ai/deepseek-llmdeepseek-ai avatar

    deepseek-ai/deepseek-LLM

    7,100View on GitHub↗

    DeepSeek-LLM is a large language model and causal language model designed for natural language generation. It functions as a multi-lingual system capable of predicting the next token in a sequence to perform text completion and conversational generation. The model is specialized for logical reasoning, specifically as a code and math LLM. This enables it to perform complex problem solving, which includes generating executable code and solving mathematical equations through step-by-step analysis. The system's broader capabilities cover conversational AI, including the generation of chat comple

    Makefile
    View on GitHub↗7,100
  • chengsong-huang/r-zeroChengsong-Huang avatar

    Chengsong-Huang/R-Zero

    822View on GitHub↗

    Check out our paper or webpage for the details

    Python
    View on GitHub↗822
  • ezelikman/starezelikman avatar

    ezelikman/STaR

    227View on GitHub↗

    1. STaR 2. Mesh Transformer JAX 1. Updates 3. Pretrained Models 1. GPT-J-6B 1. Links 2. Acknowledgments 3. License 4. Model Details 5. Zero-Shot Evaluations 4. Architecture and Usage 1. Fine-tuning 2. JAX Dependency 5. TODO

    Python
    View on GitHub↗227

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Find more with AI search
  • hkust-nlp/mstarH

    hkust-nlp/mstar

    0View on GitHub↗

    :star: Project Page

    View on GitHub↗0
  • iamhankai/forest-of-thoughtiamhankai avatar

    iamhankai/Forest-of-Thought

    55View on GitHub↗

    Forest-of-Thought: Scaling Test-Time Compute for Enhancing LLM Reasoning

    Python
    View on GitHub↗55
  • lucidrains/self-rewarding-lm-pytorchlucidrains avatar

    lucidrains/self-rewarding-lm-pytorch

    1,410View on GitHub↗

    Implementation of the training framework proposed in Self-Rewarding Language Model , from MetaAI

    Python
    View on GitHub↗1,410
  • masworks/mas-gptMASWorks avatar

    MASWorks/MAS-GPT

    80View on GitHub↗

    📑 ArXiv Paper | 🤗 Model Weights

    Python
    View on GitHub↗80
  • microsoft/codetM

    microsoft/CodeT

    0View on GitHub↗

    This repository contains projects that aims to equip large-scale pretrained language models with better programming and reasoning skills. These projects are presented by Microsoft Research Asia and Microsoft Azure AI.

    View on GitHub↗0
  • niansong1996/leverniansong1996 avatar

    niansong1996/lever

    90View on GitHub↗

    Code for paper "LEVER: Learning to Verify Language-to-Code Generation with Execution". LEVER is a simple method that improves the code generation ability of large language models trained on code (CodeLMs), by learning to verify and rerank CodeLM-generated programs with their execution results.…

    Python
    View on GitHub↗90
  • nlpxucan/wizardlmnlpxucan avatar

    nlpxucan/WizardLM

    9,486View on GitHub↗

    WizardLM is a large language model and instruction-tuning framework designed to execute sophisticated coding, mathematical, and conversational tasks. It functions as an AI system for mathematical reasoning and code generation, as well as a synthetic dataset generator used to train other language models. The project is distinguished by its evolutionary instruction tuning, which uses a method to rewrite simple instructions into complex tasks. This process expands training dataset difficulty and produces a high volume of open-domain tasks across various difficulty levels. The system covers capa

    Python
    View on GitHub↗9,486
  • osu-nlp-group/deductive-beam-searchOSU-NLP-Group avatar

    OSU-NLP-Group/Deductive-Beam-Search

    21View on GitHub↗

    Deductive Beam Search Decoding Deducible Rationale for Chain-of-Thought Reasoning

    Python
    View on GitHub↗21
  • princeton-nlp/tree-of-thought-llmP

    princeton-nlp/tree-of-thought-llm

    6,007View on GitHub↗

    Note: https://github.com/kyegomez/tree-of-thoughts CANNOT replicate paper results.

    Python
    View on GitHub↗6,007
  • sii-research/siirlsii-research avatar

    sii-research/siiRL

    361View on GitHub↗

    siiRL: Shanghai Innovation Institute RL Framework for Advanced LLMs and Multi-Agent Systems

    Python
    View on GitHub↗361
  • skyworkai/skywork-reward-v2SkyworkAI avatar

    SkyworkAI/Skywork-Reward-V2

    151View on GitHub↗

    Skywork-Reward-V2 is a series of eight reward models designed for versatility across a wide range of tasks, trained on a mixture of 26 million carefully curated preference pairs. While the Skywork-Reward-V2 series remains based on the Bradley-Terry model, we push the boundaries of training data…

    View on GitHub↗151
  • sparkjiao/dpo-trajectory-reasoningSparkJiao avatar

    SparkJiao/dpo-trajectory-reasoning

    84View on GitHub↗

    This repository contains the code for the paper "Learning Planning-based Reasoning with Trajectory Collection and Process Rewards Synthesizing" (EMNLP 2024).

    Python
    View on GitHub↗84
  • spcl/graph-of-thoughtsspcl avatar

    spcl/graph-of-thoughts

    2,805View on GitHub↗

    This is the official implementation of Graph of Thoughts: Solving Elaborate Problems with Large Language Models. This framework gives you the ability to solve complex problems by modeling them as a Graph of Operations (GoO), which is automatically executed with a Large Language Model (LLM) as…

    Python
    View on GitHub↗2,805
  • spiral-rl/spiralspiral-rl avatar

    spiral-rl/spiral

    196View on GitHub↗

    Self-Play on Zero-Sum Games Incentivizes Reasoning via Multi-Agent Multi-Turn Reinforcement Learning

    Python
    View on GitHub↗196
  • tianhongzxy/coreT

    TianHongZXY/CoRe

    0View on GitHub↗

    put the dataset under data/ Set the hyperparameters in train.slurm and execute bash train.slurm Set the hyperparameters in trainverifier.slurm and execute bash trainverifier.slurm After fine-tuning, specify the model path in mcts.slurm, execute bash mcts.slurm. Note that the provided script will…

    View on GitHub↗0
  • tsinghuac3i/ssrlTsinghuaC3I avatar

    TsinghuaC3I/SSRL

    209View on GitHub↗

    📊 Main Results ✨ Getting Started • 📨 Contact • 🎈 Citation • 🌟 Star History

    Python
    View on GitHub↗209
  • wangqinsi1/vision-zerowangqinsi1 avatar

    wangqinsi1/Vision-Zero

    136View on GitHub↗

    A domain-agnostic framework enabling VLM self-improvement through competitive visual games

    Python
    View on GitHub↗136
  • wantbook-book/serlwantbook-book avatar

    wantbook-book/SeRL

    23View on GitHub↗

    SeRL: Self-Play Reinforcement Learning for Large Language Models with Limited Data

    Python
    View on GitHub↗23
  • wecoai/aidemlWecoAI avatar

    WecoAI/aideml

    1,322View on GitHub↗

    AIDE: AI-Driven Exploration in the Space of Code. The machine Learning engineering agent that automates AI R&D.

    Python
    View on GitHub↗1,322
  • yangling0818/buffer-of-thought-llmYangLing0818 avatar

    YangLing0818/buffer-of-thought-llm

    677View on GitHub↗

    Official implementation of our Buffer of Thoughts (BoT) framework (NeurIPS 2024 Spotlight). Affiliation: Peking University, UC Berkeley, Stanford University

    Python
    View on GitHub↗677
  • allenai/open-instructallenai avatar

    allenai/open-instruct

    3,586View on GitHub↗

    Open-Instruct is a distributed training and instruction tuning framework for large language models. It functions as a coordinator for supervised fine-tuning, reinforcement learning from human feedback pipelines, and tool-use training, providing specialized roles for dataset curation and model alignment. The project distinguishes itself through a high-performance training architecture that utilizes actor-based distributed coordination and hybrid sharding to manage large GPU clusters. It implements advanced alignment techniques including direct preference optimization, group relative policy opt

    Python
    View on GitHub↗3,586
  • zhengkid/parallel-r1zhengkid avatar

    zhengkid/Parallel-R1

    259View on GitHub↗

    The official repository for "Parallel-R1: Towards Parallel Thinking via Reinforcement Learning".

    Python
    View on GitHub↗259
  • chengpengli1003/cortChengpengLi1003 avatar

    ChengpengLi1003/CoRT

    72View on GitHub↗
    Python
    View on GitHub↗72