awesome-repositories.com
Blog
MCP
awesome-repositories.com

Descubre los mejores repositorios open-source con nuestra búsqueda potenciada por IA.

ExplorarBúsquedas curadasAlternativas open-sourceSoftware autohospedableBlogMapa del sitio
ProyectoAcerca deCómo clasificamosPrensaServidor MCP
Aviso legalPrivacidadTérminos
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
N

nuprl/MultiPL-E

0
View on GitHub↗
0 estrellas·0 forks·4 vistas

MultiPL E

MultiPL-E is a system for translating unit test-driven neural code generation benchmarks to new languages. We have used MultiPL-E to translate two popular Python benchmarks (HumanEval and MBPP) to 18 other programming languages.

Features

  • Benchmarks and Datasets - Scalable approach for benchmarking neural code generation across languages.

Historial de estrellas

Gráfico del historial de estrellas de nuprl/multipl-eGráfico del historial de estrellas de nuprl/multipl-e

Búsqueda con IA

Explora más repositorios increíbles

Describe lo que necesitas en lenguaje sencillo: la IA clasifica miles de proyectos open-source curados por relevancia.

Start searching with AI

Alternativas open-source a MultiPL E

Proyectos open-source similares, clasificados según cuántas características comparten con MultiPL E.
  • rllm-org/rllmAvatar de rllm-org

    rllm-org/rllm

    5,641Ver en GitHub↗

    rllm is an asynchronous reinforcement learning framework for training language agents. It provides a unified pipeline that runs the same agent code for both evaluation and training, automatically capturing traces for gradient computation. The framework supports distributed reinforcement learning across multiple GPUs and nodes using pluggable backends, and executes agents in isolated sandboxes—either locally or in the cloud—for safe and scalable rollout collection. It trains agents built with LangGraph, SmolAgents, OpenAI Agents SDK, or custom frameworks without requiring core logic changes. T

    Pythonagent-frameworkagentic-workflowcoding-agent
    Ver en GitHub↗5,641
  • evalplus/evalplusAvatar de evalplus

    evalplus/evalplus

    1,765Ver en GitHub↗

    Rigourous evaluation of LLM-synthesized code - NeurIPS 2023 & COLM 2024

    Python
    Ver en GitHub↗1,765
  • gersteinlab/biocoderAvatar de gersteinlab

    gersteinlab/BioCoder

    58Ver en GitHub↗

    BioCoder is a challenging bioinformatics code generation benchmark for examining the capabilities of state-of-the-art large language models (LLMs).

    Jupyter Notebook
    Ver en GitHub↗58
  • bytedance/fullstackbenchAvatar de bytedance

    bytedance/FullStackBench

    121Ver en GitHub↗

    FullStack Bench: Evaluating LLMs as Full Stack Coders

    Python
    Ver en GitHub↗121
Ver las 15 alternativas a MultiPL E→

Preguntas frecuentes

¿Qué hace nuprl/multipl-e?

MultiPL-E is a system for translating unit test-driven neural code generation benchmarks to new languages. We have used MultiPL-E to translate two popular Python benchmarks (HumanEval and MBPP) to 18 other programming languages.

¿Cuáles son las características principales de nuprl/multipl-e?

Las características principales de nuprl/multipl-e son: Benchmarks and Datasets.

¿Qué alternativas de código abierto existen para nuprl/multipl-e?

Las alternativas de código abierto para nuprl/multipl-e incluyen: rllm-org/rllm — rllm is an asynchronous reinforcement learning framework for training language agents. It provides a unified pipeline… evalplus/evalplus — Rigourous evaluation of LLM-synthesized code - NeurIPS 2023 & COLM 2024. gersteinlab/biocoder — BioCoder is a challenging bioinformatics code generation benchmark for examining the capabilities of state-of-the-art… google-research/google-research — This repository serves as a comprehensive research platform and toolkit for advancing machine learning, quantum… hendrycks/apps — This is the repository for Measuring Coding Challenge Competence With APPS by Dan Hendrycks\, Steven Basart\, Saurav… bytedance/fullstackbench — FullStack Bench: Evaluating LLMs as Full Stack Coders.