awesome-repositories.com
Blog
MCP
awesome-repositories.com

Entdecke die besten Open-Source-Repositories mit KI-gestützter Suche.

EntdeckenKuratierte SuchenOpen-Source-AlternativenSelf-hosted SoftwareBlogSitemap
ProjektÜber unsRanking-MethodikPresseMCP-Server
RechtlichesDatenschutzAGB
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
N

nuprl/MultiPL-E

0
View on GitHub↗
0 Stars·0 Forks·1 Aufruf

MultiPL E

MultiPL-E is a system for translating unit test-driven neural code generation benchmarks to new languages. We have used MultiPL-E to translate two popular Python benchmarks (HumanEval and MBPP) to 18 other programming languages.

Features

  • Benchmarks and Datasets - Scalable approach for benchmarking neural code generation across languages.

Star-Verlauf

Star-Verlauf für nuprl/multipl-eStar-Verlauf für nuprl/multipl-e

KI-Suche

Entdecke weitere awesome Repositories

Beschreibe in einfachen Worten, was du brauchst — die KI bewertet tausende kuratierte Open-Source-Projekte nach Relevanz.

Start searching with AI

Open-Source-Alternativen zu MultiPL E

Ähnliche Open-Source-Projekte, sortiert nach der Anzahl der gemeinsamen Funktionen mit MultiPL E.
  • rllm-org/rllmAvatar von rllm-org

    rllm-org/rllm

    5,641Auf GitHub ansehen↗

    rllm is an asynchronous reinforcement learning framework for training language agents. It provides a unified pipeline that runs the same agent code for both evaluation and training, automatically capturing traces for gradient computation. The framework supports distributed reinforcement learning across multiple GPUs and nodes using pluggable backends, and executes agents in isolated sandboxes—either locally or in the cloud—for safe and scalable rollout collection. It trains agents built with LangGraph, SmolAgents, OpenAI Agents SDK, or custom frameworks without requiring core logic changes. T

    Pythonagent-frameworkagentic-workflowcoding-agent
    Auf GitHub ansehen↗5,641
  • evalplus/evalplusAvatar von evalplus

    evalplus/evalplus

    1,765Auf GitHub ansehen↗

    Rigourous evaluation of LLM-synthesized code - NeurIPS 2023 & COLM 2024

    Python
    Auf GitHub ansehen↗1,765
  • gersteinlab/biocoderAvatar von gersteinlab

    gersteinlab/BioCoder

    58Auf GitHub ansehen↗

    BioCoder is a challenging bioinformatics code generation benchmark for examining the capabilities of state-of-the-art large language models (LLMs).

    Jupyter Notebook
    Auf GitHub ansehen↗58
  • bytedance/fullstackbenchAvatar von bytedance

    bytedance/FullStackBench

    121Auf GitHub ansehen↗

    FullStack Bench: Evaluating LLMs as Full Stack Coders

    Python
    Auf GitHub ansehen↗121
Alle 15 Alternativen zu MultiPL E anzeigen→

Häufig gestellte Fragen

Was macht nuprl/multipl-e?

MultiPL-E is a system for translating unit test-driven neural code generation benchmarks to new languages. We have used MultiPL-E to translate two popular Python benchmarks (HumanEval and MBPP) to 18 other programming languages.

Was sind die Hauptfunktionen von nuprl/multipl-e?

Die Hauptfunktionen von nuprl/multipl-e sind: Benchmarks and Datasets.

Welche Open-Source-Alternativen gibt es zu nuprl/multipl-e?

Open-Source-Alternativen zu nuprl/multipl-e sind unter anderem: rllm-org/rllm — rllm is an asynchronous reinforcement learning framework for training language agents. It provides a unified pipeline… evalplus/evalplus — Rigourous evaluation of LLM-synthesized code - NeurIPS 2023 & COLM 2024. gersteinlab/biocoder — BioCoder is a challenging bioinformatics code generation benchmark for examining the capabilities of state-of-the-art… google-research/google-research — This repository serves as a comprehensive research platform and toolkit for advancing machine learning, quantum… hendrycks/apps — This is the repository for Measuring Coding Challenge Competence With APPS by Dan Hendrycks\, Steven Basart\, Saurav… bytedance/fullstackbench — FullStack Bench: Evaluating LLMs as Full Stack Coders.