awesome-repositories.com
Blog
MCP
awesome-repositories.com

Entdecke die besten Open-Source-Repositories mit KI-gestützter Suche.

EntdeckenKuratierte SuchenOpen-Source-AlternativenSelf-hosted SoftwareBlogSitemap
ProjektÜber unsRanking-MethodikPresseMCP-Server
RechtlichesDatenschutzAGB
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
bytedance avatar

bytedance/FullStackBench

0
View on GitHub↗
121 Stars·9 Forks·Python·Apache-2.0·4 Aufrufe

FullStackBench

FullStack Bench: Evaluating LLMs as Full Stack Coders

Features

  • Benchmarks and Datasets - Benchmark for evaluating LLMs on full-stack software development tasks.

Star-Verlauf

Star-Verlauf für bytedance/fullstackbenchStar-Verlauf für bytedance/fullstackbench

KI-Suche

Entdecke weitere awesome Repositories

Beschreibe in einfachen Worten, was du brauchst — die KI bewertet tausende kuratierte Open-Source-Projekte nach Relevanz.

Start searching with AI

Open-Source-Alternativen zu FullStackBench

Ähnliche Open-Source-Projekte, sortiert nach der Anzahl der gemeinsamen Funktionen mit FullStackBench.
  • rllm-org/rllmAvatar von rllm-org

    rllm-org/rllm

    5,641Auf GitHub ansehen↗

    rllm is an asynchronous reinforcement learning framework for training language agents. It provides a unified pipeline that runs the same agent code for both evaluation and training, automatically capturing traces for gradient computation. The framework supports distributed reinforcement learning across multiple GPUs and nodes using pluggable backends, and executes agents in isolated sandboxes—either locally or in the cloud—for safe and scalable rollout collection. It trains agents built with LangGraph, SmolAgents, OpenAI Agents SDK, or custom frameworks without requiring core logic changes. T

    Pythonagent-frameworkagentic-workflowcoding-agent
    Auf GitHub ansehen↗5,641
  • gersteinlab/biocoderAvatar von gersteinlab

    gersteinlab/BioCoder

    58Auf GitHub ansehen↗

    BioCoder is a challenging bioinformatics code generation benchmark for examining the capabilities of state-of-the-art large language models (LLMs).

    Jupyter Notebook
    Auf GitHub ansehen↗58
  • google-research/google-researchAvatar von google-research

    google-research/google-research

    38,139Auf GitHub ansehen↗

    This repository serves as a comprehensive research platform and toolkit for advancing machine learning, quantum computing, and large-scale scientific data analysis. It provides foundational frameworks for developing complex algorithmic systems, offering the necessary infrastructure for distributed training, computational graph execution, and high-performance model development. The project distinguishes itself by integrating specialized research domains with robust, privacy-preserving methodologies. It supports diverse scientific discovery through tools for quantum simulation, physics-informed

    Jupyter Notebookaimachine-learningresearch
    Auf GitHub ansehen↗38,139
  • evalplus/evalplusAvatar von evalplus

    evalplus/evalplus

    1,765Auf GitHub ansehen↗

    Rigourous evaluation of LLM-synthesized code - NeurIPS 2023 & COLM 2024

    Python
    Auf GitHub ansehen↗1,765
Alle 15 Alternativen zu FullStackBench anzeigen→

Häufig gestellte Fragen

Was macht bytedance/fullstackbench?

FullStack Bench: Evaluating LLMs as Full Stack Coders

Was sind die Hauptfunktionen von bytedance/fullstackbench?

Die Hauptfunktionen von bytedance/fullstackbench sind: Benchmarks and Datasets.

Welche Open-Source-Alternativen gibt es zu bytedance/fullstackbench?

Open-Source-Alternativen zu bytedance/fullstackbench sind unter anderem: rllm-org/rllm — rllm is an asynchronous reinforcement learning framework for training language agents. It provides a unified pipeline… gersteinlab/biocoder — BioCoder is a challenging bioinformatics code generation benchmark for examining the capabilities of state-of-the-art… google-research/google-research — This repository serves as a comprehensive research platform and toolkit for advancing machine learning, quantum… hendrycks/apps — This is the repository for Measuring Coding Challenge Competence With APPS by Dan Hendrycks\, Steven Basart\, Saurav… leolty/repobench — ✨ RepoBench: Benchmarking Repository-Level Code Auto-Completion Systems - ICLR 2024. evalplus/evalplus — Rigourous evaluation of LLM-synthesized code - NeurIPS 2023 & COLM 2024.