awesome-repositories.com
Blog
MCP
awesome-repositories.com

Descubre los mejores repositorios open-source con nuestra búsqueda potenciada por IA.

ExplorarBúsquedas curadasAlternativas open-sourceSoftware autohospedableBlogMapa del sitio
ProyectoServidor MCPAcerca deCómo clasificamosPrensa
Aviso legalPrivacidadTérminos
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

2 repositorios

Awesome GitHub RepositoriesAutomation Capability Benchmarks

Standardized benchmarks specifically designed to measure the automation efficiency of AI models.

Distinct from Model Benchmarks: Focuses on the ability to automate complex tasks rather than static model performance or pricing

Explore 2 awesome GitHub repositories matching artificial intelligence & ml · Automation Capability Benchmarks. Refine with filters or upvote what's useful.

Awesome Automation Capability Benchmarks GitHub Repositories

Encuentra los mejores repositorios con IA.Buscaremos los repositorios que mejor coincidan usando IA.
  • microsoft/jarvisAvatar de microsoft

    microsoft/JARVIS

    24,854Ver en GitHub↗

    JARVIS is a system for large language model task orchestration, deployment management, and automation benchmarking. It utilizes a task orchestrator to decompose complex requests into actionable steps and coordinates various expert models to synthesize final responses. The project includes an AI model deployment manager to handle the local deployment of expert models across different hardware scales. It further provides an AI workflow API consisting of web endpoints used to trigger automated task workflows and retrieve results from model selection stages. The framework incorporates an automat

    Evaluates the capability of large language models to automate complex tasks using standardized benchmarking datasets.

    Python
    Ver en GitHub↗24,854
  • orchestra-research/ai-research-skillsAvatar de Orchestra-Research

    Orchestra-Research/AI-Research-SKILLs

    3,641Ver en GitHub↗

    This project is an LLM research orchestrator and autonomous AI agent framework designed to automate the scientific lifecycle. It functions as an end-to-end research pipeline and model training toolkit, managing everything from initial literature reviews and hypothesis testing to the final drafting of academic papers. The system is distinguished by its ability to convert unstructured academic PDFs into machine-executable knowledge layers, allowing agents to reproduce and extend research findings. It employs a two-loop orchestration architecture and a specialized research engineering skill libr

    Evaluates the ability of AI systems to autonomously design and analyze scientific experiments with rigor.

    TeXaiai-researchclaude
    Ver en GitHub↗3,641
  1. Home
  2. Artificial Intelligence & ML
  3. Large Language Models
  4. Model Benchmarks
  5. Automation Capability Benchmarks

Explorar subetiquetas

  • Scientific Experimentation BenchmarksStandardized evaluations of an AI's ability to design and analyze scientific experiments. **Distinct from Automation Capability Benchmarks:** Specializes automation benchmarks to the scientific method and research rigor