awesome-repositories.com
ब्लॉग
MCP
awesome-repositories.com

AI-संचालित खोज के साथ बेहतरीन ओपन-सोर्स रिपॉजिटरी खोजें।

एक्सप्लोर करेंक्यूरेटेड खोजेंओपन-सोर्स विकल्पसेल्फ-होस्टेड सॉफ्टवेयरब्लॉगसाइटमैप
प्रोजेक्टMCP सर्वरहमारे बारे मेंहम रैंकिंग कैसे करते हैंप्रेस
कानूनीगोपनीयताशर्तें
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

2 रिपॉजिटरी

Awesome GitHub RepositoriesAutomation Capability Benchmarks

Standardized benchmarks specifically designed to measure the automation efficiency of AI models.

Distinct from Model Benchmarks: Focuses on the ability to automate complex tasks rather than static model performance or pricing

Explore 2 awesome GitHub repositories matching artificial intelligence & ml · Automation Capability Benchmarks. Refine with filters or upvote what's useful.

Awesome Automation Capability Benchmarks GitHub Repositories

AI के साथ बेहतरीन रिपॉजिटरी खोजें।हम AI का उपयोग करके सबसे सटीक रिपॉजिटरी खोजेंगे।
  • microsoft/jarvismicrosoft का अवतार

    microsoft/JARVIS

    24,854GitHub पर देखें↗

    JARVIS is a system for large language model task orchestration, deployment management, and automation benchmarking. It utilizes a task orchestrator to decompose complex requests into actionable steps and coordinates various expert models to synthesize final responses. The project includes an AI model deployment manager to handle the local deployment of expert models across different hardware scales. It further provides an AI workflow API consisting of web endpoints used to trigger automated task workflows and retrieve results from model selection stages. The framework incorporates an automat

    Evaluates the capability of large language models to automate complex tasks using standardized benchmarking datasets.

    Python
    GitHub पर देखें↗24,854
  • orchestra-research/ai-research-skillsOrchestra-Research का अवतार

    Orchestra-Research/AI-Research-SKILLs

    3,641GitHub पर देखें↗

    This project is an LLM research orchestrator and autonomous AI agent framework designed to automate the scientific lifecycle. It functions as an end-to-end research pipeline and model training toolkit, managing everything from initial literature reviews and hypothesis testing to the final drafting of academic papers. The system is distinguished by its ability to convert unstructured academic PDFs into machine-executable knowledge layers, allowing agents to reproduce and extend research findings. It employs a two-loop orchestration architecture and a specialized research engineering skill libr

    Evaluates the ability of AI systems to autonomously design and analyze scientific experiments with rigor.

    TeXaiai-researchclaude
    GitHub पर देखें↗3,641
  1. Home
  2. Artificial Intelligence & ML
  3. Large Language Models
  4. Model Benchmarks
  5. Automation Capability Benchmarks

सब-टैग एक्सप्लोर करें

  • Scientific Experimentation BenchmarksStandardized evaluations of an AI's ability to design and analyze scientific experiments. **Distinct from Automation Capability Benchmarks:** Specializes automation benchmarks to the scientific method and research rigor