awesome-repositories.com
المدونة
MCP
awesome-repositories.com

اكتشف أفضل مستودعات المصادر المفتوحة باستخدام بحث مدعوم بالذكاء الاصطناعي.

استكشفعمليات بحث منسقةبدائل مفتوحة المصدربرمجيات ذاتية الاستضافةالمدونةخريطة الموقع
المشروعخادم MCPحولكيفية ترتيب النتائجالصحافة
قانونيالخصوصيةالشروط
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

4 مستودعات

Awesome GitHub RepositoriesScientific Model Evaluators

Benchmarking tools for predictive models using expert-developed baselines and scientific datasets.

Distinct from Model Benchmarking: Distinct from general model benchmarking: focuses on scientific domain predictive models.

Explore 4 awesome GitHub repositories matching artificial intelligence & ml · Scientific Model Evaluators. Refine with filters or upvote what's useful.

Awesome Scientific Model Evaluators GitHub Repositories

اعثر على أفضل المستودعات باستخدام الذكاء الاصطناعي.سنبحث عن أفضل المستودعات المطابقة باستخدام الذكاء الاصطناعي.
  • google-research/google-researchالصورة الرمزية لـ google-research

    google-research/google-research

    38,139عرض على GitHub↗

    This repository serves as a comprehensive research platform and toolkit for advancing machine learning, quantum computing, and large-scale scientific data analysis. It provides foundational frameworks for developing complex algorithmic systems, offering the necessary infrastructure for distributed training, computational graph execution, and high-performance model development. The project distinguishes itself by integrating specialized research domains with robust, privacy-preserving methodologies. It supports diverse scientific discovery through tools for quantum simulation, physics-informed

    Evaluates predictive models across diverse domains by comparing results against established datasets.

    Jupyter Notebookaimachine-learningresearch
    عرض على GitHub↗38,139
  • pytorch/captumالصورة الرمزية لـ pytorch

    pytorch/captum

    5,652عرض على GitHub↗

    Captum is an open-source library for explaining model predictions by attributing them to input features, neurons, and layers using gradient-based and perturbation-based methods. It provides a modular framework for implementing, evaluating, and combining a range of explanation techniques, including gradient-based attribution, perturbation-based analysis, game-theoretic Shapley value approximation, and surrogate model explanations, with support for parallelization and noise stabilization. The library distinguishes itself through its breadth of attribution methods and its support for advanced in

    Ships tools to assess attribution reliability through sensitivity and consistency tests.

    Python
    عرض على GitHub↗5,652
  • dwzhu-pku/paperbananaالصورة الرمزية لـ dwzhu-pku

    dwzhu-pku/PaperBanana

    3,742عرض على GitHub↗

    PaperBanana is an AI research visualization tool and framework designed to generate and refine high-resolution academic illustrations from conceptual and technical descriptions. It employs an automated generation pipeline that transforms scientific text and captions into publication-quality diagrams and plots. The system utilizes a multi-stage process consisting of retrieval-augmented planning, image synthesis, and a critic-based iterative refinement mechanism. This workflow allows for the adjustment of image details and the upscaling of visual outputs to 4K resolution. The project includes

    Provides a set of metrics and tools for measuring the quality of AI-generated academic illustrations against ground-truth datasets.

    JavaScript
    عرض على GitHub↗3,742
  • stanfordnmbl/osim-rlالصورة الرمزية لـ stanfordnmbl

    stanfordnmbl/osim-rl

    944عرض على GitHub↗

    Osim-rl is a research environment designed for the development and evaluation of reinforcement learning agents within physics-based musculoskeletal simulations. It provides a standardized interface that maps physiological state observations to muscle excitation control signals, enabling the study of human movement and biomechanics through iterative policy optimization. The framework distinguishes itself by integrating high-fidelity musculoskeletal modeling with tools for scientific benchmarking and reproducible experimentation. It allows researchers to define custom reward functions and adjus

    Facilitates objective comparison of control policies against standardized metrics within a consistent and reproducible simulation framework.

    Pythonbiomechanicsdeep-reinforcement-learningkinematics
    عرض على GitHub↗944
  1. Home
  2. Artificial Intelligence & ML
  3. Machine Learning
  4. Infrastructure
  5. Evaluation & Validation
  6. Model Benchmarking
  7. Scientific Model Evaluators

استكشف الوسوم الفرعية

  • Illustration Quality Evaluators1 وسم فرعيBenchmarking tools that measure the quality of generated academic visuals against ground-truth scientific datasets. **Distinct from Scientific Model Evaluators:** Distinct from Scientific Model Evaluators by focusing on the visual output quality of illustrations rather than predictive model accuracy.