How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.
Go to https://mimic.physionet.org/ for access. Once you have the authority for the dataset, download the dataset at the data folder under the same directory as the this repository.
MedAgentsBench: Benchmarking Thinking Models and Agent Frameworks for Complex Medical Reasoning
C linical H ealthcare I n-Situ Environment Benchmark for long-horizon, policy-rich healthcare workflow agents
The code to visualize colon-bench (MICCAI 2026) data and run evaluations of MLLMs on the benchmark.
The main features of ajhamdi/colon-bench-eval are: Healthcare Agent Benchmarks, Medical Imaging Agents.
Projects with overlapping indexed features include: actava-ai/chi-bench — C linical H ealthcare I n-Situ Environment Benchmark for long-horizon, policy-rich healthcare workflow agents. clibench/clibench — Go to https://mimic.physionet.org/ for access. Once you have the authority for the dataset, download the dataset at… cuhk-aim-group/medsam-agent — 🤖 Model | 🤗 Dataset | 📖 Paper. gersteinlab/medagents-benchmark — MedAgentsBench: Benchmarking Thinking Models and Agent Frameworks for Complex Medical Reasoning. imyangc7/lungnoduleagent — Classifies lung adenocarcinoma nodules into AIS / MIA / IAC through a four-stage pipeline — nodule detection, sizing,… rajpurkarlab/rex-mle — A medical machine learning benchmark platform for evaluating automated machine learning agents on realistic healthcare…