How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.
This project is a collection of specialized toolkits and an agent skill library designed to equip large language model agents with the capabilities to perform complex scientific research across biology, chemistry, medicine, and physics. It provides a structured framework of integration paths and tools that allow agents to execute multi-step research workflows. The system is distinguished by its domain-specific toolsets, including a bioinformatics toolkit for genomic and single-cell analysis, a cheminformatics toolset for drug-target binding and lead compound optimization, and a multi-omics an
Go to https://mimic.physionet.org/ for access. Once you have the authority for the dataset, download the dataset at the data folder under the same directory as the this repository.
MedAgentsBench: Benchmarking Thinking Models and Agent Frameworks for Complex Medical Reasoning
The code to visualize colon-bench (MICCAI 2026) data and run evaluations of MLLMs on the benchmark.
DeepTumorVQA benchmark for VLMs and Agents (10k testing samples)
The main features of schuture/deeptumorvqa are: Healthcare Agent Benchmarks, Medical Imaging Analysis.
Projects with overlapping indexed features include: k-dense-ai/scientific-agent-skills — This project is a collection of specialized toolkits and an agent skill library designed to equip large language model… clibench/clibench — Go to https://mimic.physionet.org/ for access. Once you have the authority for the dataset, download the dataset at… gersteinlab/medagents-benchmark — MedAgentsBench: Benchmarking Thinking Models and Agent Frameworks for Complex Medical Reasoning. ajhamdi/colon-bench-eval — The code to visualize colon-bench (MICCAI 2026) data and run evaluations of MLLMs on the benchmark. mengyunq/meshagents — Official implementation of the MICCAI 2025 paper. actava-ai/chi-bench — C linical H ealthcare I n-Situ Environment Benchmark for long-horizon, policy-rich healthcare workflow agents.