Scientific task · Systems modeling

Reaction-network reconstruction

Reconstruct a biochemical reaction network or executable systems model.

反应网络重建

Experimental systemSystems modeling

Definition and search aliases

Permanent ID
reaction-network-reconstruction
Aliases
SBML reconstruction, pathway model reconstruction
Deprecated aliases
None
Hierarchy
Leaf task under Scientific workflow and systems analysis

Coverage

1 benchmark families cover this task

suitecomplete

SCIGYM

An agentic systems-biology suite in which language models iteratively perturb simulated SBML systems, analyze time-series observations in Python, and reconstruct hidden biological reactions.

Reaction-network reconstruction
trackcomplete

SCIGYM Large

The formally released SCIGYM track containing the 213 systems not included in the creator paper's model evaluation, with systems reaching up to 400 reactions.

Reaction-network reconstruction
trackcomplete

SCIGYM Small

The formally released and creator-evaluated SCIGYM track containing biological systems with fewer than ten reactions.

Reaction-network reconstruction

Evidence-backed count claims

Each row keeps its original unit and basis. Rows with different units or overlapping mappings are never added.

BenchmarkMapped taskCoverageCountVersionEvidence
SCIGYM
root: scigym
Reaction-network reconstruction
official-taxonomy · high
explicitly-in-scope350 systems
distinct curated BioModels systems released as SBML benchmark instances
2025 releasescigym-evidence-release-counts
scigym-evidence-taxonomy
SCIGYM Large
root: scigym
Reaction-network reconstruction
official-track · high
explicitly-in-scope213 systems
unique SBML systems in the official large Parquet split, containing the remaining systems with up to 400 reactions
2025 releasescigym-large-evidence-count
SCIGYM Small
root: scigym
Reaction-network reconstruction
official-track · high
explicitly-in-scope137 systems
unique SBML systems with fewer than 10 reactions in the official small Parquet split
2025 releasescigym-small-evidence-count

Official evaluations connected to these benchmarks

Runs are included only for benchmark records mapped here (and formal child tracks when a mapped suite is the root). A task mapping does not imply that every run isolates this task.

WorkProvider / classRelated runs
Measuring Scientific Capabilities of Language Models with a Systems Biology Dry LabUniversity of Toronto, SickKids, Axiom, Mila, Vector Institute
benchmark_creator
scigym-small-creator-paper
scigym-small-zero-shot