Exact model identity

Claude Opus 4.5

Anthropic · version status: reported

Similar model names are never merged automatically. Evaluation membership and numeric result rows only reference the exact ID claude-opus-4-5.

Evaluation settings

Numeric results

BenchmarkWorkRun / groupMetricValue
Anthropic Computational Biology EvalAdvancing Claude in healthcare and the life sciencesanthropic-computational-biology-delta
anthropic-computational-biology-delta
Accuracy improvement from Claude Opus 4.1 to Claude Opus 4.5
delta vs claude-opus-4-1
Δ 10.5 percent delta as annotated
Anthropic Protein Understanding EvalAdvancing Claude in healthcare and the life sciencesanthropic-protein-understanding-delta
anthropic-protein-understanding-delta
Accuracy improvement from Claude Opus 4.1 to Claude Opus 4.5
delta vs claude-opus-4-1
Δ 10.3 percent delta as annotated
Anthropic Scientific Figure Interpretation EvalAdvancing Claude in healthcare and the life sciencesanthropic-scientific-figure-delta
anthropic-scientific-figure-delta
Accuracy improvement from Claude Opus 4.1 to Claude Opus 4.5
delta vs claude-opus-4-1
Δ 13.2 percent delta as annotated
SpatialBenchSpatialBench: Can Agents Analyze Real-World Spatial Biology Data?spatialbench-paper-v2-base
spatialbench-paper-v2-base
Accuracy38.36 percent
SpatialBenchSpatialBench: Can Agents Analyze Real-World Spatial Biology Data?spatialbench-paper-v2-base
spatialbench-paper-v2-base
Steps2.84 steps per evaluation
SpatialBenchSpatialBench: Can Agents Analyze Real-World Spatial Biology Data?spatialbench-paper-v2-base
spatialbench-paper-v2-base
Latency123.8 seconds per evaluation
SpatialBenchSpatialBench: Can Agents Analyze Real-World Spatial Biology Data?spatialbench-paper-v2-base
spatialbench-paper-v2-base
Cost0.143 USD per evaluation
SpatialBenchSpatialBench: Can Agents Analyze Real-World Spatial Biology Data?spatialbench-paper-v2-claude-code
spatialbench-paper-v2-claude-code
Accuracy48.1 percent
SpatialBenchSpatialBench: Can Agents Analyze Real-World Spatial Biology Data?spatialbench-paper-v2-latch
spatialbench-paper-v2-latch
Accuracy61.7 percent
SpatialBenchSpatialBench 159-evaluation repository snapshotspatialbench-repo-159-mini-swe-agent
spatialbench-repo-159-mini-swe-agent
Accuracy42.77 percent
SpatialBenchSpatialBench 159-evaluation repository snapshotspatialbench-repo-159-mini-swe-agent
spatialbench-repo-159-mini-swe-agent
Cost0.4624 USD per evaluation
SpatialBenchSpatialBench 159-evaluation repository snapshotspatialbench-repo-159-mini-swe-agent
spatialbench-repo-159-mini-swe-agent
Duration376.54 seconds per evaluation