Exact model identity

grok-4.20-beta-0309-reasoning

xAI · version status: reported

Similar model names are never merged automatically. Evaluation membership and numeric result rows only reference the exact ID grok-4-20-beta-0309-reasoning.

Evaluation settings

Numeric results

BenchmarkWorkRun / groupMetricValue
BioSecBench-SurveillanceBioSecBench-Surveillance repository result snapshotbiosecbench-8d53fd8-pi
biosecbench-8d53fd8-pi
endpoint pass rate13.7 %
SpatialBenchSpatialBench 159-evaluation repository snapshotspatialbench-repo-159-mini-swe-agent
spatialbench-repo-159-mini-swe-agent
Accuracy45.91 percent
SpatialBenchSpatialBench 159-evaluation repository snapshotspatialbench-repo-159-mini-swe-agent
spatialbench-repo-159-mini-swe-agent
Cost0.1679 USD per evaluation
SpatialBenchSpatialBench 159-evaluation repository snapshotspatialbench-repo-159-mini-swe-agent
spatialbench-repo-159-mini-swe-agent
Duration342.63 seconds per evaluation