Exact model identity
Claude Code (Sonnet 4.6)
Anthropic · version status: reported
Similar model names are never merged automatically. Evaluation membership and numeric result rows only reference the exact ID
claude-code-sonnet-4-6.Evaluation settings
| Benchmark | Work | Run / group | Scope |
|---|---|---|---|
| CompBioBench | Agentic systems are adept at solving well-scoped, verifiable problems in computational biology | compbiobench-sonnet-full compbiobench-v1-sonnet-high-one-run | full · n=100 |
| CompBioBench | Agentic systems are adept at solving well-scoped, verifiable problems in computational biology | compbiobench-sonnet-hardest compbiobench-v1-hardest-sonnet-high-one-run | subset · n=17 |
Numeric results
| Benchmark | Work | Run / group | Metric | Value |
|---|---|---|---|---|
| CompBioBench | Agentic systems are adept at solving well-scoped, verifiable problems in computational biology | compbiobench-sonnet-full compbiobench-v1-sonnet-high-one-run | Accuracy | 70 percent |
| CompBioBench | Agentic systems are adept at solving well-scoped, verifiable problems in computational biology | compbiobench-sonnet-full compbiobench-v1-sonnet-high-one-run | Wall-clock time per question | 1049.2 seconds |
| CompBioBench | Agentic systems are adept at solving well-scoped, verifiable problems in computational biology | compbiobench-sonnet-full compbiobench-v1-sonnet-high-one-run | Cost per question | 1.2 USD |
| CompBioBench | Agentic systems are adept at solving well-scoped, verifiable problems in computational biology | compbiobench-sonnet-hardest compbiobench-v1-hardest-sonnet-high-one-run | Accuracy — difficulty Levels 4–5 | 53 percent |