Scopefull · n=5
ShotsNot applicable
TurnsNot applicable
System prompt publicNot applicable
Reasoning / effortNot applicable
BrowserNot applicable
InternetNot applicable
DatabasesNot applicable
Code executionNot applicable
ContainerNot reported
External toolstask-specific supervised heads over frozen or fine-tuned protein representations
Token budgetNot applicable
Time / cost budgetNot reported
TemperatureNot applicable
SeedNot reported
RepeatsNot reported
Graderdeterministic task-specific scorer · human review: no
StatisticsTask-level evaluation only; no cross-task normalized aggregate.
ContaminationSequence-identity filtering and biologically motivated held-out splits.
Metrics, results, and full protocol
Metrics
| Metric | Kind / baseline | Unit | Aggregation | Threshold / tolerance |
|---|---|---|---|---|
| Per-amino-acid accuracy | absolute | proportion | across labeled residues | Not reported |
| L/5 medium and long-range precision | absolute | proportion | per protein then task summary | Not reported |
| Fold-level accuracy | absolute | proportion | held-out test examples | Not reported |
| Spearman's rho | absolute | correlation | held-out test examples | Not reported |
No numeric result rows are published yet; the verified protocol remains useful.
Evidence
- section: Sections 4.2-5 and Table 2; Appendix A (Defines all five task splits, architectures, training procedures, native metrics, and creator comparison table.) — supports /scope, /protocol, /metrics