Scopefull · n=8
ShotsNot applicable
TurnsNot applicable
System prompt publicNot applicable
Reasoning / effortNot applicable
BrowserNot applicable
InternetNot applicable
DatabasesNot applicable
Code executionNot applicable
ContainerNot reported
External toolstask-specific 3D CNN, graph neural network, and equivariant neural network pipelines
Token budgetNot applicable
Time / cost budgetNot reported
TemperatureNot applicable
SeedNot reported
Repeats3
Graderdeterministic task-specific scorer · human review: no
StatisticsNative per-task metrics with standard deviations over three replicates; structure-ranking correlations are computed per target before summary.
ContaminationTask-specific sequence-identity, time, target, or scaffold splits.
Metrics, results, and full protocol
Metrics
| Metric | Kind / baseline | Unit | Aggregation | Threshold / tolerance |
|---|---|---|---|---|
| MAE | absolute | task-specific | held-out examples | Not reported |
| RMSE | absolute | task-specific | held-out examples | Not reported |
| AUROC | absolute | area | held-out examples | Not reported |
| Accuracy | absolute | proportion | held-out examples | Not reported |
| Pearson correlation | absolute | correlation | task-specific | Not reported |
| Spearman correlation | absolute | correlation | target-level structure ranking | Not reported |
No numeric result rows are published yet; the verified protocol remains useful.
Evidence
- table: Sections 2-5 and Tables 1-8; Appendix D (Defines all eight datasets, official splits, model classes, native metrics, three-replicate aggregation, and creator baseline results.) — supports /scope, /protocol, /metrics