This table records what the work did with each benchmark before attempting to normalize a run. Partial claims remain visible without being treated as comparable evaluations.
benchmark creation
abbibench-a-benchmark-for-antibody-binding-affinity-ma-abbibench-1-use
Non-evaluationunknown
Benchmark: AbBiBench · version Not reported
- Selection
- not applicable
- Models
- Not reported / not applicable
- Metrics
- Not reported / not applicable
- Linked runs
- None
Not reported / unresolved: A benchmark version and artifact release date are not reported.; The source conflicts on whether AbBiBench contains 14 or 16 binding-affinity assays.
AI-assisted double-pass extraction; values are limited to independently supported claims.
Evidence
- section: Introduction
Supports: /relation_type - section: Introduction
Supports: /benchmark_id
evaluation
abbibench-a-benchmark-for-antibody-binding-affinity-ma-abbibench-2-use
Partialunknown
Benchmark: AbBiBench · version Not reported
- Selection
- not reported
- Models
- ESM-IF1, AntiBERTy, AntiFold, CurrAb, DiffAb, DiffAb fixbb, dyMEAN, dyMEAN fixbb, ESM2, ESM3, MEAN, MEAN fixbb, ProGen2-large, ProSST, SaProt, ProteinMPNN, ProtGPT2
- Metrics
- Not reported / not applicable
- Linked runs
- None
Not reported / unresolved: Exact model versions, providers, seeds, repeats, and confidence intervals are not reported.; The source conflicts on light-chain inclusion and on the 1mlc evaluation count.; Exact model and tool versions, compute time, and confidence intervals are not reported.; DiffAb seeds are described only as reaching up to 15; the exact seed list is absent.; The numeric ELISA detection threshold is not reported.; The final Pareto-candidate count conflicts between 18 and 21.; Exact model versions, seeds, repeats, and exact p-values are not reported.; benchmark version; realized n/scope; metric; numeric result; prompt and tools; grader and repeats
Owner-reviewed conservative publication: the creator evaluation is retained only as a partial relationship; conflicted settings and outcomes are omitted pending manual reconciliation.
Evidence
- section: Section 3
Supports: /relation_type - section: Section 3
Supports: /benchmark_id - table: Table S2
Supports: /model_ids - table: Table S2
Supports: /model_ids - table: Table S2
Supports: /model_ids - table: Table S2
Supports: /model_ids - table: Table S2
Supports: /model_ids - table: Table S2
Supports: /model_ids - figure: Figure 3
Supports: /model_ids - table: Table S2
Supports: /model_ids - other: Table S2 and registry-context.json model proteinmpnn
Supports: /model_ids - figure: Figure 3
Supports: /model_ids - table: Table S2
Supports: /model_ids - table: Table S2
Supports: /model_ids - table: Table S2
Supports: /model_ids - table: Table S2
Supports: /model_ids - table: Table S2
Supports: /model_ids - table: Table S2
Supports: /model_ids - table: Table S2
Supports: /model_ids