preprint · benchmark creator

AbBiBench: A Benchmark for Antibody Binding Affinity Maturation and Design

McWilliams School of Biomedical Informatics, UTHealth Houston · Department of Industrial and Systems Engineering, Korea Advanced Institute of Science and Technology · Texas Therapeutics Institute, Brown Foundation Institute of Molecular Medicine, UTHealth Houston · 2025-05-23

Relationship layer

Benchmark usage

This table records what the work did with each benchmark before attempting to normalize a run. Partial claims remain visible without being treated as comparable evaluations.

benchmark creation

abbibench-a-benchmark-for-antibody-binding-affinity-ma-abbibench-1-use

Non-evaluationunknown

Benchmark: AbBiBench · version Not reported

Selection
not applicable
Models
Not reported / not applicable
Metrics
Not reported / not applicable
Linked runs
None

Not reported / unresolved: A benchmark version and artifact release date are not reported.; The source conflicts on whether AbBiBench contains 14 or 16 binding-affinity assays.

AI-assisted double-pass extraction; values are limited to independently supported claims.

Evidence
  • section: Introduction
    Supports: /relation_type
  • section: Introduction
    Supports: /benchmark_id

evaluation

abbibench-a-benchmark-for-antibody-binding-affinity-ma-abbibench-2-use

Partialunknown

Benchmark: AbBiBench · version Not reported

Selection
not reported
Metrics
Not reported / not applicable
Linked runs
None

Not reported / unresolved: Exact model versions, providers, seeds, repeats, and confidence intervals are not reported.; The source conflicts on light-chain inclusion and on the 1mlc evaluation count.; Exact model and tool versions, compute time, and confidence intervals are not reported.; DiffAb seeds are described only as reaching up to 15; the exact seed list is absent.; The numeric ELISA detection threshold is not reported.; The final Pareto-candidate count conflicts between 18 and 21.; Exact model versions, seeds, repeats, and exact p-values are not reported.; benchmark version; realized n/scope; metric; numeric result; prompt and tools; grader and repeats

Owner-reviewed conservative publication: the creator evaluation is retained only as a partial relationship; conflicted settings and outcomes are omitted pending manual reconciliation.

Evidence
  • section: Section 3
    Supports: /relation_type
  • section: Section 3
    Supports: /benchmark_id
  • table: Table S2
    Supports: /model_ids
  • table: Table S2
    Supports: /model_ids
  • table: Table S2
    Supports: /model_ids
  • table: Table S2
    Supports: /model_ids
  • table: Table S2
    Supports: /model_ids
  • table: Table S2
    Supports: /model_ids
  • figure: Figure 3
    Supports: /model_ids
  • table: Table S2
    Supports: /model_ids
  • other: Table S2 and registry-context.json model proteinmpnn
    Supports: /model_ids
  • figure: Figure 3
    Supports: /model_ids
  • table: Table S2
    Supports: /model_ids
  • table: Table S2
    Supports: /model_ids
  • table: Table S2
    Supports: /model_ids
  • table: Table S2
    Supports: /model_ids
  • table: Table S2
    Supports: /model_ids
  • table: Table S2
    Supports: /model_ids
  • table: Table S2
    Supports: /model_ids

Normalized evaluation runs

This source has no normalized model run. It may be a creator-only source or a partial/non-evaluation benchmark use.