paper · benchmark creator

Genomic benchmarks: a collection of datasets for genomic sequence classification

Masaryk University · CEITEC Masaryk University · 2023-05-01

Relationship layer

Benchmark usage

This table records what the work did with each benchmark before attempting to normalize a run. Partial claims remain visible without being treated as comparable evaluations.

No BenchmarkUse relation is normalized for this legacy work yet. Existing EvaluationRuns remain available below.

Normalized evaluation runs

Genomic Benchmarks1 run

Open benchmark record →

genomic-benchmarks-creator-nativevpackage-1.0.0-snapshot
Scopefull · n=9
ShotsNot applicable
TurnsNot applicable
System prompt publicNot applicable
Reasoning / effortNot applicable
BrowserNot applicable
InternetNot applicable
DatabasesNot applicable
Code executionNot applicable
ContainerNot reported
External toolsTensorFlow and PyTorch CNN baseline workflows
Token budgetNot applicable
Time / cost budgetNot reported
TemperatureNot applicable
SeedNot reported
RepeatsNot reported
Graderdeterministic classification scorer · human review: no
StatisticsAccuracy and F1 are reported independently for each dataset and framework.
ContaminationDataset-specific train/test splits; duplicate and background-generation controls follow each construction notebook.
Metrics, results, and full protocol

Metrics

MetricKind / baselineUnitAggregationThreshold / tolerance
Accuracyabsolutepercentheld-out test sequences per datasetNot reported
F1 scoreabsolutepercentheld-out test sequences per datasetNot reported

No numeric result rows are published yet; the verified protocol remains useful.

Evidence

  • table: Methods and Tables 1-2 (Lists all nine datasets, train/test design, CNN workflows, and accuracy/F1 results.) — supports /scope, /protocol, /metrics