suite · audited-with-caveats · verified 2026-07-31
crafted experiments
Real single-cell RNA-seq data augmented with known gene perturbations for comparing feature-selection methods.
Audited with caveats: 2 field(s) are marked provisional or conflicted. Warnings are shown next to affected values and these claims are excluded from unqualified comparisons.
Benchmark definition
What is counted
- Version
- initial-release
- Total
- 24 (Explicitly reported generated inventory)
- Task formats
- unclassified
- Capabilities
Data analysis
Version history
| Version | Status | Release / as-of | Total | Formal tracks |
|---|
initial-release
crafted-experiments-initial-release-version | current | 2025-01-07 | 24 (Explicitly reported generated inventory) | None registered |
Scientific Task Atlas
Scientific task classification
partial for initial-release. No Scientific Task claim passed independent high-confidence verification; task mapping remains pending a targeted official-source audit.
The source names only a broad direction; no more specific leaf task can be assigned without inference.
Relationship registry
How works use this benchmark
Partial claims, non-evaluation uses, and third-party summaries stay visible without entering model comparisons.
Partial evaluation claims
evaluation
crafted-experiments-to-evaluate-feature-selection-meth-crafted-experiments-2-use
Partialunknown · n=24
Work: Crafted experiments to evaluate feature selection methods for single-cell RNA-seq data · source version crafted-experiments-to-evaluate-feature-selection-meth-2025-03-19
- Selection
- not reported · Guided tutorials and default parameters; top 2000 ranked features per method, except HIPPO used its zero-proportion-test cutoff
- Metrics
- MultiK low-resolution rank, DiProPerm Z scores, average rank, proportion of crafted genes selected, number of crafted genes selected, number of crafted genes not selected
- Linked runs
- None
Not reported / unresolved: Benchmark version is not reported.; Exact software versions and random seeds are not reported.; Figure 4 bar values are unlabeled, so no numeric model results were extracted.; Conflicted model claim omitted after independent verification; the evaluation relationship is published conservatively.; Conflicted repeats claim omitted after independent verification; the evaluation relationship is published conservatively.; benchmark version; numeric result
AI-assisted double-pass extraction; values are limited to independently supported claims.
Evidence
- section: Results — Crafted experiment applications
Supports: /relation_type - section: Results — Crafted experiment applications
Supports: /benchmark_id - section: Results — Crafted experiment applications
Supports: /scope - section: Results — Crafted experiment applications
Supports: /scope - section: Results — Crafted experiment applications
Supports: /scope - section: Materials and methods — Benchmarking of feature selection methods
Supports: /scope - section: Results — Crafted experiment applications; Reference 16
Supports: /model_ids - section: Results — Crafted experiment applications; Reference 15
Supports: /model_ids - section: Results — Crafted experiment applications; Reference 18
Supports: /model_ids - section: Results — Crafted experiment applications
Supports: /model_ids - other: Article author list; Results — Crafted experiment applications
Supports: /model_ids - other: Article author list; Results — Crafted experiment applications
Supports: /model_ids - other: Article author list; Results — Crafted experiment applications
Supports: /model_ids - section: Materials and methods — Benchmarking of feature selection methods
Supports: /metric_labels - section: Materials and methods — Benchmarking of feature selection methods
Supports: /metric_labels - section: Results — Crafted experiment applications
Supports: /metric_labels - section: Results — Crafted experiment applications
Supports: /metric_labels - section: Results — Crafted experiment applications
Supports: /metric_labels - section: Results — Crafted experiment applications
Supports: /metric_labels
Creation, training, validation, or model-selection uses
benchmark creation
crafted-experiments-to-evaluate-feature-selection-meth-crafted-experiments-1-use
Non-evaluationunknown
Work: Crafted experiments to evaluate feature selection methods for single-cell RNA-seq data · source version crafted-experiments-to-evaluate-feature-selection-meth-2025-03-19
- Selection
- not applicable
- Models
- Not reported / not applicable
- Metrics
- Not reported / not applicable
- Linked runs
- None
Not reported / unresolved: Benchmark version is not reported.; An explicit benchmark release date is not reported.; The released crafted-dataset artifact's license is not reported.
AI-assisted double-pass extraction; values are limited to independently supported claims.
Evidence
- section: Introduction
Supports: /relation_type - section: Introduction
Supports: /benchmark_id
Evaluation registry
Works and run settings
A setting change—scope, prompt, tools, budget, grader, or repeats—creates a separate run. Charts never cross a comparability group.
No normalized evaluation run is published yet. Creator evidence is still attached below.
Evidence and change history
Source locators remain visible; expand an item to inspect the exact Registry fields it supports.
Crafted experiments to evaluate feature selection methods for single-cell RNA-seq data · section: Data availability; Code availability · Supports 1 field
Open source →
Crafted experiments to evaluate feature selection methods for single-cell RNA-seq data · section: Data availability · Supports 1 field
Open source →
Crafted experiments to evaluate feature selection methods for single-cell RNA-seq data · section: Results — Crafted experiment applications · Supports 1 field
Open source →
Crafted experiments to evaluate feature selection methods for single-cell RNA-seq data · section: Introduction · Supports 1 field
Open source →
Crafted experiments to evaluate feature selection methods for single-cell RNA-seq data · section: Introduction · Supports 1 field
Open source →
Crafted experiments to evaluate feature selection methods for single-cell RNA-seq data · section: Results — Construction of crafted experiments · Supports 1 field
Open source →
Crafted experiments to evaluate feature selection methods for single-cell RNA-seq data · section: Results — Construction of crafted experiments · Supports 1 field
Open source →
Crafted experiments to evaluate feature selection methods for single-cell RNA-seq data · section: Abstract · Supports 1 field
Open source →
Crafted experiments to evaluate feature selection methods for single-cell RNA-seq data · section: Contributor Information · Supports 1 field
Open source →
Crafted experiments to evaluate feature selection methods for single-cell RNA-seq data · section: Introduction · Supports 1 field
Open source →
Crafted experiments to evaluate feature selection methods for single-cell RNA-seq data · other: Crossref bibliographic metadata (Resolved from the canonical paper identifier during intake.) · Supports 1 field
Open source →
Crafted experiments to evaluate feature selection methods for single-cell RNA-seq data · section: Results — Construction of crafted experiments · Supports 4 fields
Open source →
/task_counts/total/task_counts/basis/versions/0/task_counts/total/versions/0/task_counts/basis
Crafted experiments to evaluate feature selection methods for single-cell RNA-seq data · section: Results — Construction of crafted experiments · Supports 2 fields
Open source →
/task_counts/subsets/versions/0/task_counts/subsets
Crafted experiments to evaluate feature selection methods for single-cell RNA-seq data · section: Data availability · Supports 2 fields
Open source →
/resources/access/license
Crafted experiments to evaluate feature selection methods for single-cell RNA-seq data · other: Article DOI · Supports 1 field
Open source →
Crafted experiments to evaluate feature selection methods for single-cell RNA-seq data · other: Article DOI · Supports 2 fields
Open source →
/latest_version/versions/0
Unresolved field claims
/task_counts/subsetsProvisional · high — The double-pass review established the root item total but did not establish an exhaustive formal-subset inventory; the empty list must not be interpreted as evidence that the benchmark has no subsets.
Evidence: crafted-experiments-automated-subset-coverage-evidence/access/licenseProvisional · high — The double-pass review verified the official resource identity but did not establish a redistributable benchmark license; the value remains null pending source-level license verification.
Evidence: crafted-experiments-automated-resource-evidence
View source-level modification history on GitHub →