0.932 paired_accuracy
Probing-head-like (ab initio) · DART-Eval REI-PAIR · Paired accuracy
- Tested method
- Probing-head-like (ab initio)
- Task
- DART-Eval REI-PAIR: Regulatory element identification, paired accuracy
- Dataset subset
- ENCODE cCREs against dinucleotide-shuffled backgrounds (DART-Eval split)
- Procedure
- Distinguish ENCODE cCREs from dinucleotide-shuffled background sequences. The setting the model was run in is part of the method name.
- Evaluation
- Probing-head-like (ab initio) on DART-Eval REI-PAIR: Regulatory element identification, paired accuracy
- Evidence
- Author-reported evaluation · source checkedDART-Eval: A Comprehensive DNA Language Model Evaluation Benchmark on Regulatory DNA · Table 3, row(Ab initio Probing-head-like), column(ab initio paired_accuracy)
A source-checked result verifies the numerical transcription, not every model or protocol detail. Evaluation metadata: source checked. Source checked does not mean independently reproduced.
Methods and reproduction
DART-Eval evaluation of Probing-head-like (ab initio) on Regulatory element identification, paired accuracy, scored with Paired accuracy.
- task
- DART-Eval REI-PAIR: Regulatory element identification, paired accuracy
- method
- Probing-head-like (ab initio)
- dataset subset
- ENCODE cCREs against dinucleotide-shuffled backgrounds (DART-Eval split)
- Split
- Not reported
- Adaptation
- Not reported
- Scoring implementation
- Paired accuracy
No execution recipe has been verified for this exact configuration and evaluation. A benchmark's general instructions may use different inputs, splits or model settings.
Reproducing this published result requires matching its model configuration, data, split and scorer. Source checking or a successful smoke test does not establish score reproduction.
Evaluation results
Release 2026-09-17-134cd1815de8 · 1 evaluation · 1 metric row. Different protocols are not a single leaderboard. Where several source tables report the same metric, the published comparisons above offer a pooled view that names what it does not hold constant.
| Metric and finding | Coverage and uncertainty | Evidence |
|---|---|---|
| Probing-head-like (ab initio) on DART-Eval REI-PAIR: Regulatory element identification, paired accuracy Method: Probing-head-like (ab initio)Task: DART-Eval REI-PAIR: Regulatory element identification, paired accuracyDataset subset: ENCODE cCREs against dinucleotide-shuffled backgrounds (DART-Eval split) Distinguish ENCODE cCREs from dinucleotide-shuffled background sequences. The setting the model was run in is part of the method name. Author-reported evaluation · Evaluation metadata: source checked | ||
| 0.932 paired_accuracy Unit: fraction · Direction: higher | Uncertainty: Not reported Scored: Not reported · Eligible: Not reported | source checkedDART-Eval: A Comprehensive DNA Language Model Evaluation Benchmark on Regulatory DNA · Table 3, row(Ab initio Probing-head-like), column(ab initio paired_accuracy) Source checking is not independent reproduction. |
Evidence table
Inspect claims, sources and review details
Trace each statement to its source and review. A context-only reference supports the record generally; it does not verify an individual field. Source checking does not reproduce an experiment.
One row per statement and cited source. Multiple citations are not independent evaluations. Shared locators are labelled explicitly.
1 evidence row matching the loaded filters
| Property and statement | Original source and location | Review and provenance |
|---|---|---|
| Reported result 0.932 Individual claims | DART-Eval: A Comprehensive DNA Language Model Evaluation Benchmark on Regulatory DNA Table 3, row(Ab initio Probing-head-like), column(ab initio paired_accuracy) Version: 2412.05430v1 | source checked Deterministic parse of the pinned HTML tables, with row and column counts asserted · 2026-09-18 author reported Audit detailsSource checked, not reproduced. Metrics differ by task, so no composite score across tasks is computed or implied. Field: Source artifact SHA-256: Hash scope: Exact retrieved primary paper artifact bytes. Extraction artifact SHA-256: |
Sources and history
Release 2026-09-17-134cd1815de8 · Record review: source checked
1 source records and release history
- DART-Eval: A Comprehensive DNA Language Model Evaluation Benchmark on Regulatory DNA · Original source · 2412.05430v1
Technical metadata and extraction receipts
Stable ID: dart-eval-result-probing-head-like-ab-initio-rei-pair-paired-accuracy
- areas
- dna-genomes
- tasks
- Regulatory element identification, paired accuracy
- metric
- paired_accuracy
- metric direction
- higher
- unit
- fraction
- printed value
- 0.932
- numeric value
- 0.932
- uncertainty
- Not reported
- source locator
- Table 3, row(Ab initio Probing-head-like), column(ab initio paired_accuracy)
- missing metadata
- denominator: unextracted; seeds: unreported
- review
- method: Deterministic parse of the pinned HTML tables, with row and column counts asserted; reviewer: Codex research agent; no human review claimed; date: 2026-09-18; artifact sha256: 4194b137ba55c9a2c269d119a9afec6ae1bb0feaf17d91433ae483c41221a56b; retrieval url: https://arxiv.org/html/2412.05430v1; notes: Source checked, not reproduced. Metrics differ by task, so no composite score across tasks is computed or implied.