Model type
DNA sequence transformer; this record is the paper-specific evaluated configuration.
DART-Eval tests DNABERT-2 on regulatory DNA under separately defined zero-shot, probed and fine-tuned protocols.
Conceptual input–method–output guide. Check the procedure text and linked evaluation for fitted components, additional inputs and exact settings.
DNA sequence transformer; this record is the paper-specific evaluated configuration.
Regulatory DNA and matched controls defined by the DART-Eval task
Regulatory-element scores or task-specific predictions, according to the evaluation protocol
limited source coverage · Automated source review, 2026-09-16. All specifications and missing details
Release 2026-09-17-d277315f7d76 · 1 evaluation · 1 metric row. Different protocols are not a single leaderboard.
| Metric and finding | Coverage and uncertainty | Evidence |
|---|---|---|
| DNABERT-2: regulatory element identification Configuration: DNABERT-2Task: regulatory element identificationDataset: DART-Eval cCREs versus matched shuffled controls zero-shot likelihood ranking: higher likelihood for cCRE than matched control Independent external evaluation · Evaluation metadata: needs review | ||
| 0.876 accuracy Unit: fraction · Direction: unknown | Uncertainty: not reported in legacy extract Scored: Not reported · Eligible: Not reported | source checkedDART-Eval: A Comprehensive DNA Language Model Evaluation Benchmark on Regulatory DNA · Table 3 (PDF page 5), DNABERT-2 row, Zero-Shot Accuracy column Source checking is not independent reproduction. |
The masked DNA transformer uses byte-pair tokenisation. The linked regulatory-element row uses the paper’s zero-shot procedure; trained probing or fine-tuning rows are distinct evaluations.
DNABERT-2 replaces overlapping k-mer tokens with byte-pair encoding and uses ALiBi positional biases. The official 117M model produces 768-dimensional token representations; downstream classifiers and pooling choices are separate configuration details.
The linked evaluation record identifies DNABERT-2: regulatory element identification. Its dataset, split, adaptation and evidence origin remain attached to the reported results.
Primary full text and the available official implementation/model documentation were inspected. Explanatory claims are source-backed; unresolved exact-configuration metadata is labelled explicitly. This is automated review, not a human review or independent benchmark reproduction.
Stable record: reported-model-28413ae1766316Explanatory profile: limited source coverage · Automated source review, 2026-09-16. Review applies to the cited claims; unresolved fields are listed below. Numerical results retain their own review status.
| Property | Description and evidence |
|---|---|
| Model type | DNA sequence transformer; this record is the paper-specific evaluated configuration.SourcesMAGICS-LAB/DNABERT_2 README.md · README.md model description |
| Architecture / procedure | The masked DNA transformer uses byte-pair tokenisation. The linked regulatory-element row uses the paper’s zero-shot procedure; trained probing or fine-tuning rows are distinct evaluations.SourcesDART-Eval: A Comprehensive DNA Language Model Evaluation Benchmark on Regulatory DNA · Table 2 (models); Section 3.2 zero-shot analysis; Table 3 regulatory-element identification; Appendix B evaluation procedures |
| Biological inputs | Regulatory DNA and matched controls defined by the DART-Eval taskSourcesDART-Eval: A Comprehensive DNA Language Model Evaluation Benchmark on Regulatory DNA · Table 2 (models); Section 3.2 zero-shot analysis; Table 3 regulatory-element identification; Appendix B evaluation procedures |
| Outputs | Regulatory-element scores or task-specific predictions, according to the evaluation protocolSourcesDART-Eval: A Comprehensive DNA Language Model Evaluation Benchmark on Regulatory DNA · Table 2 (models); Section 3.2 zero-shot analysis; Table 3 regulatory-element identification; Appendix B evaluation procedures |
| Parameters | 117 million parametersSourcesDART-Eval: A Comprehensive DNA Language Model Evaluation Benchmark on Regulatory DNA · Table 2 (models); Section 3.2 zero-shot analysis; Table 3 regulatory-element identification; Appendix B evaluation procedures |
| Known versions / configuration | DNABERT-2 is the comparison-table label; that label does not specify an immutable weight revision. · Not reported in inspected sourcesSourcesDART-Eval: A Comprehensive DNA Language Model Evaluation Benchmark on Regulatory DNA · Model identification in the comparison table and corresponding Methods; immutable checkpoint revision is not supplied by the table label. |
| Training data / fitting | The benchmark identifies multispecies genomic pretraining; task adaptation is reported separately.SourcesDART-Eval: A Comprehensive DNA Language Model Evaluation Benchmark on Regulatory DNA · Table 2 (models); Section 3.2 zero-shot analysis; Table 3 regulatory-element identification; Appendix B evaluation procedures |
| Context limits | A maximum input/context length for this exact evaluated configuration is not established by the inspected sources. · Not reported in inspected sourcesSources (2)DART-Eval: A Comprehensive DNA Language Model Evaluation Benchmark on Regulatory DNA; MAGICS-LAB/DNABERT_2 README.md · Complete primary text and named comparison table; inspected for explicit maximum input length (dataset lengths and family-wide limits are not substituted); README.md at pinned repository revision |
| Access | Official upstream implementation and usage documentation: https://github.com/MAGICS-LAB/DNABERT_2/blob/f25bed9ee20db966dff39e5c1571249d04e36404/README.md. This pinned documentation revision is not automatically the evaluated weight revision.SourcesMAGICS-LAB/DNABERT_2 README.md · README.md; installation, model download and usage instructions |
| Code licence | Apache 2.0 (upstream repository code at the cited revision; this does not establish every dependency or historical checkpoint licence).SourcesMAGICS-LAB/DNABERT_2 LICENSE · LICENSE; complete licence text |
| Weights licence | The inspected model-access documentation does not explicitly identify terms for this exact evaluated checkpoint or fitted head; repository code terms are shown separately. · Not reported in inspected sourcesSourcesMAGICS-LAB/DNABERT_2 README.md · README.md; checkpoint/access documentation and licence scope |
Trace each statement to its source and review. A context-only reference supports the record generally; it does not verify an individual field. Source checking does not reproduce an experiment.
One row per statement and cited source. Multiple citations are not independent evaluations. Shared locators are labelled explicitly.
21 evidence rows matching the loaded filters
| Property and statement | Original source and location | Review and provenance |
|---|---|---|
| Diagram caption Conceptual input–method–output guide. Check the procedure text and linked evaluation for fitted components, additional inputs and exact settings. Individual claims | DART-Eval: A Comprehensive DNA Language Model Evaluation Benchmark on Regulatory DNA Table 2 (models); Section 3.2 zero-shot analysis; Table 3 regulatory-element identification; Appendix B evaluation procedures Version: NeurIPS 2024 Datasets and Benchmarks Track proceedings | source checked automated source review · 2026-09-16 Audit detailsPrimary full text and the available official implementation/model documentation were inspected. Explanatory claims are source-backed; unresolved exact-configuration metadata is labelled explicitly. This is automated review, not a human review or independent benchmark reproduction. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| Diagram steps ["Regulatory DNA and matched controls defined by the DART-Eval task","DNABERT-2","Regulatory-element scores or task-specific predictions, according to the evaluation protocol"] Individual claims | DART-Eval: A Comprehensive DNA Language Model Evaluation Benchmark on Regulatory DNA Table 2 (models); Section 3.2 zero-shot analysis; Table 3 regulatory-element identification; Appendix B evaluation procedures Version: NeurIPS 2024 Datasets and Benchmarks Track proceedings | source checked automated source review · 2026-09-16 Audit detailsPrimary full text and the available official implementation/model documentation were inspected. Explanatory claims are source-backed; unresolved exact-configuration metadata is labelled explicitly. This is automated review, not a human review or independent benchmark reproduction. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| Diagram title Evaluated procedure (conceptual) Individual claims | DART-Eval: A Comprehensive DNA Language Model Evaluation Benchmark on Regulatory DNA Table 2 (models); Section 3.2 zero-shot analysis; Table 3 regulatory-element identification; Appendix B evaluation procedures Version: NeurIPS 2024 Datasets and Benchmarks Track proceedings | source checked automated source review · 2026-09-16 Audit detailsPrimary full text and the available official implementation/model documentation were inspected. Explanatory claims are source-backed; unresolved exact-configuration metadata is labelled explicitly. This is automated review, not a human review or independent benchmark reproduction. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| Model type DNA sequence transformer; this record is the paper-specific evaluated configuration. Individual claims | MAGICS-LAB/DNABERT_2 README.md README.md model description Version: f25bed9ee20db966dff39e5c1571249d04e36404 | source checked automated source review · 2026-09-16 Audit detailsPrimary full text and the available official implementation/model documentation were inspected. Explanatory claims are source-backed; unresolved exact-configuration metadata is labelled explicitly. This is automated review, not a human review or independent benchmark reproduction. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| Architecture / procedure The masked DNA transformer uses byte-pair tokenisation. The linked regulatory-element row uses the paper’s zero-shot procedure; trained probing or fine-tuning rows are distinct evaluations. Individual claims | DART-Eval: A Comprehensive DNA Language Model Evaluation Benchmark on Regulatory DNA Table 2 (models); Section 3.2 zero-shot analysis; Table 3 regulatory-element identification; Appendix B evaluation procedures Version: NeurIPS 2024 Datasets and Benchmarks Track proceedings | source checked automated source review · 2026-09-16 Audit detailsPrimary full text and the available official implementation/model documentation were inspected. Explanatory claims are source-backed; unresolved exact-configuration metadata is labelled explicitly. This is automated review, not a human review or independent benchmark reproduction. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| Weights licence The inspected model-access documentation does not explicitly identify terms for this exact evaluated checkpoint or fitted head; repository code terms are shown separately. Individual claims | MAGICS-LAB/DNABERT_2 README.md README.md; checkpoint/access documentation and licence scope Version: f25bed9ee20db966dff39e5c1571249d04e36404 | unreported automated source review · 2026-09-16 Audit detailsPrimary full text and the available official implementation/model documentation were inspected. Explanatory claims are source-backed; unresolved exact-configuration metadata is labelled explicitly. This is automated review, not a human review or independent benchmark reproduction. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| Biological inputs Regulatory DNA and matched controls defined by the DART-Eval task Individual claims | DART-Eval: A Comprehensive DNA Language Model Evaluation Benchmark on Regulatory DNA Table 2 (models); Section 3.2 zero-shot analysis; Table 3 regulatory-element identification; Appendix B evaluation procedures Version: NeurIPS 2024 Datasets and Benchmarks Track proceedings | source checked automated source review · 2026-09-16 Audit detailsPrimary full text and the available official implementation/model documentation were inspected. Explanatory claims are source-backed; unresolved exact-configuration metadata is labelled explicitly. This is automated review, not a human review or independent benchmark reproduction. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| Outputs Regulatory-element scores or task-specific predictions, according to the evaluation protocol Individual claims | DART-Eval: A Comprehensive DNA Language Model Evaluation Benchmark on Regulatory DNA Table 2 (models); Section 3.2 zero-shot analysis; Table 3 regulatory-element identification; Appendix B evaluation procedures Version: NeurIPS 2024 Datasets and Benchmarks Track proceedings | source checked automated source review · 2026-09-16 Audit detailsPrimary full text and the available official implementation/model documentation were inspected. Explanatory claims are source-backed; unresolved exact-configuration metadata is labelled explicitly. This is automated review, not a human review or independent benchmark reproduction. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| Parameters 117 million parameters Individual claims | DART-Eval: A Comprehensive DNA Language Model Evaluation Benchmark on Regulatory DNA Table 2 (models); Section 3.2 zero-shot analysis; Table 3 regulatory-element identification; Appendix B evaluation procedures Version: NeurIPS 2024 Datasets and Benchmarks Track proceedings | source checked automated source review · 2026-09-16 Audit detailsPrimary full text and the available official implementation/model documentation were inspected. Explanatory claims are source-backed; unresolved exact-configuration metadata is labelled explicitly. This is automated review, not a human review or independent benchmark reproduction. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| Known versions / configuration DNABERT-2 is the comparison-table label; that label does not specify an immutable weight revision. Individual claims | DART-Eval: A Comprehensive DNA Language Model Evaluation Benchmark on Regulatory DNA Model identification in the comparison table and corresponding Methods; immutable checkpoint revision is not supplied by the table label. Version: NeurIPS 2024 Datasets and Benchmarks Track proceedings | unreported automated source review · 2026-09-16 Audit detailsPrimary full text and the available official implementation/model documentation were inspected. Explanatory claims are source-backed; unresolved exact-configuration metadata is labelled explicitly. This is automated review, not a human review or independent benchmark reproduction. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
Release 2026-09-17-d277315f7d76 · Record review: needs review
Stable ID: reported-model-28413ae1766316