39.19(0.37)% top_k_accuracy
UTRBERT-5mer · BEACON SPL · Top-k ACC
- Tested configuration
- UTRBERT-5mer
- Task
- BEACON SPL: Splice site prediction
- Dataset subset
- SpliceAI (BEACON split)
- Procedure
- BEACON harness, fixed downstream head per task. Train/validation/test 144,628/18,078/16,505; dataset SpliceAI.
- Evaluation
- UTRBERT-5mer on BEACON SPL: Splice site prediction
- Evidence
- Author-reported evaluation · source checkedBEACON: Benchmark for Comprehensive RNA Tasks and Language Models (arXiv:2406.10391v2) · Table3,p.8,row(UTRBERT-5mer),column(SPL)
A source-checked result verifies the numerical transcription, not every model or protocol detail. Evaluation metadata: source checked. Source checked does not mean independently reproduced.
Methods and reproduction
BEACON harness evaluation of UTRBERT-5mer on Splice site prediction, scored with Top-k ACC.
- task
- BEACON SPL: Splice site prediction
- configuration
- UTRBERT-5mer
- dataset subset
- SpliceAI (BEACON split)
- Split
- 144,628/18,078/16,505
- Adaptation
- Published pre-trained RNA language model, fine-tuned by the BEACON authors
- Scoring implementation
- Top-k ACC
No execution recipe has been verified for this exact configuration and evaluation. A benchmark's general instructions may use different inputs, splits or model settings.
Reproducing this published result requires matching its model configuration, data, split and scorer. Source checking or a successful smoke test does not establish score reproduction.
Evaluation results
Release 2026-09-17-1a18ca2c038a · 1 evaluation · 1 metric row. Different protocols are not a single leaderboard. Where several source tables report the same metric, the published comparisons above offer a pooled view that names what it does not hold constant.
| Metric and finding | Coverage and uncertainty | Evidence |
|---|---|---|
| UTRBERT-5mer on BEACON SPL: Splice site prediction Configuration: UTRBERT-5merTask: BEACON SPL: Splice site predictionDataset subset: SpliceAI (BEACON split) BEACON harness, fixed downstream head per task. Train/validation/test 144,628/18,078/16,505; dataset SpliceAI. Author-reported evaluation · Evaluation metadata: source checked | ||
| 39.19(0.37)% top_k_accuracy Unit: percent · Direction: higher | Uncertainty: type: standard deviation; value: 0.37 Scored: Not reported · Eligible: Not reported | source checkedBEACON: Benchmark for Comprehensive RNA Tasks and Language Models (arXiv:2406.10391v2) · Table3,p.8,row(UTRBERT-5mer),column(SPL) Source checking is not independent reproduction. |
Evidence table
Inspect claims, sources and review details
Trace each statement to its source and review. A context-only reference supports the record generally; it does not verify an individual field. Source checking does not reproduce an experiment.
One row per statement and cited source. Multiple citations are not independent evaluations. Shared locators are labelled explicitly.
1 evidence row matching the loaded filters
| Property and statement | Original source and location | Review and provenance |
|---|---|---|
| Reported result 39.19(0.37) Individual claims | BEACON: Benchmark for Comprehensive RNA Tasks and Language Models (arXiv:2406.10391v2) Table3,p.8,row(UTRBERT-5mer),column(SPL) Version: v2, 12 December 2024 | source checked Deterministic parse of the pinned PDF text layer, with row and cell counts asserted and ten values cross-checked against the printed table · 2026-09-18 author reported Audit detailsSource checked, not reproduced. Metrics differ by task, so no composite score across tasks is computed or implied. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record Extraction artifact SHA-256: |
Sources and history
Release 2026-09-17-1a18ca2c038a · Record review: source checked
1 source records and release history
- BEACON: Benchmark for Comprehensive RNA Tasks and Language Models (arXiv:2406.10391v2) · Original source · v2, 12 December 2024
Technical metadata and extraction receipts
Stable ID: beacon-result-utrbert-5mer-spl-top-k-accuracy
- areas
- rna-transcriptomes
- tasks
- Splice site prediction
- metric
- top_k_accuracy
- metric direction
- higher
- unit
- percent
- printed value
- 39.19(0.37)
- numeric value
- 39.19
- uncertainty
- type: standard_deviation; value: 0.37
- source locator
- Table3,p.8,row(UTRBERT-5mer),column(SPL)
- missing metadata
- denominator: unextracted; seeds: unreported
- review
- method: Deterministic parse of the pinned PDF text layer, with row and cell counts asserted and ten values cross-checked against the printed table; reviewer: Codex research agent; no human review claimed; date: 2026-09-18; artifact sha256: 1370d75fe591bb8f994bc67f100a1a3fd557a62b9edb962b21f2fb038e9dea16; retrieval url: https://arxiv.org/pdf/2406.10391; notes: Source checked, not reproduced. Metrics differ by task, so no composite score across tasks is computed or implied.