UTR-LM-MRL on BEACON SPL: Splice site prediction
BEACON harness evaluation of UTR-LM-MRL on Splice site prediction, scored with Top-k ACC.
Methods and reproduction
BEACON harness evaluation of UTR-LM-MRL on Splice site prediction, scored with Top-k ACC.
- task
- BEACON SPL: Splice site prediction
- configuration
- UTR-LM-MRL
- dataset subset
- SpliceAI (BEACON split)
- Split
- 144,628/18,078/16,505
- Adaptation
- Published pre-trained RNA language model, fine-tuned by the BEACON authors
- Scoring implementation
- Top-k ACC
No execution recipe has been verified for this exact configuration and evaluation. A benchmark's general instructions may use different inputs, splits or model settings.
Reproducing this published result requires matching its model configuration, data, split and scorer. Source checking or a successful smoke test does not establish score reproduction.
Evaluation procedure
BEACON harness, fixed downstream head per task. Train/validation/test 144,628/18,078/16,505; dataset SpliceAI.
- Configuration
- UTR-LM-MRL
- Task
- BEACON SPL: Splice site prediction
- Dataset subset
- SpliceAI (BEACON split)
- origin
- Author-reported evaluation
- configuration
- Not reported
- protocol id
- beacon-task-spl
- split
- 144,628/18,078/16,505
- adaptation
- Published pre-trained RNA language model, fine-tuned by the BEACON authors
- metric implementation
- Top-k ACC
Metadata review: source checked. Unreported conditions prevent automatic comparisons.
Evaluation results
Release 2026-09-17-1a18ca2c038a · 1 evaluation · 1 metric row. Different protocols are not a single leaderboard. Where several source tables report the same metric, the published comparisons above offer a pooled view that names what it does not hold constant.
| Metric and finding | Coverage and uncertainty | Evidence |
|---|---|---|
| UTR-LM-MRL on BEACON SPL: Splice site prediction Configuration: UTR-LM-MRLTask: BEACON SPL: Splice site predictionDataset subset: SpliceAI (BEACON split) BEACON harness, fixed downstream head per task. Train/validation/test 144,628/18,078/16,505; dataset SpliceAI. Author-reported evaluation · Evaluation metadata: source checked | ||
| 36.20(1.84)% top_k_accuracy Unit: percent · Direction: higher | Uncertainty: type: standard deviation; value: 1.84 Scored: Not reported · Eligible: Not reported | source checkedBEACON: Benchmark for Comprehensive RNA Tasks and Language Models (arXiv:2406.10391v2) · Table3,p.8,row(UTR-LM-MRL),column(SPL) Source checking is not independent reproduction. |
Evidence table
Inspect claims, sources and review details
Trace each statement to its source and review. A context-only reference supports the record generally; it does not verify an individual field. Source checking does not reproduce an experiment.
One row per statement and cited source. Multiple citations are not independent evaluations. Shared locators are labelled explicitly.
12 evidence rows matching the loaded filters
| Property and statement | Original source and location | Review and provenance |
|---|---|---|
| attributes.comparison.adaptation Published pre-trained RNA language model, fine-tuned by the BEACON authors Context-only references | BEACON: Benchmark for Comprehensive RNA Tasks and Language Models (arXiv:2406.10391v2) Table3,p.8,row(UTR-LM-MRL),column(SPL) Version: v2, 12 December 2024 | not individually reviewed No individual claim review recorded author reported Audit detailsField: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| attributes.comparison.metric_implementation Top-k ACC Context-only references | BEACON: Benchmark for Comprehensive RNA Tasks and Language Models (arXiv:2406.10391v2) Table3,p.8,row(UTR-LM-MRL),column(SPL) Version: v2, 12 December 2024 | not individually reviewed No individual claim review recorded author reported Audit detailsField: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| attributes.comparison.protocol_id beacon-task-spl Context-only references | BEACON: Benchmark for Comprehensive RNA Tasks and Language Models (arXiv:2406.10391v2) Table3,p.8,row(UTR-LM-MRL),column(SPL) Version: v2, 12 December 2024 | not individually reviewed No individual claim review recorded author reported Audit detailsField: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| attributes.comparison.split 144,628/18,078/16,505 Context-only references | BEACON: Benchmark for Comprehensive RNA Tasks and Language Models (arXiv:2406.10391v2) Table3,p.8,row(UTR-LM-MRL),column(SPL) Version: v2, 12 December 2024 | not individually reviewed No individual claim review recorded author reported Audit detailsField: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| attributes.origin author_reported Context-only references | BEACON: Benchmark for Comprehensive RNA Tasks and Language Models (arXiv:2406.10391v2) Table3,p.8,row(UTR-LM-MRL),column(SPL) Version: v2, 12 December 2024 | not individually reviewed No individual claim review recorded author reported Audit detailsField: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| attributes.protocol BEACON harness, fixed downstream head per task. Train/validation/test 144,628/18,078/16,505; dataset SpliceAI. Context-only references | BEACON: Benchmark for Comprehensive RNA Tasks and Language Models (arXiv:2406.10391v2) Table3,p.8,row(UTR-LM-MRL),column(SPL) Version: v2, 12 December 2024 | not individually reviewed No individual claim review recorded author reported Audit detailsField: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| attributes.source_locator Table3,p.8,row(UTR-LM-MRL),column(SPL) Context-only references | BEACON: Benchmark for Comprehensive RNA Tasks and Language Models (arXiv:2406.10391v2) Table3,p.8,row(UTR-LM-MRL),column(SPL) Version: v2, 12 December 2024 | not individually reviewed No individual claim review recorded author reported Audit detailsField: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| description BEACON harness evaluation of UTR-LM-MRL on Splice site prediction, scored with Top-k ACC. Context-only references | BEACON: Benchmark for Comprehensive RNA Tasks and Language Models (arXiv:2406.10391v2) Table3,p.8,row(UTR-LM-MRL),column(SPL) Version: v2, 12 December 2024 | not individually reviewed No individual claim review recorded author reported Audit detailsField: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| Relationship: benchmark beacon-task-spl Context-only references | BEACON: Benchmark for Comprehensive RNA Tasks and Language Models (arXiv:2406.10391v2) Table3,p.8,row(UTR-LM-MRL),column(SPL) Version: v2, 12 December 2024 | not individually reviewed No individual claim review recorded author reported Audit detailsField: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| Relationship: dataset beacon-dataset-spliceai Context-only references | BEACON: Benchmark for Comprehensive RNA Tasks and Language Models (arXiv:2406.10391v2) Table3,p.8,row(UTR-LM-MRL),column(SPL) Version: v2, 12 December 2024 | not individually reviewed No individual claim review recorded author reported Audit detailsField: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
Sources and history
Release 2026-09-17-1a18ca2c038a · Record review: source checked
1 source records and release history
- BEACON: Benchmark for Comprehensive RNA Tasks and Language Models (arXiv:2406.10391v2) · Original source · v2, 12 December 2024
Technical metadata and extraction receipts
Stable ID: beacon-evaluation-utr-lm-mrl-spl
- areas
- rna-transcriptomes
- tasks
- Splice site prediction
- origin
- author_reported
- protocol
- BEACON harness, fixed downstream head per task. Train/validation/test 144,628/18,078/16,505; dataset SpliceAI.
- comparison
- protocol id: beacon-task-spl; split: 144,628/18,078/16,505; adaptation: Published pre-trained RNA language model, fine-tuned by the BEACON authors; metric implementation: Top-k ACC
- missing metadata
- checkpoint revision: unreported; seeds: unreported; budget: unreported; split manifest: unextracted
- source locator
- Table3,p.8,row(UTR-LM-MRL),column(SPL)
Related records
- benchmark: BEACON SPL: Splice site prediction
- model: UTR-LM-MRL
- dataset: SpliceAI (BEACON split)
- evaluation: UTR-LM-MRL · BEACON SPL · Top-k ACC