evaluation · needs review
Caduceus (character tokens): regulatory sequence classification
Evaluation procedure
task-category MCC across benchmark datasets
- Model
- Caduceus (character tokens)
- Benchmark
- regulatory sequence classification
- Dataset
- genomic benchmark categories
- origin
- Independent external evaluation
- configuration
- 3.9M parameter variant
- protocol id
- Not reported
- dataset version
- Not reported
- split
- paper benchmark summary
- population
- Not reported
- inputs
- Not reported
- adaptation
- Not reported
- metric implementation
- Not reported
- aggregation
- Not reported
- budget
- Not reported
Metadata review: needs review. Unreported conditions prevent automatic comparisons.
Evaluation results
Release 2026-09-16-d74d282221a9 · 1 evaluation · 1 metric row. Different protocols are not a single leaderboard.
| Metric and finding | Coverage and uncertainty | Evidence |
|---|---|---|
| Caduceus (character tokens): regulatory sequence classification Model: Caduceus (character tokens) · Benchmark: regulatory sequence classification · Dataset: genomic benchmark categories task-category MCC across benchmark datasets Independent external evaluation · Evaluation metadata: needs review | ||
| 0.778 MCC Unit: unitless · Direction: unknown | Uncertainty: not reported in legacy extract Scored: Not reported · Eligible: Not reported | source checkedThe impact of tokenizer selection in genomic language models · Table 2, Regulatory row, Caduceus (char) MCC column Source checking is not independent reproduction. |
Sources and history
Release 2026-09-16-d74d282221a9 · Record review: needs review
- The impact of tokenizer selection in genomic language models · Original source · journal full text in PMC
Technical metadata and extraction receipts
Stable ID: evaluation-b2-genomic-tokenizer-selection-2025
- areas
- dna-genomes
- tasks
- regulatory sequence classification
- origin
- independent_paper
- protocol
- task-category MCC across benchmark datasets
- version
- 3.9M parameter variant
- comparison
- protocol id: Not reported; dataset version: Not reported; split: paper benchmark summary; population: Not reported; inputs: Not reported; adaptation: Not reported; metric implementation: Not reported; aggregation: Not reported; budget: Not reported
- missing metadata
- dataset version: not_reported_in_legacy_extract
Related records
- model: Caduceus (character tokens)
- benchmark: regulatory sequence classification
- dataset: genomic benchmark categories
- evaluation: Caduceus (character tokens) · MCC · genomic benchmark categories