Strengths and considerations
- The linked evaluation identifies the paper-specific dataset and assessment rather than treating the task name as a universal benchmark.Benchmarking DNA large language models on quadruplexes · Table 5, DNABERT-2 (117 M) row, Accuracy column; Table 5, Caduceus (8 M) row, Accuracy column