evaluation · needs review
C2S (GPT-2 Large): Combinatorial cell-label classification
Evaluation procedure
Partial-credit labels including cell type, perturbation, and dose.
- Model
- C2S (GPT-2 Large)
- Benchmark
- Combinatorial cell-label classification
- Dataset
- L1000
- origin
- Author-reported evaluation
- configuration
- GPT-2 Large
- protocol id
- Not reported
- dataset version
- Not reported
- split
- Not reported
- population
- Not reported
- inputs
- Not reported
- adaptation
- Not reported
- metric implementation
- Not reported
- aggregation
- Not reported
- budget
- Not reported
Metadata review: needs review. Unreported conditions prevent automatic comparisons.
Evaluation results
Release 2026-09-16-d74d282221a9 · 1 evaluation · 1 metric row. Different protocols are not a single leaderboard.
| Metric and finding | Coverage and uncertainty | Evidence |
|---|---|---|
| C2S (GPT-2 Large): Combinatorial cell-label classification Partial-credit labels including cell type, perturbation, and dose. Author-reported evaluation · Evaluation metadata: needs review | ||
| 0.631 Partial-label accuracy Unit: unitless · Direction: unknown | Uncertainty: ± 0.0031 Scored: Not reported · Eligible: Not reported | source checkedCell2Sentence: Teaching Large Language Models the Language of Biology · Table 3, Partial label / C2S (GPT-2 Large) row, L1000 Acc column Source checking is not independent reproduction. |
Sources and history
Release 2026-09-16-d74d282221a9 · Record review: needs review
- Cell2Sentence: Teaching Large Language Models the Language of Biology · Original source · preprint archived 2024-10-29
Technical metadata and extraction receipts
Stable ID: evaluation-lit-027
- areas
- cells-tissues
- tasks
- Combinatorial cell-label classification
- origin
- author_reported
- protocol
- Partial-credit labels including cell type, perturbation, and dose.
- version
- GPT-2 Large
- comparison
- protocol id: Not reported; dataset version: Not reported; split: Not reported; population: Not reported; inputs: Not reported; adaptation: Not reported; metric implementation: Not reported; aggregation: Not reported; budget: Not reported
- missing metadata
- dataset version: not_reported_in_legacy_extract; split: not_reported_in_legacy_extract
Related records
- model: C2S (GPT-2 Large)
- benchmark: Combinatorial cell-label classification
- dataset: L1000
- evaluation: C2S (GPT-2 Large) · Partial-label accuracy · L1000