rewire.it
Task

flu-vaccine mRNA property prediction

The flu-vaccine entry refers to a supervised mRNA-property evaluation dataset within the CodonBERT paper.

SourcesCodonBERT large language model for mRNA vaccines · Results: supervised learning datasets; cached text lines 14–17; task metric definitions and corresponding results table; matching task comparison table/ablation captions

1 evaluation · 1 metric row

At a glance

Inputs, training, access and other details

Explanatory profile: limited source coverage · Automated source review, 2026-09-16. Review applies to the cited claims; unresolved fields are listed below. Numerical results retain their own review status.

Data, procedure and scoring
PropertyDescription and evidence
DatasetsA separately generated flu-vaccine mRNA dataset with measured expression-related labels.
SourcesCodonBERT large language model for mRNA vaccines · Results: supervised learning datasets; cached text lines 14–17; task metric definitions and corresponding results table; matching task comparison table/ablation captions
SplitsEach downstream dataset uses a shared 70:15:15 training/validation/test split across methods, according to Table 1.
SourcesCodonBERT large language model for mRNA vaccines · Results: supervised learning datasets; cached text lines 14–17; task metric definitions and corresponding results table; matching task comparison table/ablation captions
MetricsSpearman rank correlation for the vaccine-expression regression task; classification accuracy belongs to a different E. coli task.
SourcesCodonBERT large language model for mRNA vaccines · Results: supervised learning datasets; cached text lines 14–17; task metric definitions and corresponding results table; matching task comparison table/ablation captions
BaselinesPlain TextCNN, RNABERT/TextCNN, RNA-FM/TextCNN, TF-IDF and Codon2vec/TextCNN configurations in the expression-regression table.
SourcesCodonBERT large language model for mRNA vaccines · Results: supervised learning datasets; cached text lines 14–17; task metric definitions and corresponding results table; matching task comparison table/ablation captions
Leakage controlsThe shared split controls partition differences between methods. Homology or construct-family exclusion is not specified in the checked downstream-task split statement.
SourcesCodonBERT large language model for mRNA vaccines · Results: supervised learning datasets; cached text lines 14–17; task metric definitions and corresponding results table; matching task comparison table/ablation captions
UncertaintyThe cited text-accessible evaluation sections give no confidence-interval, resampling or repeat-run error-bar specification. Image-only tables and uninspected supplements are outside this absence claim. · Not reported in inspected sources
SourcesCodonBERT large language model for mRNA vaccines · Results: supervised learning datasets; cached text lines 14–17; task metric definitions and corresponding results table; matching task comparison table/ablation captions
Entity typePaper-specific computational evaluation protocol.
SourcesCodonBERT large language model for mRNA vaccines · Results: supervised learning datasets; cached text lines 14–17; task metric definitions and corresponding results table; matching task comparison table/ablation captions
OrganismsInfluenza-vaccine mRNA constructs.
SourcesCodonBERT large language model for mRNA vaccines · Results: supervised learning datasets; cached text lines 14–17; task metric definitions and corresponding results table; matching task comparison table/ablation captions
AssaysMeasured expression-related mRNA labels.
SourcesCodonBERT large language model for mRNA vaccines · Results: supervised learning datasets; cached text lines 14–17; task metric definitions and corresponding results table; matching task comparison table/ablation captions
Allowed inputsmRNA sequence representations.
SourcesCodonBERT large language model for mRNA vaccines · Results: supervised learning datasets; cached text lines 14–17; task metric definitions and corresponding results table; matching task comparison table/ablation captions
AdaptationSupervised downstream expression regression; all methods use the same task split.
SourcesCodonBERT large language model for mRNA vaccines · Results: supervised learning datasets; cached text lines 14–17; task metric definitions and corresponding results table; matching task comparison table/ablation captions

How it works

How it worksComputational evaluation flow
Computational evaluation flow1. Input: mRNA sequence representations.. Then: 2. Evaluation: Supervised downstream expression regression; all methods use the same task split.. Then: 3. Readout: Spearman rank correlation for the vaccine-expression regression task; classification accuracy belongs to a different E. coli task.Computational evaluation flow1. Input: mRNA sequence representations.. Then: 2. Evaluation: Supervised downstream expression regression; all methods use the same task split.. Then: 3. Readout: Spearman rank correlation for the vaccine-expression regression task; classification accuracy belongs to a different E. coli task.Computational evaluation flow1. Input: mRNA sequence representations.. Then: 2. Evaluation: Supervised downstream expression regression; all methods use the same task split.. Then: 3. Readout: Spearman rank correlation for the vaccine-expression regression task; classification accuracy belongs to a different E. coli task.

Conceptual summary of the cited evaluation; exact task configuration and source version remain part of the protocol.

SourcesCodonBERT large language model for mRNA vaccines · Results: supervised learning datasets; cached text lines 14–17; task metric definitions and corresponding results table; matching task comparison table/ablation captions
Evaluation methodology

A separately generated flu-vaccine mRNA dataset with measured expression-related labels. Each downstream dataset uses a shared 70:15:15 training/validation/test split across methods, according to Table 1. Spearman rank correlation for the vaccine-expression regression task; classification accuracy belongs to a different E. coli task. Plain TextCNN, RNABERT/TextCNN, RNA-FM/TextCNN, TF-IDF and Codon2vec/TextCNN configurations in the expression-regression table. The shared split controls partition differences between methods. Homology or construct-family exclusion is not specified in the checked downstream-task split statement. The cited text-accessible evaluation sections give no confidence-interval, resampling or repeat-run error-bar specification. Image-only tables and uninspected supplements are outside this absence claim.

SourcesCodonBERT large language model for mRNA vaccines · Results: supervised learning datasets; cached text lines 14–17; task metric definitions and corresponding results table; matching task comparison table/ablation captions

Recorded evaluations

Each evaluation records what was tested and under which conditions.

Tested entities and results

Release 2026-09-17-d277315f7d76 · 1 evaluation · 1 metric row. Different protocols are not a single leaderboard.

Results grouped by the exact reported evaluation
Metric and findingCoverage and uncertaintyEvidence
CodonBERT: flu-vaccine mRNA property prediction

codon-based model fine-tuned for downstream regression

Author-reported evaluation · Evaluation metadata: needs review

0.81 Spearman rho

Unit: unitless · Direction: unknown

Uncertainty: not reported in legacy extract

Scored: Not reported · Eligible: Not reported

source checkedCodonBERT large language model for mRNA vaccines · Table 2, CodonBERT row, Flu vaccines Spearman correlation column

Source checking is not independent reproduction.

Papers and result coverage

Last literature check: 2026-09-17. Dated primary-source discovery and protocol/table screening. Source checking does not mean experimental reproduction. Only separately extracted and independently reviewed numeric batches are publishable.

Paper or primary resourceVersionReference
CodonBERT large language model for mRNA vaccinesjournal full text in PMCRead source

What is still missing

  • complete numerical transcription and independent cell review: Full primary artifact and table inventory preserved; no new numeric row is published from this audit alone.
  • exact checkpoint hashes and per-method scored denominators: Table labels alone do not establish these fields; do not infer checkpoint or scored count from model name or dataset size.
Search and extraction details

primary comparison tables located

Searches

  • CodonBERT large language model for mRNA vaccines 10.1101/gr.278870.123

Evidence locations

  • Table 2.; XML table GR278870LITB2

Strengths and limitations

Strengths and considerations

No source-reviewed explanatory claims are recorded here yet.

Limitations and conditions

  • Exact split membership, regression aggregation, baseline execution remain unresolved for the separate vaccine-expression dataset.
    SourcesCodonBERT large language model for mRNA vaccines · Results: supervised learning datasets; cached text lines 14–17; task metric definitions and corresponding results table; matching task comparison table/ablation captions
Profile review details

Task-specific computational methodology and field context checked in the cited primary-source artifact. Source-backed fields, inapplicable evaluator dimensions and unresolved details are distinguished. Numerical results were not reproduced.

Stable record: reported-task-1c74661df2c401

Evidence table

Inspect claims, sources and review details

Trace each statement to its source and review. A context-only reference supports the record generally; it does not verify an individual field. Source checking does not reproduce an experiment.

One row per statement and cited source. Multiple citations are not independent evaluations. Shared locators are labelled explicitly.

17 evidence rows matching the loaded filters

Claims, original sources and review scope · Release 2026-09-17-d277315f7d76
Property and statementOriginal source and locationReview and provenance
Diagram caption

Conceptual summary of the cited evaluation; exact task configuration and source version remain part of the protocol.

Individual claims
CodonBERT large language model for mRNA vaccines

Original source ↗

Results: supervised learning datasets; cached text lines 14–17; task metric definitions and corresponding results table; matching task comparison table/ablation captions

Version: journal full text in PMC
Retrieved: 2026-09-16T10:38:57.558201+00:00

source checked

automated source review · 2026-09-16

Audit details

Task-specific computational methodology and field context checked in the cited primary-source artifact. Source-backed fields, inapplicable evaluator dimensions and unresolved details are distinguished. Numerical results were not reproduced.

Field: attributes.profile.diagram.caption

Source artifact SHA-256: 2968073753e6d44feff9c08b131edf23145e95b171434539dddf77bb92847033

Hash scope: Hash scope not separately documented; inspect source record

Inspected artifact

Diagram steps

["Input: mRNA sequence representations.","Evaluation: Supervised downstream expression regression; all methods use the same task split.","Readout: Spearman rank correlation for the vaccine-expression regression task; classification accuracy belongs to a different E. coli task."]

Individual claims
CodonBERT large language model for mRNA vaccines

Original source ↗

Results: supervised learning datasets; cached text lines 14–17; task metric definitions and corresponding results table; matching task comparison table/ablation captions

Version: journal full text in PMC
Retrieved: 2026-09-16T10:38:57.558201+00:00

source checked

automated source review · 2026-09-16

Audit details

Task-specific computational methodology and field context checked in the cited primary-source artifact. Source-backed fields, inapplicable evaluator dimensions and unresolved details are distinguished. Numerical results were not reproduced.

Field: attributes.profile.diagram.steps

Source artifact SHA-256: 2968073753e6d44feff9c08b131edf23145e95b171434539dddf77bb92847033

Hash scope: Hash scope not separately documented; inspect source record

Inspected artifact

Diagram title

Computational evaluation flow

Individual claims
CodonBERT large language model for mRNA vaccines

Original source ↗

Results: supervised learning datasets; cached text lines 14–17; task metric definitions and corresponding results table; matching task comparison table/ablation captions

Version: journal full text in PMC
Retrieved: 2026-09-16T10:38:57.558201+00:00

source checked

automated source review · 2026-09-16

Audit details

Task-specific computational methodology and field context checked in the cited primary-source artifact. Source-backed fields, inapplicable evaluator dimensions and unresolved details are distinguished. Numerical results were not reproduced.

Field: attributes.profile.diagram.title

Source artifact SHA-256: 2968073753e6d44feff9c08b131edf23145e95b171434539dddf77bb92847033

Hash scope: Hash scope not separately documented; inspect source record

Inspected artifact

Datasets

A separately generated flu-vaccine mRNA dataset with measured expression-related labels.

Individual claims
CodonBERT large language model for mRNA vaccines

Original source ↗

Results: supervised learning datasets; cached text lines 14–17; task metric definitions and corresponding results table; matching task comparison table/ablation captions

Version: journal full text in PMC
Retrieved: 2026-09-16T10:38:57.558201+00:00

source checked

automated source review · 2026-09-16

Audit details

Task-specific computational methodology and field context checked in the cited primary-source artifact. Source-backed fields, inapplicable evaluator dimensions and unresolved details are distinguished. Numerical results were not reproduced.

Field: attributes.profile.facts.0.value

Source artifact SHA-256: 2968073753e6d44feff9c08b131edf23145e95b171434539dddf77bb92847033

Hash scope: Hash scope not separately documented; inspect source record

Inspected artifact

Splits

Each downstream dataset uses a shared 70:15:15 training/validation/test split across methods, according to Table 1.

Individual claims
CodonBERT large language model for mRNA vaccines

Original source ↗

Results: supervised learning datasets; cached text lines 14–17; task metric definitions and corresponding results table; matching task comparison table/ablation captions

Version: journal full text in PMC
Retrieved: 2026-09-16T10:38:57.558201+00:00

source checked

automated source review · 2026-09-16

Audit details

Task-specific computational methodology and field context checked in the cited primary-source artifact. Source-backed fields, inapplicable evaluator dimensions and unresolved details are distinguished. Numerical results were not reproduced.

Field: attributes.profile.facts.1.value

Source artifact SHA-256: 2968073753e6d44feff9c08b131edf23145e95b171434539dddf77bb92847033

Hash scope: Hash scope not separately documented; inspect source record

Inspected artifact

Adaptation

Supervised downstream expression regression; all methods use the same task split.

Individual claims
CodonBERT large language model for mRNA vaccines

Original source ↗

Results: supervised learning datasets; cached text lines 14–17; task metric definitions and corresponding results table; matching task comparison table/ablation captions

Version: journal full text in PMC
Retrieved: 2026-09-16T10:38:57.558201+00:00

source checked

automated source review · 2026-09-16

Audit details

Task-specific computational methodology and field context checked in the cited primary-source artifact. Source-backed fields, inapplicable evaluator dimensions and unresolved details are distinguished. Numerical results were not reproduced.

Field: attributes.profile.facts.10.value

Source artifact SHA-256: 2968073753e6d44feff9c08b131edf23145e95b171434539dddf77bb92847033

Hash scope: Hash scope not separately documented; inspect source record

Inspected artifact

Metrics

Spearman rank correlation for the vaccine-expression regression task; classification accuracy belongs to a different E. coli task.

Individual claims
CodonBERT large language model for mRNA vaccines

Original source ↗

Results: supervised learning datasets; cached text lines 14–17; task metric definitions and corresponding results table; matching task comparison table/ablation captions

Version: journal full text in PMC
Retrieved: 2026-09-16T10:38:57.558201+00:00

source checked

automated source review · 2026-09-16

Audit details

Task-specific computational methodology and field context checked in the cited primary-source artifact. Source-backed fields, inapplicable evaluator dimensions and unresolved details are distinguished. Numerical results were not reproduced.

Field: attributes.profile.facts.2.value

Source artifact SHA-256: 2968073753e6d44feff9c08b131edf23145e95b171434539dddf77bb92847033

Hash scope: Hash scope not separately documented; inspect source record

Inspected artifact

Baselines

Plain TextCNN, RNABERT/TextCNN, RNA-FM/TextCNN, TF-IDF and Codon2vec/TextCNN configurations in the expression-regression table.

Individual claims
CodonBERT large language model for mRNA vaccines

Original source ↗

Results: supervised learning datasets; cached text lines 14–17; task metric definitions and corresponding results table; matching task comparison table/ablation captions

Version: journal full text in PMC
Retrieved: 2026-09-16T10:38:57.558201+00:00

source checked

automated source review · 2026-09-16

Audit details

Task-specific computational methodology and field context checked in the cited primary-source artifact. Source-backed fields, inapplicable evaluator dimensions and unresolved details are distinguished. Numerical results were not reproduced.

Field: attributes.profile.facts.3.value

Source artifact SHA-256: 2968073753e6d44feff9c08b131edf23145e95b171434539dddf77bb92847033

Hash scope: Hash scope not separately documented; inspect source record

Inspected artifact

Leakage controls

The shared split controls partition differences between methods. Homology or construct-family exclusion is not specified in the checked downstream-task split statement.

Individual claims
CodonBERT large language model for mRNA vaccines

Original source ↗

Results: supervised learning datasets; cached text lines 14–17; task metric definitions and corresponding results table; matching task comparison table/ablation captions

Version: journal full text in PMC
Retrieved: 2026-09-16T10:38:57.558201+00:00

source checked

automated source review · 2026-09-16

Audit details

Task-specific computational methodology and field context checked in the cited primary-source artifact. Source-backed fields, inapplicable evaluator dimensions and unresolved details are distinguished. Numerical results were not reproduced.

Field: attributes.profile.facts.4.value

Source artifact SHA-256: 2968073753e6d44feff9c08b131edf23145e95b171434539dddf77bb92847033

Hash scope: Hash scope not separately documented; inspect source record

Inspected artifact

Uncertainty

The cited text-accessible evaluation sections give no confidence-interval, resampling or repeat-run error-bar specification. Image-only tables and uninspected supplements are outside this absence claim.

Individual claims
CodonBERT large language model for mRNA vaccines

Original source ↗

Results: supervised learning datasets; cached text lines 14–17; task metric definitions and corresponding results table; matching task comparison table/ablation captions

Version: journal full text in PMC
Retrieved: 2026-09-16T10:38:57.558201+00:00

unreported

automated source review · 2026-09-16

Audit details

Task-specific computational methodology and field context checked in the cited primary-source artifact. Source-backed fields, inapplicable evaluator dimensions and unresolved details are distinguished. Numerical results were not reproduced.

Field: attributes.profile.facts.5.value

Source artifact SHA-256: 2968073753e6d44feff9c08b131edf23145e95b171434539dddf77bb92847033

Hash scope: Hash scope not separately documented; inspect source record

Inspected artifact

Sources and history

Release 2026-09-17-d277315f7d76 · Record review: needs review

2 source records and release historyDownload this release
Technical metadata and extraction receipts

Stable ID: reported-task-1c74661df2c401

areas
rna-transcriptomes
tasks
flu-vaccine mRNA property prediction
entity level
task
version
Not reported
task
flu-vaccine mRNA property prediction
scope note
Paper-specific evaluation task; protocol completeness requires further extraction.
benchmark research
review date: 2026-09-17; status: primary_comparison_tables_located; primary sources: evidence-expansion-codonbert-vaccines-2024-29680737; inspected locators: Table 2.; XML table GR278870LITB2; searched queries: CodonBERT large language model for mRNA vaccines 10.1101/gr.278870.123; gaps: complete numerical transcription and independent cell review: Full primary artifact and table inventory preserved; no new numeric row is published from this audit alone.; exact checkpoint hashes and per-method scored denominators: Table labels alone do not establish these fields; do not infer checkpoint or scored count from model name or dataset size.; claim scope: Dated primary-source discovery and protocol/table screening. Source checking does not mean experimental reproduction. Only separately extracted and independently reviewed numeric batches are publishable.
historical missing metadata
protocol version: not_reported_in_legacy_extract
metadata review scope
historical_missing_metadata preserves the original discovery state. Current descriptive evidence and missingness are recorded in profile.facts; numerical-result review is separate.
legacy kinds
benchmark
entity classification
review date: 2026-09-17; rationale: This source-scoped record identifies the biological prediction task and holds its paper context. Preserve the existing task identity; exact split, model adaptation and scoring remain in linked evaluations or separate protocol records.; source ids: codonbert-vaccines-2024; source locator: Results: supervised learning datasets; cached text lines 14–17; task metric definitions and corresponding results table; matching task comparison table/ablation captions; ambiguities: A paper- or suite-specific task may constrain some inputs or metrics; that alone does not make it interchangeable with a complete versioned protocol. No protocol equivalence is inferred.; Some legacy profile Entity type facts use the generic phrase computational evaluation protocol. That boilerplate is not sufficient to establish a single fixed protocol identity or to merge this task with another protocol record.
Related records

Suggest a correction