Strengths supported by sources
- Abundance agreement and taxon detection are separate evaluation dimensions.
Sources
CAMI-challenge/OPAL official source · Pinned README: introduction; Computed metrics; input format; example evaluations
The CAMI taxonomic-profiling task can be understood through its documented assessment tool; this guide does not identify a challenge-specific run.
Explanatory profile: source reviewed · Automated source review, 2026-09-16. Review applies to the cited claims; unresolved fields are listed below. Numerical results retain their own review status.
| Property | Description and evidence |
|---|---|
| Datasets | A selected CAMI challenge dataset and matching gold standard are required; this entry does not fix the edition.SourcesCAMI-challenge/OPAL official source · Pinned README: introduction; Computed metrics; input format; example evaluations |
| Splits | CAMI II supplied public-genome practice datasets with ground truth before its blinded challenge. Challenge datasets were marine, strain-madness and plant-associated communities; these are challenge conditions, not a standard supervised train/validation/test partition.Sourcescami2 primary benchmark evidence · Methods: Challenge datasets, Challenge organization, Evaluation metrics; Figures 2–4; Table 1 |
| Metrics | Precision, recall, F1, Jaccard, L1 error, UniFrac, Bray–Curtis and diversity measures.SourcesCAMI-challenge/OPAL official source · Pinned README: introduction; Computed metrics; input format; example evaluations |
| Baselines | Submitted programs are compared under the same data condition. Gold-standard assemblies and MEGAHIT assemblies separate binning performance from upstream assembly error; published method identities and versions are listed in Table 1.Sourcescami2 primary benchmark evidence · Methods: Challenge datasets, Challenge organization, Evaluation metrics; Figures 2–4; Table 1 |
| Leakage controls | Challenge genome data and metadata were kept confidential until the challenge ended. Public reference collections dated 8 January 2019 were supplied for reference-based methods. CAMI II also includes public genomes, so novelty is stratified rather than assumed for every organism.Sourcescami2 primary benchmark evidence · Methods: Challenge datasets, Challenge organization, Evaluation metrics; Figures 2–4; Table 1 |
| Uncertainty | Uncertainty is task-specific: taxonomic binning Figure 3 uses standard errors across bins; taxonomic profiling Figure 4 reports means across samples with standard deviations. These are not a common seed-based interval for every CAMI metric.Sourcescami2 primary benchmark evidence · Figure 3 and Figure 4 captions: standard error across taxonomic bins versus standard deviation across samples |
| Entity type | Constituent benchmark task: CAMI taxonomic profilingSourcesCAMI-challenge/OPAL official source · Pinned README: introduction; Computed metrics; input format; example evaluations |
| Organisms | Taxa are supplied by the selected reference community. · Not applicableSourcesCAMI-challenge/OPAL official source · Pinned README: introduction; Computed metrics; input format; example evaluations |
| Assays | Challenge-specific metagenomic sequence data and reference composition.SourcesCAMI-challenge/OPAL official source · Pinned README: introduction; Computed metrics; input format; example evaluations |
| Allowed inputs | Predicted and reference abundance profiles.SourcesCAMI-challenge/OPAL official source · Pinned README: introduction; Computed metrics; input format; example evaluations |
| Adaptation | OPAL scores profiles; it does not train the submitting method. · Not applicableSourcesCAMI-challenge/OPAL official source · Pinned README: introduction; Computed metrics; input format; example evaluations |
Conceptual procedure. Task variants and protocol versions retain their separate scoring conditions.
CAMI is a series of blinded metagenomic software challenges. In CAMI II, participants received simulated short and long reads from defined communities and submitted assemblies, genome bins, taxonomic assignments or abundance profiles. Reference truth was used only for scoring, and software versions, input read types and community conditions were kept distinct.
Benchmarks bring together tasks and protocols. A task describes the biological question; a protocol defines a particular test.
These source-backed links do not make different protocols or scores interchangeable.
Release 2026-09-17-d277315f7d76 · 0 evaluations · 0 metric rows. Different protocols are not a single leaderboard.
No evaluations linked in this release.
Last literature check: 2026-09-17. Primary-source discovery and table/protocol screening; source checked is not independently reproduced. Raw acquisitions not automatically numerical publication approval.
| Paper or primary resource | Version | Reference |
|---|---|---|
| Critical Assessment of Metagenome Interpretation: the second round of challenges | PMC9007738 | Read source DOI: 10.1038/s41592-022-01431-4 |
source found structured extraction pending
Primary paper and/or task implementation reviewed for the explicitly cited methodology claims. Scope-limited absence is recorded only after the documented source search; no model runs or independent reproduction. Independent automated spot review corrected interval terminology and sequence-identity scope against the original figure captions and methods.
Stable record: discovery-benchmark-cami-taxonomic-profilingApplicability is distinct from a completed evaluation.
Trace each statement to its source and review. A context-only reference supports the record generally; it does not verify an individual field. Source checking does not reproduce an experiment.
One row per statement and cited source. Multiple citations are not independent evaluations. Shared locators are labelled explicitly.
22 evidence rows matching the loaded filters
| Property and statement | Original source and location | Review and provenance |
|---|---|---|
| Diagram caption Conceptual procedure. Task variants and protocol versions retain their separate scoring conditions. Individual claims | cami2 primary benchmark evidence Pinned README: introduction; Computed metrics; input format; example evaluations; Methods: Challenge datasets, Challenge organization, Evaluation metrics; Figures 2–4; Table 1 Shared locator for this statement’s cited sources; not a separate locator for each citation. Version: PMC9007738 | source checked automated source review · 2026-09-16 Audit detailsPrimary paper and/or task implementation reviewed for the explicitly cited methodology claims. Scope-limited absence is recorded only after the documented source search; no model runs or independent reproduction. Independent automated spot review corrected interval terminology and sequence-identity scope against the original figure captions and methods. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| Diagram caption Conceptual procedure. Task variants and protocol versions retain their separate scoring conditions. Individual claims | CAMI-challenge/OPAL official source Pinned README: introduction; Computed metrics; input format; example evaluations; Methods: Challenge datasets, Challenge organization, Evaluation metrics; Figures 2–4; Table 1 Shared locator for this statement’s cited sources; not a separate locator for each citation. Version: 98120c326eef08e391899e4bd3a362e0e6558b4a | source checked automated source review · 2026-09-16 Audit detailsPrimary paper and/or task implementation reviewed for the explicitly cited methodology claims. Scope-limited absence is recorded only after the documented source search; no model runs or independent reproduction. Independent automated spot review corrected interval terminology and sequence-identity scope against the original figure captions and methods. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| Diagram steps ["Allowed inputs: Predicted and reference abundance profiles.","Splits: CAMI II supplied public-genome practice datasets with ground truth before its blinded challenge. Challenge datasets were marine, strain-madness and plant-associated communities; these are challenge conditions, not a standard supervised train/validation/test partition.","Metrics: Precision, recall, F1, Jaccard, L1 error, UniFrac, Bray–Curtis and diversity measures."] Individual claims | cami2 primary benchmark evidence Pinned README: introduction; Computed metrics; input format; example evaluations; Methods: Challenge datasets, Challenge organization, Evaluation metrics; Figures 2–4; Table 1 Shared locator for this statement’s cited sources; not a separate locator for each citation. Version: PMC9007738 | source checked automated source review · 2026-09-16 Audit detailsPrimary paper and/or task implementation reviewed for the explicitly cited methodology claims. Scope-limited absence is recorded only after the documented source search; no model runs or independent reproduction. Independent automated spot review corrected interval terminology and sequence-identity scope against the original figure captions and methods. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| Diagram steps ["Allowed inputs: Predicted and reference abundance profiles.","Splits: CAMI II supplied public-genome practice datasets with ground truth before its blinded challenge. Challenge datasets were marine, strain-madness and plant-associated communities; these are challenge conditions, not a standard supervised train/validation/test partition.","Metrics: Precision, recall, F1, Jaccard, L1 error, UniFrac, Bray–Curtis and diversity measures."] Individual claims | CAMI-challenge/OPAL official source Pinned README: introduction; Computed metrics; input format; example evaluations; Methods: Challenge datasets, Challenge organization, Evaluation metrics; Figures 2–4; Table 1 Shared locator for this statement’s cited sources; not a separate locator for each citation. Version: 98120c326eef08e391899e4bd3a362e0e6558b4a | source checked automated source review · 2026-09-16 Audit detailsPrimary paper and/or task implementation reviewed for the explicitly cited methodology claims. Scope-limited absence is recorded only after the documented source search; no model runs or independent reproduction. Independent automated spot review corrected interval terminology and sequence-identity scope against the original figure captions and methods. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| Diagram title Evaluation procedure Individual claims | cami2 primary benchmark evidence Pinned README: introduction; Computed metrics; input format; example evaluations; Methods: Challenge datasets, Challenge organization, Evaluation metrics; Figures 2–4; Table 1 Shared locator for this statement’s cited sources; not a separate locator for each citation. Version: PMC9007738 | source checked automated source review · 2026-09-16 Audit detailsPrimary paper and/or task implementation reviewed for the explicitly cited methodology claims. Scope-limited absence is recorded only after the documented source search; no model runs or independent reproduction. Independent automated spot review corrected interval terminology and sequence-identity scope against the original figure captions and methods. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| Diagram title Evaluation procedure Individual claims | CAMI-challenge/OPAL official source Pinned README: introduction; Computed metrics; input format; example evaluations; Methods: Challenge datasets, Challenge organization, Evaluation metrics; Figures 2–4; Table 1 Shared locator for this statement’s cited sources; not a separate locator for each citation. Version: 98120c326eef08e391899e4bd3a362e0e6558b4a | source checked automated source review · 2026-09-16 Audit detailsPrimary paper and/or task implementation reviewed for the explicitly cited methodology claims. Scope-limited absence is recorded only after the documented source search; no model runs or independent reproduction. Independent automated spot review corrected interval terminology and sequence-identity scope against the original figure captions and methods. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| Datasets A selected CAMI challenge dataset and matching gold standard are required; this entry does not fix the edition. Individual claims | CAMI-challenge/OPAL official source Pinned README: introduction; Computed metrics; input format; example evaluations Version: 98120c326eef08e391899e4bd3a362e0e6558b4a | source checked automated source review · 2026-09-16 Audit detailsPrimary paper and/or task implementation reviewed for the explicitly cited methodology claims. Scope-limited absence is recorded only after the documented source search; no model runs or independent reproduction. Independent automated spot review corrected interval terminology and sequence-identity scope against the original figure captions and methods. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| Splits CAMI II supplied public-genome practice datasets with ground truth before its blinded challenge. Challenge datasets were marine, strain-madness and plant-associated communities; these are challenge conditions, not a standard supervised train/validation/test partition. Individual claims | cami2 primary benchmark evidence Methods: Challenge datasets, Challenge organization, Evaluation metrics; Figures 2–4; Table 1 Version: PMC9007738 | source checked automated source review · 2026-09-16 Audit detailsPrimary paper and/or task implementation reviewed for the explicitly cited methodology claims. Scope-limited absence is recorded only after the documented source search; no model runs or independent reproduction. Independent automated spot review corrected interval terminology and sequence-identity scope against the original figure captions and methods. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| Adaptation OPAL scores profiles; it does not train the submitting method. Individual claims | CAMI-challenge/OPAL official source Pinned README: introduction; Computed metrics; input format; example evaluations Version: 98120c326eef08e391899e4bd3a362e0e6558b4a | inapplicable automated source review · 2026-09-16 Audit detailsPrimary paper and/or task implementation reviewed for the explicitly cited methodology claims. Scope-limited absence is recorded only after the documented source search; no model runs or independent reproduction. Independent automated spot review corrected interval terminology and sequence-identity scope against the original figure captions and methods. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
| Metrics Precision, recall, F1, Jaccard, L1 error, UniFrac, Bray–Curtis and diversity measures. Individual claims | CAMI-challenge/OPAL official source Pinned README: introduction; Computed metrics; input format; example evaluations Version: 98120c326eef08e391899e4bd3a362e0e6558b4a | source checked automated source review · 2026-09-16 Audit detailsPrimary paper and/or task implementation reviewed for the explicitly cited methodology claims. Scope-limited absence is recorded only after the documented source search; no model runs or independent reproduction. Independent automated spot review corrected interval terminology and sequence-identity scope against the original figure captions and methods. Field: Source artifact SHA-256: Hash scope: Hash scope not separately documented; inspect source record |
Release 2026-09-17-d277315f7d76 · Record review: discovered
Stable ID: discovery-benchmark-cami-taxonomic-profiling