Skip to main content
Benchmark claim · developer-reported

AIDO Cell 1.0: Cell perturbation prediction

A structured benchmark claim is recorded; consult the linked source for numeric values and protocol details.

Model versionVersion history not yet curated
TaskCell perturbation prediction
DatasetVirtual Cell Benchmark 1.0
SplitSplit details not yet normalized
Metric31 metrics across five task families
Replicationunknown
Reported bySource authors
Review statuscurated
Evidence confidence · limited

Confidence is multidimensional, not a universal model score.

Evidence completenessstrongModel version, task, dataset, split, metric, source and provenance fields.
Independent validationunknownunknown
Source qualitylimitedDeveloper-reported; independent reproduction pending
Reproducibility evidenceunknownReflects documented replication status, not a universal reproducibility score.
Version specificitystrongVersion history not yet curated
Context applicabilitymoderateDepends on task, split and explicit caveats; users must still validate their own context.
Contradiction reviewclearNo direct contradiction signal is currently queued.

BioAtlas reports evidence dimensions separately so a strong source cannot hide weak applicability, incomplete replication or unresolved contradiction.

Why?

Why should this evidence influence a decision?

Why this evidence?

It is linked to a specific model version, scientific task, dataset, split, metric and source. That makes the claim inspectable rather than a detached marketing score.

Why not a universal score?

Performance can change with dataset, split, preprocessing, metric and context of use. BioAtlas therefore keeps confidence dimensions separate.

What could change the conclusion?

Independent replication, a better matched prospective dataset, a version change, a contradictory result or a more relevant validation protocol can reopen this evidence record.

Evidence boundary

What this claim does not prove.

  • Protocol, split and implementation details must match before comparing this claim with another result.

BioAtlas groups benchmark claims only when task, dataset, split, metric and protocol context align. This record is not a universal model score.

Model context

AIDO Cell 1.0 is GenBio AI's general-purpose virtual-cell simulator. It integrates multiple biological scales and modalities into a shared world-model framework and is presented for intervention-conditioned simulation across cell biology tasks. BioAtlas treats performance claims from Virtual Cell Benchmark 1.0 as developer-reported pending independent reproduction.

Full evidence passport →

Known model limitations

  • Benchmark leadership claims are developer-reported and should not be treated as independently established.
  • The released system currently demonstrates prototype virtual cells and does not constitute a solved general model of cell biology.
  • Wet-lab validation of novel predictions is still emerging and remains essential for scientific use.