
Illumina's Billion Cell Atlas alliance, a genome-wide genetic perturbation dataset launched in January 2026, has added three new members including Formation Bio, an AI-driven drug developer.
The alliance now holds over 350 million sequenced cells and upwards of six petabytes of genomic data, enabling drug makers to better understand how diseases originate, how candidate drugs interact with disease biology, and which patient populations are most likely to respond to treatments.
What happened
Illumina announced three new members for its Billion Cell Atlas alliance, including AI-native drug developer Formation Bio. The program, launched in January 2026 with founding members AstraZeneca, Merck, and Eli Lilly, has now sequenced over 350 million cells and generated upwards of six petabytes of genomic data.
Why it matters
The alliance gives drug developers unprecedented access to genome-wide perturbation data to identify which patient populations will respond to candidate drugs, validate drug targets, and make better decisions about which medicines to advance—potentially reducing risk and accelerating the path from biological discovery to approved drugs.
What to watch
Formation Bio intends to use the Atlas to credential therapeutic-area hypotheses and target-indication pairs, helping it choose which assets, indications, patients, and trial designs have the strongest probability of clinical success. The program will eventually capture how one billion individual cells respond to genetic changes via CRISPR across more than 200 disease-relevant cell lines.
Ask the AI about this article →
The Illumina Billion Cell Atlas represents a shift in how drug discovery leverages biological data at scale. Launched in January 2026, the program was built to address bottlenecks across drug development: identifying what targets to pursue, understanding why drugs succeed or fail, and selecting the right patient populations for clinical trials. The addition of Formation Bio and two other AI-driven companies signals confidence that machine-learning models trained on large, diverse biological datasets can improve the rigor of asset selection—a critical decision point where many first-in-class medicines fail because their target biology is compelling but clinical evidence is sparse.
Formation Bio's stated model—acquiring promising drugs close to or in the clinic and using AI to develop them faster—aligns directly with what the Atlas enables: using single-cell perturbation data to build more precise models of how candidate drugs interact with disease biology and to identify patient subgroups most likely to respond. By integrating the Atlas's cell-state-specific data with genetic, translational, and clinical evidence, developers can more rigorously credential target-indication pairs before investing in expensive clinical trials. This capability is particularly valuable for first-in-class programs, where the clinical precedent is limited and the stakes are high.
For example, today's edition would include:
AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.
Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. Q&As are published on this page for other readers too.
Israeli startup DataAgent Ltd
Taoyuan is positioning itself as a northern hub for AI data centers (AIDC), citing the Tatan area and an LNG c…

SK Hynix presented a custom HBM concept at SEMICON Taiwan 2026, where compute functions are placed in the base…

The U.S. Department of Defense announced on August 31 that it has deployed ChatGPT Mil, a customized version o…

Nvidia reported earnings that were both remarkable and boring, reflecting its focus on avoiding a consolidated…

Anthropic has agreed to a $35bn cloud-computing contract with Lambda, a Nvidia-backed cloud provider
