Chai Discovery: Scaling Drug Design as a Software Problem

Chai Discovery's co-founders discuss their approach to AI-driven drug design, emphasizing simplicity, scaling laws, and the transformation of biology into an engineering discipline.

9 min read
Four people sitting in chairs in a discussion setting, microphones present.
Sequoia Capital

Visual TL;DR. Complex Drug Discovery addressed by Chai Discovery. Chai Discovery uses approach Software Problem. Software Problem guided by Bitter Lesson. Software Problem enables AI-Driven Design. AI-Driven Design leads to Predictable Engineering. Chai Discovery requires Avengers Squad.

  1. Complex Drug Discovery: historically complex, relying on serendipity and experimentation for new drugs
  2. Chai Discovery: co-founders Josh and Matt transforming biology into an engineering discipline
  3. Software Problem: treating molecule engineering as a scalable software problem, like LLMs
  4. Bitter Lesson: emphasizing simplicity and the power of scaling data, compute, and models
  5. AI-Driven Design: developing abstraction layers for rapid iteration and faster drug development
  6. Predictable Engineering: transforming drug discovery from trial-and-error to a predictable process
  7. Avengers Squad: building a multidisciplinary team to overcome the scary nature of biology
Visual TL;DR
Visual TL;DR, startuphub.ai Complex Drug Discovery addressed by Chai Discovery. Chai Discovery uses approach Software Problem. Software Problem enables AI-Driven Design. AI-Driven Design leads to Predictable Engineering addressed by uses approach enables leads to Complex Drug Discovery Chai Discovery Software Problem AI-Driven Design Predictable Engineering From startuphub.ai · The publishers behind this format
Visual TL;DR, startuphub.ai Complex Drug Discovery addressed by Chai Discovery. Chai Discovery uses approach Software Problem. Software Problem enables AI-Driven Design. AI-Driven Design leads to Predictable Engineering addressed by uses approach enables leads to Complex DrugDiscovery Chai Discovery Software Problem AI-Driven Design PredictableEngineering From startuphub.ai · The publishers behind this format
Visual TL;DR, startuphub.ai Complex Drug Discovery addressed by Chai Discovery. Chai Discovery uses approach Software Problem. Software Problem enables AI-Driven Design. AI-Driven Design leads to Predictable Engineering addressed by uses approach enables leads to Complex Drug Discovery historically complex, relying onserendipity and experimentation for newdrugs Chai Discovery co-founders Josh and Matt transformingbiology into an engineering discipline Software Problem treating molecule engineering as ascalable software problem, like LLMs AI-Driven Design developing abstraction layers for rapiditeration and faster drug development Predictable Engineering transforming drug discovery fromtrial-and-error to a predictable process From startuphub.ai · The publishers behind this format
Visual TL;DR, startuphub.ai Complex Drug Discovery addressed by Chai Discovery. Chai Discovery uses approach Software Problem. Software Problem enables AI-Driven Design. AI-Driven Design leads to Predictable Engineering addressed by uses approach enables leads to Complex DrugDiscovery historicallycomplex, relying onserendipity and… Chai Discovery co-founders Joshand Matttransforming… Software Problem treating moleculeengineering as ascalable software… AI-Driven Design developingabstraction layersfor rapid iteration… PredictableEngineering transforming drugdiscovery fromtrial-and-error to… From startuphub.ai · The publishers behind this format
Visual TL;DR, startuphub.ai Complex Drug Discovery addressed by Chai Discovery. Chai Discovery uses approach Software Problem. Software Problem guided by Bitter Lesson. Software Problem enables AI-Driven Design. AI-Driven Design leads to Predictable Engineering. Chai Discovery requires Avengers Squad addressed by uses approach guided by enables leads to requires Complex Drug Discovery historically complex, relying onserendipity and experimentation for newdrugs Chai Discovery co-founders Josh and Matt transformingbiology into an engineering discipline Software Problem treating molecule engineering as ascalable software problem, like LLMs Bitter Lesson emphasizing simplicity and the power ofscaling data, compute, and models AI-Driven Design developing abstraction layers for rapiditeration and faster drug development Predictable Engineering transforming drug discovery fromtrial-and-error to a predictable process Avengers Squad building a multidisciplinary team toovercome the scary nature of biology From startuphub.ai · The publishers behind this format
Visual TL;DR, startuphub.ai Complex Drug Discovery addressed by Chai Discovery. Chai Discovery uses approach Software Problem. Software Problem guided by Bitter Lesson. Software Problem enables AI-Driven Design. AI-Driven Design leads to Predictable Engineering. Chai Discovery requires Avengers Squad addressed by uses approach guided by enables leads to requires Complex DrugDiscovery historicallycomplex, relying onserendipity and… Chai Discovery co-founders Joshand Matttransforming… Software Problem treating moleculeengineering as ascalable software… Bitter Lesson emphasizingsimplicity and thepower of scaling… AI-Driven Design developingabstraction layersfor rapid iteration… PredictableEngineering transforming drugdiscovery fromtrial-and-error to… Avengers Squad building amultidisciplinaryteam to overcome… From startuphub.ai · The publishers behind this format

In the rapidly evolving field of AI-driven drug discovery, Chai Discovery is carving out a unique niche by treating molecule engineering as a scalable software problem. Co-founders Josh and Matt sat down to discuss their 'bitter lesson' philosophy, which emphasizes the power of scaling data, compute, and models, drawing parallels to the breakthroughs seen in large language models. Their goal is to transform drug discovery from a trial-and-error process into a more predictable, engineering-like discipline.

Chai Discovery: Scaling Drug Design as a Software Problem - Sequoia Capital
Chai Discovery: Scaling Drug Design as a Software Problem — from Sequoia Capital

Engineering Biology with AI

Matt explained Chai's core mission: to make drug discovery look more like engineering. He highlighted the success of LLMs in code generation, attributing it to code's simple abstraction. Biology, however, has historically been more complex, relying heavily on serendipity and experimentation. Chai aims to industrialize this process by developing the necessary abstraction layers, akin to those in modern software engineering, to enable rapid iteration and faster development cycles in biology.

The conversation touched upon the historical progression of AI in biology, noting significant leaps like AlphaFold's impact on protein folding prediction. This evolution has paved the way for new sub-problems like designing protein sequences that fold into specific structures or perform particular functions. The advent of diffusion models has been particularly impactful, allowing for the simultaneous generation of protein structures and sequences, and enabling more realistic prompts that incorporate real-world constraints.

The "Bitter Lesson" of Simplicity

Josh emphasized Chai's guiding principle of simplicity, contrasting it with earlier models like Chai 1, which featured 23 distinct sub-modules. He explained that iterating on such complex systems becomes challenging, as understanding each module's behavior independently doesn't scale well. By simplifying and identifying core important components, Chai aims to streamline the research process and identify effective scaling directions.

From Protein Folding to Drug Design

The discussion traced the evolution of AI in biology, starting with protein folding competitions and progressing to more intricate tasks like designing proteins with specific functions. The key breakthrough, according to Matt, was the emergence of diffusion models, which enabled the simultaneous generation of protein structures and sequences. This advancement allows for more sophisticated prompts, such as specifying a target protein shape and incorporating real-world constraints.

The Significance of 2024 for Chai's Launch

The co-founders identified 2024 as a pivotal year, fueled by advancements that made antibody design feasible. Previously, the complexity and data requirements for antibody folding were considered insurmountable. Chai's progress, particularly with diffusion models, demonstrated that predicting and designing antibodies computationally was within reach. They noted that if one cannot predict an antibody's structure, designing one effectively is an uphill battle.

Overcoming the "Scary" Nature of Biology

Both co-founders admitted to initially being intimidated by biology, with Matt coming from a background in theoretical computer science and pure math. However, they found that the underlying problems in biology are more interconnected and simpler than they might appear. By reframing these challenges as sequences of amino acids that can be represented and manipulated by models, they demystified the process.

Building a Multidisciplinary "Avengers Squad"

Chai's success hinges on assembling a diverse team of experts from chemistry, biology, and AI. They've adopted a pragmatic approach, initially focusing on AI researchers and gradually expanding to include top-tier antibody engineers and scientists. The company's growth strategy involves bringing in individuals who have a proven track record in building exciting products, ensuring that the powerful AI models are translated into user-friendly and impactful applications.

The Power of Iteration and "Dogfooding"

Chai's approach emphasizes continuous iteration and leveraging internal use cases, or "dogfooding," to refine their models. By using their own AI tools to generate data and test hypotheses, they gain valuable insights that feedback into model improvement. This iterative cycle, similar to that seen with LLMs, allows them to push the boundaries of what's possible in molecule design.

Achieving High Success Rates in Molecule Generation

A critical milestone for Chai was achieving high success rates in de novo molecule generation. They highlighted that early state-of-the-art methods for antibody design had a binding rate of only 0.1%. Chai's models have significantly improved this, reaching a 15% success rate with their Chai 2 model. This increased accuracy provides richer statistical data for property analysis and enables more efficient optimization, ultimately leading to the ability to bake desired properties into molecules from the outset.

The "Bitter Lesson" Applied to Scaling Laws

The company's philosophy is rooted in the "bitter lesson" of machine learning: relying on scaling laws for compute, data, and models is crucial for progress. This principle is applied to biology by identifying how to tokenize and represent biological data effectively, ensuring that the models can generalize and improve with scale. The emphasis on simplicity in their model architecture is key to achieving this scalability.

Verifiability and Rigor in Biology

While biology can seem less verifiable than domains like code generation, Chai emphasizes that it is, in fact, a highly objective field. The ability to obtain specific readouts from lab experiments, even if they take longer than typical software tests, allows for honest self-assessment and rigorous validation of model progress. This rigor is essential for building products that genuinely advance the field.

Unlocking Novel Biology and Computer-Aided Design

Chai's focus on de novo generation and their belief in scaling laws suggest that they can unlock targets previously considered undruggable. They aim to build a comprehensive computer-aided design suite for molecules, allowing for rapid iteration from idea to testable hypothesis, ultimately accelerating the discovery of better medicines.

© 2026 StartupHub.ai. All rights reserved. Do not enter, scrape, copy, reproduce, or republish this article in whole or in part. Use as input to AI training, fine-tuning, retrieval-augmented generation, or any machine-learning system is prohibited without written license. Substantially-similar derivative works will be pursued to the fullest extent of applicable copyright, database, and computer-misuse laws. See our terms.