Kenny Workman on Building Verifiable AI Evals for Biology
LatchBio CTO Kenny Workman explains how verifiable evaluation frameworks and multi-omics benchmarks drive AI agent progress in biological research.

Visual TL;DR
modern techniques generate massive, complex datasets at an unprecedented rate
From the article 3 mentionsBiological research produces vast volumes of complex data every day.
co-founder and CTO building data infrastructure for biotech and pharma
From the article 2 mentionsAt the AI Engineer World's Fair, Kenny Workman, co-founder and CTO of LatchBio, detailed how his team builds verifiable evaluation environments to train and test AI agents in life sciences.
frameworks to train and test AI agents in life sciences research
From the article 3 mentions"Just like code provided a verifiable substrate for complex software tasks that are not inherently verifiable, data analysis might do the same thing in bio," Workman noted.
a specific benchmarking framework for measuring agent performance on tasks
From the articleTo solve this, LatchBio built SpatialBench, an evaluation suite containing 146 verifiable problems derived from real spatial biology workflows.
long-horizon benchmarks for complex tasks across genomics and drug discovery
From the article 5 mentionsBy breaking scientific research down into data processing steps, developers can systematically benchmark model performance.
From the article 2 mentionsWorkman explained how treating biological data analysis like code execution creates a natural path to improve AI reasoning across genomics and drug discovery.
driving AI agent progress in biological research and drug discovery
From the article 2 mentionsOver five years, LatchBio evolved from an enterprise data management provider into a specialized research lab for biological AI agents.
considering model refusals and biosecurity implications in AI development
From the article 2 mentionsAs AI capabilities expand, assessing safety and biosecurity risks becomes critical.
Contents(6)
© 2026 StartupHub.ai. All rights reserved. You may not republish this article in full without a license. Search engines and AI research tools may crawl and summarize for reference. Bulk reproduction or model training requires a license. See our terms.
Written by
Daniel SingerEditor, StartupHub.ai
Daniel Singer is the editor of StartupHub.ai, a technology expert and thought leader on AI and its applications across sectors, from fintech and healthcare to developer tooling and consumer software. He writes and tests the tools covered here thoroughly and regularly, and built StartupHub.ai to give founders, operators and buyers a clearer read on what they are actually being sold.