LatchBio
Subscribe
Sign in
Home
Archive
About
Latest
Top
Discussions
VariantBench: An Agentic Benchmark for Genetic Variant Discovery and Interpretation
A verifiable benchmark for variant discovery, statistical genetics, and personal genomics
Jul 16
•
Kenny Workman
and
Christopher Zou
61
Why An Open Source Harness Outperforms Claude Code On Frontier Biology Tasks
In current generation models, behavioral priors introduce by harnesses such as Claude Code, and Codex cause substantial performance swings on our…
Jul 10
•
Joel Simonoff
44
3
1
Benchmarking AI Agents on Pathogen Genomic Surveillance
A verifiable benchmark for practical decisions about taxonomy, variants, AMR, source tracking, anomaly detection, and engineered sequences in pathogen…
Jul 9
•
Kenny Workman
,
Harmon Bhasin
,
Dianzhuo Wang
, and
Evan Seeyave
9
2
Benchmarking Refusals in Agentic Biology
A paired benchmark for capability and caution in agentic biosecurity risk assessment
Jul 1
•
Kenny Workman
,
Harmon Bhasin
, and
Dianzhuo Wang
74
June 2026
Latch MCP: Agent-Native Data Infrastructure
Access the Latch platform using Claude Code, Codex, or any agent
Jun 30
•
Anirudh Narsipur
,
Hannah Le
, and
Kenny Workman
8
LatchBio acquires TwentyTwo to form Latch Biosecurity
Welcoming Harmon Bhasin, Evan Seeyave, and John Wang to the team.
Jun 29
•
Alfredo Andere
73
2
Benchmarking AI Agents on Long-Horizon Single-Cell Biology
A verifiable benchmark for recovering scientific conclusions from raw single-cell data
Jun 25
•
Kenny Workman
83
1
1
Benchmarking AI Agents on Small-Molecule Preclinical Pharmacology
A verifiable benchmark for practical decisions about potency, mechanism, exposure, safety, and efficacy
Jun 17
•
Kenny Workman
73
1
EpiBench: AI agents still struggle with epigenomics analysis
Benchmarking frontier models on practical CUT&Tag/CUT&RUN, ATAC-seq, ChIP-seq, and DNA methylation workflows
Jun 11
•
Kenny Workman
49
May 2026
Human Verification of SpatialBench
Two rounds of independent expert attempts define a verified subset of 115 spatial biology tasks and expose ambiguity in benchmark specification and…
May 29
•
Kenny Workman
82
Verifiable Benchmarking of Long-Horizon Spatial Biology
Evaluating whether AI agents can recover complex scientific conclusions from raw spatial biology data
May 27
•
Kenny Workman
99
scBench Updates: Opus 4.7, GPT 5.5, Gemini 3.1
Benchmarking frontier models on messy, real-world single cell data analysis
May 12
•
Kenny Workman
90
This site requires JavaScript to run correctly. Please
turn on JavaScript
or unblock scripts