TxBench: Antibody Discovery (13 minute read)
Summary
TxBench is a benchmark for antibody discovery, likely assessing computational or AI-driven methods in this biotech domain.
View Cached Full Text
Cached at: 09/03/26, 11:55 PM
Similar Articles
TxBench-PP: Analyzing AI Agent Performance on Small-Molecule Preclinical Pharmacology
TxBench-PP is a benchmark for evaluating AI agents on small-molecule preclinical pharmacology tasks. Across 16 model-harness configurations, the best system achieved only 59.3% accuracy, indicating significant room for improvement.
EpiBench: Can LLMs Understand Epitopes for Antibody Drug Discovery?
EpiBench is a new closed-book, sequence-based benchmark for evaluating how well LLMs understand epitopes across five antibody-drug-discovery tasks, finding that current models capture partial signals but struggle with antibody-specific reasoning.
Introducing GeneBench-Pro
OpenAI introduces GeneBench-Pro, a research-level benchmark designed to test AI agents' ability to perform judgment-heavy analyses in computational biology, covering genomics, quantitative biology, and translational medicine.
PhenoBench: Mapping What a Deeply Phenotyped Human Cohort Can Tell Us
PhenoBench is an executable benchmark for evaluating AI models on clinical tasks using a deeply phenotyped human cohort, demonstrating limited aggregate improvements of foundation models over standard methods.
BixBench3: Frontier AI Agents Can Now Reproduce ~48% of Real Computational Biology Research Workflows
This paper introduces BixBench3, a benchmark for evaluating AI agents on computational biology tasks, revealing that frontier LLMs can reproduce approximately 48% of real research workflows but struggle with large datasets and sequential steps.