EpiBench: Can LLMs Understand Epitopes for Antibody Drug Discovery?

arXiv cs.CL Papers

Summary

EpiBench is a new closed-book, sequence-based benchmark for evaluating how well LLMs understand epitopes across five antibody-drug-discovery tasks, finding that current models capture partial signals but struggle with antibody-specific reasoning.

arXiv:2608.06022v1 Announce Type: new Abstract: Epitopes determine where antibodies bind antigens and shape downstream therapeutic properties such as functional blockade and escape resistance, making epitope understanding central to antibody drug discovery. Although large language models (LLMs) have shown strong biomedical reasoning ability, it remains unclear whether they can infer epitope information directly from antigen and antibody sequences. Existing epitope resources typically focus on isolated prediction tasks or rely on specialized structural settings, while general protein benchmarks do not evaluate epitope-centered decisions across the antibody development workflow. To address this gap, we introduce EpiBench, a closed-book, sequence-based, and automatically scorable benchmark for evaluating epitope reasoning in LLMs. EpiBench contains 1,609 curated samples grounded in structural antibody--antigen contacts, curated functional B-cell assays, and deep mutational scanning escape measurements. It covers five connected tasks: targetable region discovery, antibody-conditioned epitope identification, epitope binning, functional epitope assessment, and antibody escape assessment, with controlled sampling to reduce shortcut-based evaluation artifacts. We evaluate nine general-purpose LLMs and analyze their behavior through task-specific baselines, antigen length stratification, explicit-reasoning comparison, and failure-mode inspection. The results show that current LLMs capture partial epitope-related signals but remain limited in antibody-specific sequence grounding, long-context residue localization, and biologically grounded reasoning. Therefore, EpiBench provides a diagnostic testbed for measuring and improving sequence-aware biomedical LLMs toward reliable LLM-assisted antibody discovery.
Original Article
View Cached Full Text

Cached at: 08/07/26, 07:52 AM

# EpiBench: Can LLMs Understand Epitopes for Antibody Drug Discovery?
Source: [https://arxiv.org/abs/2608.06022](https://arxiv.org/abs/2608.06022)
[View PDF](https://arxiv.org/pdf/2608.06022)

> Abstract:Epitopes determine where antibodies bind antigens and shape downstream therapeutic properties such as functional blockade and escape resistance, making epitope understanding central to antibody drug discovery\. Although large language models \(LLMs\) have shown strong biomedical reasoning ability, it remains unclear whether they can infer epitope information directly from antigen and antibody sequences\. Existing epitope resources typically focus on isolated prediction tasks or rely on specialized structural settings, while general protein benchmarks do not evaluate epitope\-centered decisions across the antibody development workflow\. To address this gap, we introduce EpiBench, a closed\-book, sequence\-based, and automatically scorable benchmark for evaluating epitope reasoning in LLMs\. EpiBench contains 1,609 curated samples grounded in structural antibody\-\-antigen contacts, curated functional B\-cell assays, and deep mutational scanning escape measurements\. It covers five connected tasks: targetable region discovery, antibody\-conditioned epitope identification, epitope binning, functional epitope assessment, and antibody escape assessment, with controlled sampling to reduce shortcut\-based evaluation artifacts\. We evaluate nine general\-purpose LLMs and analyze their behavior through task\-specific baselines, antigen length stratification, explicit\-reasoning comparison, and failure\-mode inspection\. The results show that current LLMs capture partial epitope\-related signals but remain limited in antibody\-specific sequence grounding, long\-context residue localization, and biologically grounded reasoning\. Therefore, EpiBench provides a diagnostic testbed for measuring and improving sequence\-aware biomedical LLMs toward reliable LLM\-assisted antibody discovery\.

## Submission history

From: Jiaqi Wang \[[view email](https://arxiv.org/show-email/d9b8b306/2608.06022)\] **\[v1\]**Thu, 6 Aug 2026 13:29:53 UTC \(1,307 KB\)

Similar Articles