arXiv CS AI
techCenter
Agentic BAIM-LLM Evaluation (ABLE): Benchmarking LLM Use of Protein Design Toolstranslating…
arXiv:2609.05818v1 Announce Type: new
Abstract: We introduce ABLE, a benchmark for evaluating LLM agents' ability to use biological AI models (BAIMs), such as ProteinMPNN and AlphaFold3, in dual-use protein design workflows. ABLE assesses agent performance through a set of…
Keywords#ABLE#Protein Design#LLM#Agentic#Agentic BAIM-LLM