We Can Detect New Pathogens in Days. Can AI Help Us Understand Them?
AI Summary: Researchers have introduced BioSecBench-Function, a benchmark for testing whether AI agents can infer the functional properties of biological threats, such as viruses, bacteria, and toxins. The benchmark evaluates AI agents across five threat axes, including transmissibility and drug resistance, and assesses their ability to interpret evidence and report conclusions. The results show that current AI agents struggle with this task, with a top pass rate of 50.3% and significant variation in performance across different biological functions and organisms. The benchmark aims to facilitate the development of AI agents that can support biodefense efforts by rapidly characterizing novel pathogens.