Measuring Biological Capabilities and Risks of AI Agents: Generating and Interpreting Evidence from Agentic Evaluations

Patricia Paskov, Jeffrey Lee, Kyle Brady, Alyssa M. Worland · RAND Corporation eBooks · 2026

This perspective examines biological agentic evaluations as an emerging tool for assessing the capabilities and risks of autonomous AI systems in biological contexts. Drawing on hands-on evaluation experience, it offers practical guidance on defining, designing, running, scoring, and interpreting evaluations, highlighting how design choices shape conclusions and policy relevance.

Read the paper · More papers on PaperTik