Skip to page content
Toggle Menu
Site-wide navigation
Topics
Trending
International Trade
Biotechnology
Education Policy
Ukraine
Artificial Intelligence
Topics
Children, Families, and Communities
Cyber and Data Sciences
Education and Literacy
Energy and Environment
Health, Health Care, and Aging
Homeland Security and Public Safety
Infrastructure and Transportation
International Affairs
Law and Business
National Security
Science and Technology
Workers and the Workplace
All Topics
Research & Commentary
Experts
About
Research Divisions
RAND's divisions conduct research on a uniquely broad front for clients around the globe.
U.S. research divisions
RAND Army Research Division
RAND Education, Employment, and Infrastructure
RAND Global and Emerging Risks
RAND Health
RAND Homeland Security Research Division
RAND National Security Research Division
RAND Project AIR FORCE
International research divisions
RAND Australia
RAND Europe
Services & Impact
Careers
Graduate School
Subscribe
Give
Cart
Toggle
Search
Search terms
Submit
RAND
Research & Commentary
Browse by Author
D
Sunishchal Dev — Publications
Sunishchal Dev — Publications
View Profile »
Can LLM Agents Select and Engage with Biological Tools? An Initial Biosecurity Assessment
2026
Interpreting Dual-Use Biology Benchmarks for Frontier AI: Measurement, Saturation, and Generational Analysis
2026
Judge Reliability Harness
2026
Open-Weight AI Models Require Proportional Evaluation Approaches
2026
The Science and Practice of Proportionality in AI Risk Evaluations: AI Evaluations Should Provide Meaningful Risk Information Without Imposing Excessive Burden
2026
Simpler Is Better for Autograders: Toward Cost-Effective LLM Evaluations for Open-Ended Tasks
2026
Position: Human Baselines in Model Evaluations Need Rigor and Transparency: (With Recommendations & Reporting Checklist)
2025
Toward Comprehensive Benchmarking of the Biological Knowledge of Frontier Large Language Models
2025