Toward Comprehensive Benchmarking of the Biological Knowledge of Frontier Large Language ModelsRAND Corporation2025-12-1567 页
AI Agents in Action: Foundations for Evaluation and Governance人工智能/ChatGPT/AIGC/生成式AI世界经济论坛(World Economic Forum)2025-12-1534 页
Manipulating Minds: Security Implications of AI-Induced Psychosis人工智能/ChatGPT/AIGC/生成式AIRAND Corporation2025-12-2159 页