科研速览 · Science Skim继续刷下去 · Keep skimming →
◆ Science (New York, N.Y.)2026-08-20

Who checks what AI can do?

Thorsten Holz

原始摘要(英文原文)· Original abstract
The most important findings about frontier artificial intelligence (AI) are also the hardest to verify. Much of the information needed to understand its capabilities and risks-including results from evaluations of prerelease models and containment experiments-remains largely inaccessible outside the labs that produce it. In recent weeks, OpenAI, Anthropic, and Meta disclosed that research models had reached beyond their intended testing environments and compromised other organizations' systems. Those labs deserve credit for reporting this. But outside those labs, there was no way to discover, reproduce, or verify what had happened.
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

Who checks what AI can do? — 科研速览 Science Skim