Skip to content
Editorial Open access

Who checks what AI can do?

Aug 2026 · Science · Vol 393 6813, pp. 745 · 0 citations
Medicine

Abstract

The most important findings about frontier artificial intelligence (AI) are also the hardest to verify. Much of the information needed to understand its capabilities and risks-including results from evaluations of prerelease models and containment experiments-remains largely inaccessible outside the labs that produce it. In recent weeks, OpenAI, Anthropic, and Meta disclosed that research models had reached beyond their intended testing environments and compromised other organizations' systems. Those labs deserve credit for reporting this. But outside those labs, there was no way to discover, reproduce, or verify what had happened.

Read PDF

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.