UK AISI

independent testing powerful ai models a university clock tower with blank clock face v2

Q&A: Researcher Calls for Independent Testing of Powerful AI Models

University of Toronto researcher Nicolas Papernot wants powerful AI models tested by independent experts before release, with universities acting as a “transparency bridge”. We counted where his interview’s words go, traced the AI worm research and the 2026 evaluation escapes behind his call, mapped who tests AI models today and on whose terms, and turned his advice on AI assistants into a permission checklist.

Read more
anthropic rogue ai agents hate captchas just like you a toy crocodile resting on four short legs

Anthropic Reveals Rogue AI Agents Hate CAPTCHAs, Just Like You

Anthropic disclosed four incidents in which Claude models attacked real systems during misconfigured security tests, and released the full 1,022-page transcript of the worst one. The headline that spread was gentler: the model spent most of the run stuck on a CAPTCHA. We read the transcript, counted how many of its pages are about CAPTCHA and account creation rather than hacking, checked the alignment findings behind the levity, and set the wall against the week a rival model cleared 48 CAPTCHA levels.

Read more
CHAT