Skip to content

Author

Kellin Pelrine

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Aug 2026

AI Security Leaderboard: Methodology, Results and Minimal Standard

The AI Security Leaderboard is an independent benchmark that ranks the safeguards of frontier AI models from least to most secure. It tests models against the FAR$.$AI Minimal Standard for Safeguards, which represents a minimum bar for security: meeting it does not guarantee a secure model, but failing to meet it guara...

Jasper Timm, Lukas Struppek, Ziwei Xu et al. · 0 citations
#artificial intelligence Review Sep 2026

AI Persuasion as a Threat to Human Control

The threat that AI persuasion poses to human control has been acknowledged in the literature, but not yet systematically studied. Now that persuasion attacks are no longer theoretical - with Anthropic's Claude Mythos 5 recently making headlines for trying to convince people involved in an open-source project to merge m...

Joshua H. Levy, Mick Yang, Kellin Pelrine · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.