Skip to content

Author

Collin Francel

1 paper indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Jul 2026

ToxScreen: Detecting Whether an LLM Has Been Poisoned

No method reliably surfaces every backdoor, but a broadly jailbreakable model is itself anomalous, a useful signal even when the exact trigger is not recovered, allowing defenders to filter jailbreaks.

A. Hughes, N. Xing, Collin Francel et al. · 0 citations