Skip to content

Author

Faiza Medjek

1 paper indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Review Sep 2026

How Reliable are Automated Jailbreak Evaluators? A Study of Human-Machine Agreement in Cybersecurity Multi-Turn LLM Attacks

Evaluating the effectiveness of multi-turn jailbreak attacks on large language models (LLMs) increasingly relies on automated evaluators, yet their reliability and agreement with human judgments in cybersecurity conversations have not been systematically examined. This paper presents a systematic human-machine agreemen...

Michael Tchuindjang, Nathan Duran, Phil Legg et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.