Open access
Jun 2026
A red teaming framework for large language models: a case study on faithfulness evaluation
A red teaming framework that systematically uncovers vulnerabilities in LLM outputs and provides both actionable insights into current LLM vulnerabilities and a scalable methodology for ongoing safety evaluation as models continue to evolve is presented.
Abrar Alotaibi, Raed Mughus, Moataz Ahmed
· Software quality journal · 1 citation