Conference
Open access
2026
SecureBreak: A Dataset towards Safe and Secure Models
SecureBreak is introduced, a safety-oriented dataset designed to support the development of AI-driven solutions for detecting harmful LLM outputs caused by residual weaknesses in security alignment and is valuable not only for constructing post-generation filtering modules that act as a last-line defense, but also for building additional supervisory intelligence for alignment optimization.
Marco Arazzi, Vignesh Kumar Kembu, Antonino Nocera
· Proceedings of the 15th Inte... · 0 citations