Skip to content

Author

Xiyu Zeng

We have 2 of 5 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

2026

SafeSteer: Adaptive Subspace Steering for Efficient Jailbreak Defense in Vision Language Models

As the capabilities of Vision Language Models (VLMs) continue to improve, they are increasingly targeted by jailbreak attacks. Existing defense methods face two major limitations: (1) they struggle to ensure safety without compromising the model’s utility; and (2) many defense mechanisms significantly reduce the model’...

Xiyu Zeng, Siyuan Liang, Liming Lu et al. · 4 citations
Sep 2025

SafeSteer: Adaptive Subspace Steering for Efficient Jailbreak Defense in Vision-Language Models

SafeSteer is a lightweight, inference-time steering framework that effectively defends against diverse jailbreak attacks without modifying model weights, using the innovative use of Singular Value Decomposition to construct a low-dimensional safety subspace during inference.

Xiyu Zeng, Siyuan Liang, Liming Lu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.