Skip to content

Author

Zongjie Li

We have 4 of 70 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#natural language process... Preprint Feb 2025

GuidedBench: Measuring and Mitigating the Evaluation Discrepancies of In-the-wild LLM Jailbreak Methods

GuidedBench, a novel benchmark comprising a curated harmful question dataset and GuidedEval, an evaluation system integrated with detailed case-by-case evaluation guidelines are introduced, ensuring reliable and reproducible evaluations.

Ruixuan Huang, Xun-Guang Wang, Zongjie Li et al. · 8 citations
Preprint Aug 2026

Reassembling Distributed Risk: Trajectory-Conditioned Action Generation for Multi-Turn Agent Safety

This work proposes ReDiR, a generation-time defense that conditions action generation on trajectory-level security evidence and reduces attack success rates to below 8%, transfers to unseen tool domains, and preserves benign fidelity with low computational overhead.

Yanbo Dai, Zhenlan Ji, Zongjie Li et al. · 0 citations
Preprint Aug 2026

Uncovering and Understanding Hidden Dependencies in the LLM API Reseller Ecosystem via Prefix-Cache Side Channels

These findings reveal substantial hidden dependencies among seemingly independent API resellers, which can create a large potential blast radius, where a confidentiality or integrity failure along a common upstream path may affect users across multiple downstream resellers.

Zimo Ji, Xin Wei, Congying Xu et al. · 0 citations
Open access Sep 2025

Empirical Study of Code Large Language Models for Binary Security Patch Detection

This initial study demonstrates that directly prompting off-the-shelf code LLMs remains ineffective; even advanced prompting strategies cannot compensate for the lack of task-specific knowledge, and fine-tuning proves highly effective, with pseudo-code representation consistently yielding the best performance.

Qingyuan Li, Bin-Chang Li, Cuiyun Gao et al. · 3 citations · ⚡1

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.