Skip to content

Author

Sangdon Park

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

The Interplay of Harness Design and Post-Training in LLM Agents

It is shown that harness-aware post-training not only improves in-distribution performance but also enables agents to robustly adapt to OOD settings, highlighting the importance of harness-aware post-training under such shifts.

Kyungmin Kim, Youngbin Choi, Seoyeon Lee et al. · 7 citations
Jul 2026

MJ: Multi-turn LLM Jailbreaking via Decomposed Credit Assignment

Modern large language models (LLMs) operate in interactive multi-turn settings, making multi-turn jailbreaking a realistic threat model and an important setting for automated red teaming. A core challenge in learning multi-turn jailbreak attackers is credit assignment: different turns contribute differently to the fina...

Junyoung Park, N. Park, Sechan Lee et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.