Skip to content

Author

Sylvia L. Herbert

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#machine learning Preprint Sep 2026

Safe Score Matching: Diffusion Policies with Hamilton-Jacobi Reachability for Online Safe Reinforcement Learning

Online safe reinforcement learning (RL) seeks policies that maximize reward while satisfying safety constraints. A popular line of research in safe RL relaxes safety to a soft expected-cost constraint and solves the resulting Constrained Markov Decision Process via primal-dual Lagrangian updates that only enforce safet...

Bo-Yang Li, Matthew Kim, Sylvia L. Herbert · 0 citations
#machine learning Preprint Sep 2026

Constrained Flow Policy Updates: A Generalized Schr\"odinger Bridge View

This work builds on the density-free kinetic-energy regularizer of FLAC, a recent reward-only method, and proposes Reparameterized Augmented-Lagrangian Flow Actor with Least Energy (RAFALE), an off-policy actor-critic method for safe RL.

Bo-Yan Li, Matthew Kim, Sylvia L. Herbert · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.