Steering Generative Robot Policies with Lexicographic Preferences
It is shown that a frozen generative robot policy---based on either diffusion or flow matching---can be steered at inference time to respect such lexicographically ordered deployment objectives, and a controlled manipulation study shows that the dynamic barrier reaches comparable best performance over a substantially w...