Generative Support Realignment for Cross-Domain Offline Reinforcement Learning
A performance gap bound is derived that characterizes the interplay between generation error and source--target dynamics gap, providing theoretical guidance for effective coverage expansion and source utilization.
Minung Kim, Jeongmo Kim, G. Choi et al.
· 0 citations