Generative Support Realignment for Cross-Domain Offline Reinforcement Learning
A performance gap bound is derived that characterizes the interplay between generation error and source--target dynamics gap, providing theoretical guidance for effective coverage expansion and source utilization.