כתבה
arXiv cs.LG ·
Generative Support Realignment for Cross-Domain Offline Reinforcement Learning
תקציר מקורי באנגליתarXiv:2605.13054v2 Announce Type: replace Abstract: Cross-domain offline reinforcement learning learns a target policy from pre-collected source and target datasets with different dynamics. When target data are scarce, effectively compensating for their limited coverage using source data remains challenging due to the discrepancy between domains. We propose Target-aligned Coverage Expansion (TCE), which leverages source states to generatively realign and expand the limited target support while controlling the generation error induced by this expansion. We further derive a performance gap bound that characterizes the interplay between generation error and source--target dynamics gap, providing theoretical guidance for effective coverage expansion and source utilization. Across diverse cross
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית