CoDrift proposes compositional action fields for offline reinforcement learning
A new arXiv preprint introduces CoDrift, a one-step generative policy framework that combines behavioral compatibility and value-seeking objectives for offline reinforcement learning. The authors report the best average rank across 73 OGBench and D4RL tasks in both offline and offline-to-online settings.