I'm currently a PhD student at The University of Hong Kong (HKU), and I'm a continual RL learner.
I dream of a small, focused team to create root-level work
with aspiration of craft over hype.
Algorithms define my life. :)
I'm currently a PhD student at The University of Hong Kong (HKU), and I'm a continual RL learner.
I dream of a small, focused team to create root-level work
with aspiration of craft over hype.
Algorithms define my life. :)
[2606.21136] Horizon Adaptive Offline Policy Learning via Value Stitching
Forked from ZhengYinan-AIR/Diffusion-Planner
[ICLR 2025 Oral] The official implementation of "Diffusion-Based Planning for Autonomous Driving with Flexible Guidance"
Python