Zhennan Jiang

Ph.D. Student · Reinforcement Learning & Robotics · CASIA

prof_pic.jpg

I am Zhennan Jiang (江震南), a Ph.D. student in Computer Applications Technology at the Institute of Automation, Chinese Academy of Sciences (CASIA). I am fortunate to be advised by Prof. Dongbin Zhao (IEEE Fellow) and Dr. Haoran Li.

I received my Bachelor’s degree in Automation from Central South University. I am also a joint-training student at Zhongguancun Academy (ZGCA), working with Prof. Chao Yu.

My research focuses on building reliable, efficient, and generalizable learning systems for real-world robots, with particular interests in reinforcement learning, generative policies, and world models.

Reinforcement Learning Robot Learning Generative Policies World Models

Research

Selected Publications

Google Scholar ↗

Peer-reviewed & Accepted

Published research
  1. TeViR: Text-to-Video Reward with Diffusion Models for Efficient Reinforcement Learning
    Yuhui Chen, Haoran Li, Zhennan Jiang, Haowei Wen, and Dongbin Zhao
    In IEEE Transactions on Systems Man Cybernetics-Systems, 2024
  2. Generalizing Consistency Policy to Visual RL with Prioritized Proximal Experience Regularization
    Haoran Li, Zhennan Jiang, Yuhui Chen, and Dongbin Zhao
    In The 38th Annual Conference on Neural Information Processing Systems, NIPS, Sep 2024
  3. RL-100: Performant Robotic Manipulation with Real-World Reinforcement Learning
    Kun Lei*, Huanyu Li*, Dongjie Yu*, Zhenyu Wei*, Lingxiao Guo, Zhennan Jiang, Ziyu Wang, Shiyu Liang, and Huazhe Xu
    Science Robotics, Sep 2026
    Accepted for publication
  4. World4RL: Diffusion World Models for Policy Refinement with Reinforcement Learning for Robotic Manipulation
    Zhennan Jiang, Kai Liu, Yuxin Qin, Shuai Tian, Yupeng Zheng, Mingcai Zhou, Chao Yu, Haoran Li, and Dongbin Zhao
    IEEE Robotics and Automation Letters (RA-L), Sep 2026
    Accepted for publication
  5. MARS-Sep: Multimodal-Aligned Reinforced Sound Separation
    Zihan Zhang, Xize Cheng, Zhennan Jiang, Dongjie Fu, Jingyuan Chen, Zhou Zhao, and Tao Jin
    In International Conference on Learning Representations (ICLR), Sep 2026

Preprints & Under Review

Latest work
  1. Under Review
    Beyond Action Residuals: Real-World Robot Policy Steering via Bottleneck Latent Reinforcement Learning
    Beyond Action Residuals: Real-World Robot Policy Steering via Bottleneck Latent Reinforcement Learning
    Dongjie Yu*, Kun Lei*Zhennan Jiang, Jia Pan, and Huazhe Xu
    May 2026
  2. Under Review
    Posterior Optimization with Clipped Objective for Bridging Efficiency and Stability in Generative Policy Learning
    Posterior Optimization with Clipped Objective for Bridging Efficiency and Stability in Generative Policy Learning
    Yuhui Chen, Haoran Li, Zhennan Jiang, Yuxing Qin, Yuxuan Wan, Weiheng Liu, and Dongbin Zhao
    Apr 2026
  3. Under Review
    wovr.png
    WoVR: World Models as Reliable Simulators for Post-Training VLA Policies with RL
    Zhennan Jiang, Shangqing Zhou, Yutong Jiang, Zefang Huang, Mingjie Wei, Yuhui Chen, Tianxing Zhou, Zhen Guo, Hao Lin, Quanlu Zhang, Yu Wang, Haoran Li, Chao Yu, and Dongbin Zhao
    Apr 2026

Invited Talks