About

I am a direct-track Ph.D. student in Computer Science at Shanghai Jiao Tong University and Shanghai Innovation Institute. My research focuses on robot learning, reinforcement learning, humanoid robots, legged locomotion, loco-manipulation, and sim-to-real transfer.

I received my B.Eng. in Automation from Tsinghua University in 2023.

News

  • CoRL 2026 · Accepted UniLab and GeoAlign
  • RSS 2026 · Accepted HiWET and GS-Playground
  • ICRA 2026 · Accepted A2CF, HierKick, and Disturbance-Aware
  • IROS 2025 · Accepted MUTE and UniLegs

Highlights

Selected first-author and co-first-author papers.

2026

UniLab: A Heterogeneous Architecture for Robot RL Beyond GPU-Dominant Paradigms

Y Jia*, Z Cao*, M Yu*, H Zhang*, S Chen*, D Jiang*, M Li, X Li, Y Liu, J Wu, Z Li, ...

CoRL 2026 · Accepted Conference on Robot Learning, 2026. * Equal contribution.

2026

HiWET: Hierarchical World-Frame End-Effector Tracking for Long-Horizon Humanoid Loco-Manipulation

Z Cao, L Yan, Y Zhang, S Chen, J Ma, T Zhan, S Fu, Y Jia, C Lu, Y Gao

RSS 2026 · Accepted Robotics: Science and Systems (RSS), 2026.

2025

Learning Motion Skills with Adaptive Assistive Curriculum Force in Humanoid Robots

Z Cao, Y Zhang, B Nie, H Lin, H Li, Y Gao

ICRA 2026 · Accepted IEEE International Conference on Robotics and Automation (ICRA), 2026.

2025

MUTE: Minimizing Acoustic Noise: Enhancing Quiet Locomotion for Quadruped Robots in Indoor Applications

Z Cao, B Nie, Y Zhang, Y Gao

IROS 2025 · Published IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), 2025, pp. 17972-17979.

2025

UniLegs: Universal Multi-Legged Robot Control through Morphology-Agnostic Policy Distillation

W Xi*, Z Cao*, C Ming, J Zheng, G Zhou

IROS 2025 · Published IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), 2025, pp. 10698-10703. * Equal contribution.

Other Publications

Additional publications, listed by year.

2026

2026

GeoAlign: Beyond Semantics with State-Guided Spatial Alignment in VLA Models

Y Chen, Z Cao, X Peng, Y Zheng, X Si, Y Li, L Yan, K Zhu, X Chen, S Fu, T Zhan, Y Jia, J Yao, Y Xie, K Wang, C Lu, Y Gao

CoRL 2026 · Accepted Conference on Robot Learning, 2026.

2026

GS-Playground: A High-Throughput Photorealistic Simulator for Vision-Informed Robot Learning

Y Jia, H Zhang, Z Zhang, J Wu, M Yu, Z Wang, D Jiang, Z Li, C Cao, Z Yu, ...

RSS 2026 · Accepted Robotics: Science and Systems (RSS), 2026.

2026

HierKick: Hierarchical Reinforcement Learning for Vision-Guided Soccer Robot Control

Y Chen, Z Zhang, Z Cao, Y Chen, S Fu, L Yan, Y Zhang, J Liu, H Li, Y Gao

ICRA 2026 · Accepted IEEE International Conference on Robotics and Automation (ICRA), 2026.

2026

FocusNav: Spatial Selective Attention with Waypoint Guidance for Humanoid Local Navigation

Y Zhang, J Ma, L Yan, Z Cao, Y Zhang, H Li, Y Gao

arXiv preprint arXiv:2601.12790, 2026.

2026

Keep On Going: Learning Robust Humanoid Motion Skills via Selective Adversarial Training

Y Zhang, Z Cao, B Nie, H Li, Z Jiangwei, Q Sun, X Hu, X Yang, Y Gao

AAAI 2026 · Published Proceedings of the AAAI Conference on Artificial Intelligence, 40 (22), 18800-18808, 2026.

2026

Coordinated Humanoid Robot Locomotion with Symmetry Equivariant Reinforcement Learning Policy

B Nie, Y Zhang, R Jin, Z Cao, H Lin, X Yang, Y Gao

AAAI 2026 · Published Proceedings of the AAAI Conference on Artificial Intelligence, 40 (22), 18523-18531, 2026.

2025

2025

Disturbance-Aware Adaptive Compensation in Hybrid Force-Position Locomotion Policy for Legged Robots

Y Zhang, B Nie, Z Cao, Y Fu, Y Gao

ICRA 2026 · Accepted IEEE International Conference on Robotics and Automation (ICRA), 2026.

2025

Contrastive Forward Prediction Reinforcement Learning for Adaptive Fault-Tolerant Legged Robots

Y Fu, Y Zhang, Q Yang, L Yan, Z Cao, Y Gao

CoRL 2025 · Published Proceedings of The 9th Conference on Robot Learning, PMLR 305:3285-3303, 2025.

2025

Anticipate Before Act: Prediction Based Constrained Reinforcement Learning Framework for Skiing Robot Control

H Li, X Yang, J Zhu, Z Cao, Y Zhang, Y Gao

RA-L 2025 · Published IEEE Robotics and Automation Letters 10 (12), 13169-13176, 2025.

2025

Stochastic Trajectory Optimization for Robotic Skill Acquisition From a Suboptimal Demonstration

C Ming, Z Wang, B Zhang, Z Cao, X Duan, J He

RA-L 2025 · Published IEEE Robotics and Automation Letters 10 (6), 6127-6134, 2025.

2024

2024

Constrained Dirichlet Distribution Policy: Guarantee Zero Constraint Violation Reinforcement Learning for Continuous Robotic Control

J Ma, Z Cao, Y Gao

RA-L 2024 · Published IEEE Robotics and Automation Letters 9 (12), 11690-11697, 2024.

2023

2023

Real is Better than Perfect: Sim-to-Real Robotic System in Secondary School Education

J Gao, H Guo, Z Cao, P Huang, G Zhou

IROS 2023 · Published IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), 2023, pp. 4480-4487.

Preprints and Manuscripts

SSRN

Safety Protection for Reinforcement Learning-Based Motion Policy Improvement in Legged Robots

Y Zhang, Y Gao, Z Cao

SSRN working paper 5067499, 2024.

Education

Shanghai Jiao Tong University / Shanghai Innovation Institute Direct-track Ph.D. student in Computer Science, 2023 - Present
Tsinghua University B.Eng. in Automation, 2019 - 2023