About
News
Highlights
Publications
Education
EN
/
中文
About
I am a direct-track Ph.D. student in Computer Science at Shanghai
Jiao Tong University and Shanghai Innovation Institute. My research
focuses on robot learning, reinforcement learning, humanoid robots,
legged locomotion, loco-manipulation, and sim-to-real transfer.
I received my B.Eng. in Automation from Tsinghua University in 2023.
News
2026-09-05
2026-04-27 RSS 2026 · Accepted HiWET and GS-Playground
2026-01-31 ICRA 2026 · Accepted A2CF, HierKick, and Disturbance-Aware
2025-06 (mid) IROS 2025 · Accepted MUTE and UniLegs
Highlights
Selected first-author and co-first-author papers.
2026
UniLab: A Heterogeneous Architecture for Robot RL Beyond GPU-Dominant Paradigms
Y Jia*, Z Cao* , M Yu*, H Zhang*, S Chen*, D Jiang*, M Li, X Li, Y Liu, J Wu, Z Li, ...
CoRL 2026 · Accepted Conference on Robot Learning, 2026. * Equal contribution.
Scholar arXiv Project
2026
HiWET: Hierarchical World-Frame End-Effector Tracking for Long-Horizon Humanoid Loco-Manipulation
Z Cao , L Yan, Y Zhang, S Chen, J Ma, T Zhan, S Fu, Y Jia, C Lu, Y Gao
RSS 2026 · Accepted Robotics: Science and Systems (RSS), 2026.
Scholar arXiv
2025
Learning Motion Skills with Adaptive Assistive Curriculum Force in Humanoid Robots
Z Cao , Y Zhang, B Nie, H Lin, H Li, Y Gao
ICRA 2026 · Accepted IEEE International Conference on Robotics and Automation (ICRA), 2026.
Scholar arXiv
2025
MUTE: Minimizing Acoustic Noise: Enhancing Quiet Locomotion for Quadruped Robots in Indoor Applications
Z Cao , B Nie, Y Zhang, Y Gao
IROS 2025 · Published IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), 2025, pp. 17972-17979.
Scholar DOI arXiv
2025
UniLegs: Universal Multi-Legged Robot Control through Morphology-Agnostic Policy Distillation
W Xi*, Z Cao* , C Ming, J Zheng, G Zhou
IROS 2025 · Published IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), 2025, pp. 10698-10703. * Equal contribution.
Scholar DOI arXiv
Other Publications
Additional publications, listed by year.
2026
2026
GeoAlign: Beyond Semantics with State-Guided Spatial Alignment in VLA Models
Y Chen, Z Cao , X Peng, Y Zheng, X Si, Y Li, L Yan, K Zhu, X Chen, S Fu, T Zhan, Y Jia, J Yao, Y Xie, K Wang, C Lu, Y Gao
CoRL 2026 · Accepted Conference on Robot Learning, 2026.
arXiv
2026
GS-Playground: A High-Throughput Photorealistic Simulator for Vision-Informed Robot Learning
Y Jia, H Zhang, Z Zhang, J Wu, M Yu, Z Wang, D Jiang, Z Li, C Cao, Z Yu, ...
RSS 2026 · Accepted Robotics: Science and Systems (RSS), 2026.
Scholar arXiv Project
2026
HierKick: Hierarchical Reinforcement Learning for Vision-Guided Soccer Robot Control
Y Chen, Z Zhang, Z Cao , Y Chen, S Fu, L Yan, Y Zhang, J Liu, H Li, Y Gao
ICRA 2026 · Accepted IEEE International Conference on Robotics and Automation (ICRA), 2026.
Scholar arXiv
2026
FocusNav: Spatial Selective Attention with Waypoint Guidance for Humanoid Local Navigation
Y Zhang, J Ma, L Yan, Z Cao , Y Zhang, H Li, Y Gao
arXiv preprint arXiv:2601.12790, 2026.
Scholar arXiv
2026
Keep On Going: Learning Robust Humanoid Motion Skills via Selective Adversarial Training
Y Zhang, Z Cao , B Nie, H Li, Z Jiangwei, Q Sun, X Hu, X Yang, Y Gao
AAAI 2026 · Published Proceedings of the AAAI Conference on Artificial Intelligence, 40 (22), 18800-18808, 2026.
Scholar DOI arXiv
2026
Coordinated Humanoid Robot Locomotion with Symmetry Equivariant Reinforcement Learning Policy
B Nie, Y Zhang, R Jin, Z Cao , H Lin, X Yang, Y Gao
AAAI 2026 · Published Proceedings of the AAAI Conference on Artificial Intelligence, 40 (22), 18523-18531, 2026.
Scholar DOI arXiv
2025
2025
Disturbance-Aware Adaptive Compensation in Hybrid Force-Position Locomotion Policy for Legged Robots
Y Zhang, B Nie, Z Cao , Y Fu, Y Gao
ICRA 2026 · Accepted IEEE International Conference on Robotics and Automation (ICRA), 2026.
Scholar arXiv
2025
Contrastive Forward Prediction Reinforcement Learning for Adaptive Fault-Tolerant Legged Robots
Y Fu, Y Zhang, Q Yang, L Yan, Z Cao , Y Gao
CoRL 2025 · Published Proceedings of The 9th Conference on Robot Learning, PMLR 305:3285-3303, 2025.
Scholar PMLR
2025
Anticipate Before Act: Prediction Based Constrained Reinforcement Learning Framework for Skiing Robot Control
H Li, X Yang, J Zhu, Z Cao , Y Zhang, Y Gao
RA-L 2025 · Published IEEE Robotics and Automation Letters 10 (12), 13169-13176, 2025.
Scholar DOI
2025
Stochastic Trajectory Optimization for Robotic Skill Acquisition From a Suboptimal Demonstration
C Ming, Z Wang, B Zhang, Z Cao , X Duan, J He
RA-L 2025 · Published IEEE Robotics and Automation Letters 10 (6), 6127-6134, 2025.
Scholar DOI
2024
2024
Constrained Dirichlet Distribution Policy: Guarantee Zero Constraint Violation Reinforcement Learning for Continuous Robotic Control
J Ma, Z Cao , Y Gao
RA-L 2024 · Published IEEE Robotics and Automation Letters 9 (12), 11690-11697, 2024.
Scholar DOI
2023
2023
Real is Better than Perfect: Sim-to-Real Robotic System in Secondary School Education
J Gao, H Guo, Z Cao , P Huang, G Zhou
IROS 2023 · Published IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), 2023, pp. 4480-4487.
Scholar DOI
Preprints and Manuscripts
SSRN
Safety Protection for Reinforcement Learning-Based Motion Policy Improvement in Legged Robots
Y Zhang, Y Gao, Z Cao
SSRN working paper 5067499, 2024.
Scholar SSRN
Education
Shanghai Jiao Tong University / Shanghai Innovation Institute
Direct-track Ph.D. student in Computer Science, 2023 - Present
Tsinghua University
B.Eng. in Automation, 2019 - 2023