Robot learning · Shanghai机器人学习 · 上海

Puze Liu 刘普泽

I am an Associate Professor at the Shanghai Research Institute for Intelligent Autonomous Systems, Tongji University.

Previously, I was Deputy Head of System AI for Robot Learning at the German Research Center for Artificial Intelligence (DFKI). I received my Ph.D. from Intelligent Autonomous Systems, TU Darmstadt, advised by Prof. Jan Peters.

My research equips robots with complex skills through modern machine learning, with a particular focus on on-robot learning from real-world interaction.

我现任同济大学上海自主智能无人系统科学中心副教授。

此前,我曾任德国人工智能研究中心(DFKI)机器人学习系统人工智能研究组副组长。我在达姆施塔特工业大学智能自主系统实验室获得博士学位,导师为 Jan Peters 教授

我的研究致力于利用先进的机器学习方法赋予机器人复杂技能,尤其关注基于真实世界交互的机器人在线学习

Updates近期动态

  • 2026-04-28

    Our paper "Mind Your Steps: A General Learning Framework for Accurate Humanoid Foothold Tracking" has been accepted in Robotics: Science and Systems (RSS) 2026!我们的论文《Mind Your Steps: A General Learning Framework for Accurate Humanoid Foothold Tracking》被 Robotics: Science and Systems(RSS)2026 接收!

  • 2026-03-01

    I joined the Shanghai Research Institute for Intelligent Autonomous Systems (SRIAS) as an Associate Professor!我加入同济大学上海自主智能无人系统科学中心(SRIAS),担任副教授!

  • 2025-10-02  

    Our workshop LeaPRiDE has been successfully organized at IROS 2025!我们在 IROS 2025 成功举办了 LeaPRiDE 研讨会!

  • 2025-05-01    

    Our Paper: "Morphologically Symmetric Reinforcement Learning for Ambidextrous Bimanual Manipulation" has been accepted in CoRL 2025!我们的论文《Morphologically Symmetric Reinforcement Learning for Ambidextrous Bimanual Manipulation》被 CoRL 2025 接收!

  • 2025-05-09  

    Our Workshop Proposal: "LeaPRiDE: Learning, Planning, and Reasoning in Dynamic Environments" has been accepted in IROS 2025!我们的研讨会提案“LeaPRiDE:动态环境中的学习、规划与推理”被 IROS 2025 接收!

  • 2025-05-01  

    Our Paper: "Maximum Total Correlation Reinforcement Learning" has been accepted in ICML 2025!我们的论文《Maximum Total Correlation Reinforcement Learning》被 ICML 2025 接收!

Research highlights研究精选

All publications全部论文

Journal Articles

  1. A robot operating system framework for using large language models in embodied AI illustration
    2026 Nature Machine Intelligence

    A robot operating system framework for using large language models in embodied AI

    An open-source bridge between large language models and ROS turns natural-language goals into robust robot behaviors, while learning and refining new skills through feedback.

    该开源框架连接大语言模型与 ROS,将自然语言目标转化为可靠的机器人行为,并能通过反馈持续学习和优化新技能。

  2. Safe Reinforcement Learning on the Constraint Manifold: Theory and Applications illustration
    2025 IEEE Transactions on Robotics (T-RO)

    Safe Reinforcement Learning on the Constraint Manifold: Theory and Applications

    A principled framework projects arbitrary reinforcement-learning actions onto a constraint manifold, enabling safe, high-dimensional learning directly on a real robot air-hockey platform.

    该框架将强化学习动作投影到约束流形的安全空间中,使机器人能够在真实空气曲棍球平台上完成安全的高维在线学习。

  3. Adaptive control based friction estimation for tracking control of robot manipulators illustration
    2025 IEEE Robotics and Automation Letters

    Adaptive control based friction estimation for tracking control of robot manipulators

    An adaptive controller estimates joint friction online and compensates for it during motion, improving manipulator tracking without relying on a fixed friction model.

    自适应控制器在运动过程中在线估计并补偿关节摩擦,无需固定摩擦模型即可提升机器人操作臂的轨迹跟踪精度。

  4. Fast Kinodynamic Planning on the Constraint Manifold With Deep Neural Networks illustration
    2024 IEEE Transactions on Robotics (T-RO)

    Fast Kinodynamic Planning on the Constraint Manifold With Deep Neural Networks

    A neural planner generates dynamically feasible motions on constrained manifolds at constant inference time, supporting rapid replanning in demanding robot air-hockey scenarios.

    神经规划器以固定推理时间生成满足动力学与流形约束的运动轨迹,可在高动态机器人空气曲棍球任务中快速重规划。

Conference Papers

  1. Morphologically Symmetric Reinforcement Learning for Ambidextrous Bimanual Manipulation illustration
    2025 Conference on Robot Learning (CoRL)

    Morphologically Symmetric Reinforcement Learning for Ambidextrous Bimanual Manipulation

    Symmetry-aware reinforcement learning transfers skills across mirrored morphologies, allowing a bimanual robot to manipulate dexterously with either arm.

    利用形态对称性的强化学习可在镜像结构间迁移技能,使双臂机器人能够使用任一手臂完成灵巧操作。

  2. Handling Long-Term Safety and Uncertainty in Safe Reinforcement Learning illustration
    2024 8th Annual Conference on Robot Learning

    Handling Long-Term Safety and Uncertainty in Safe Reinforcement Learning

    The method explicitly models long-term uncertainty and safety, helping reinforcement-learning agents remain reliable beyond short planning horizons.

    该方法显式建模长期不确定性与安全约束,使强化学习智能体在超越短期规划范围后仍能保持可靠。

  3. A Retrospective on the Robot Air Hockey Challenge: Benchmarking Robust, Reliable, and Safe Learning Techniques for Real-World Robotics illustration
    2024 Proceedings of the 38th International Conference on Neural Information Processing Systems

    A Retrospective on the Robot Air Hockey Challenge: Benchmarking Robust, Reliable, and Safe Learning Techniques for Real-World Robotics

    A large-scale real-robot benchmark distills the practical lessons of the Robot Air Hockey Challenge, revealing when hybrid learning and prior knowledge outperform purely data-driven systems.

    该真实机器人基准总结了空气曲棍球挑战赛的实践经验,揭示了融合先验知识与学习的方法何时优于纯数据驱动系统。

  4. Safe Reinforcement Learning of Dynamic High-Dimensional Robotic Tasks: Navigation, Manipulation, Interaction illustration
    2023 Proceedings of the IEEE International Conference on Robotics and Automation (ICRA)

    Safe Reinforcement Learning of Dynamic High-Dimensional Robotic Tasks: Navigation, Manipulation, Interaction

    A unified safe-exploration formulation handles complex, learned collision constraints and transfers to real navigation, manipulation, and human–robot interaction tasks.

    统一的安全探索方法可处理由数据学习的复杂碰撞约束,并在真实机器人的导航、操作与人机交互任务中完成验证。

  5. Regularized Deep Signed Distance Fields for Reactive Motion Generation illustration
    2022 Proceedings of the IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS)

    Regularized Deep Signed Distance Fields for Reactive Motion Generation

    ReDSDF learns smooth, multi-scale distance fields for articulated bodies, enabling fast collision reasoning and reactive motion around people in shared workspaces.

    ReDSDF 学习面向关节物体的平滑多尺度距离场,使机器人能够在共享空间中快速进行碰撞推理与人机协作运动。

  6. Robot Reinforcement Learning on the Constraint Manifold illustration
    2022 Proceedings of the 5th Conference on Robot Learning (CoRL)

    Robot Reinforcement Learning on the Constraint Manifold

    Best Paper Award Finalist

    Constraint-manifold exploration embeds known robot models and safety limits into reinforcement learning, improving sample efficiency while guaranteeing safe exploration.

    约束流形探索将机器人模型与安全限制嵌入强化学习,在保证探索安全的同时提高了样本效率。

  7. Efficient and Reactive Planning for High Speed Robot Air Hockey illustration
    2021 2021 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS)

    Efficient and Reactive Planning for High Speed Robot Air Hockey

    Best Entertainment and Amusement Paper Award Finalist

    A fast, reactive planning stack pushes general-purpose robot arms to execute precise hits and competitive play in a highly dynamic air-hockey environment.

    快速反应式规划系统使通用机器人操作臂能够在高动态空气曲棍球环境中完成精准击球与对抗。