Embodied Decision Intelligence Lab (EDI Lab) 清华大学具身决策智能实验室

Introduction

Embodied Decision Intelligence Lab (EDI Lab) belongs to Tsinghua University Shenzhen International Graduate School. It is supersised by Prof. Chao Yu (于超). Our lab is committed to research related to reinforcement learning infrastructure, strategic agent and embodied AI.

Chao Yu is an Associate Professor at Tsinghua University, SIGS. She is also the chairman of the Tsinghua Shenzhen International Graduate School - AgiBot Joint Research Center for Embodied Cognition and Decision Systems (JCES) and the co-founder of Striding AI. As first author or corresponding author, Prof. Chao Yu has published more than 50 papers in top-tier international conferences and journals, including ICML, NeurIPS, ICLR, CVPR, ECCV, CoRL, IROS, ICRA, TMLR, and RAL, with over 7,000 citations on Google Scholar.My representative works include the multi-agent reinforcement learning algorithm MAPPO, which has received more than 4,000 Google Scholar citations, and RLinf, a large-scale reinforcement learning training framework for embodied intelligence, which has accumulated over 4,000 GitHub stars.

Our lab is currently recruiting Master’s students, Ph.D. students, joint-program Ph.D. students with Zhongguancun Academy, postdoctoral researchers, and undergraduate research assistants. We warmly welcome students who are interested in recommended admission or applying through the graduate entrance examination to SIGS programs such as Artificial Intelligence, Data Science and Information Technology, Big Data Engineering, and Electronic Engineering, as well as applicants for the above Ph.D. and postdoctoral positions, to join us.

Highlights

VS-Bench: Evaluating VLMs for Strategic Reasoning and Decision-Making in Multi-Agent Environments
VS-Bench: Evaluating VLMs for Strategic Reasoning and Decision-Making in Multi-Agent Environments
Zelai Xu, Zhexuan Xu, Xiangmin Yi, Huining Yuan, Xinlei Chen, Yongji Wu, Chao Yu, Yu Wang
Proceedings of CVPR  ·  2026CCF-A
RLinf: Flexible and Efficient Large-scale Reinforcement Learning via Macro-to-Micro Flow Transformation
RLinf: Flexible and Efficient Large-scale Reinforcement Learning via Macro-to-Micro Flow Transformation
Chao Yu, Yuanqing Wang, Zhen Guo, Hao Lin, Si Xu, Hongzhi Zang, Quanlu Zhang, Yongji Wu, Chunyang Zhu, Junhao Hu, Zixiao Huang, Mingjie Wei, Yuqing Xie, Ke Yang, Bo Dai, Zhexuan Xu, Jiakun Du, Xiangyuan Wang, Xu Fu, Letong Shi, Zhihao Liu, Kang Chen, Weilin Liu, Gang Liu, Boxun Li, Jianlei Yang, Zhi Yang, Guohao Dai, Yu Wang
Proceedings of OSDI  ·  2025CCF-A
RLinf-VLA: A Unified and Efficient Framework for VLA+RL Training
RLinf-VLA: A Unified and Efficient Framework for VLA+RL Training
Hongzhi Zang, Mingjie Wei, Si Xu, Yongji Wu, Zhen Guo, Yuanqing Wang, Hao Lin, Liangzhi Shi, Yuqing Xie, Zhexuan Xu, Zhihao Liu, Kang Chen, Wenhao Tang, Quanlu Zhang, Weinan Zhang, Chao Yu, Yu Wang
Proceedings of RSS  ·  2025CCF-A
Language Agents with Reinforcement Learning for Strategic Play in the Werewolf Game
Language Agents with Reinforcement Learning for Strategic Play in the Werewolf Game
Zelai Xu, Chao Yu, Fei Fang, Yu Wang, Yi Wu
Proceedings of ICML  ·  2024CCF-A
Is DPO Superior to PPO for LLM Alignment? A Comprehensive Study
Is DPO Superior to PPO for LLM Alignment? A Comprehensive Study
Shusheng Xu, Wei Fu, Jiaxuan Gao, Wenjie Ye, Weilin Liu, Zhiyu Mei, Guangju Wang, Chao Yu, Yi Wu
Proceedings of ICML  ·  2024CCF-A
The Surprising Effectiveness of PPO in Cooperative Multi-Agent Games
The Surprising Effectiveness of PPO in Cooperative Multi-Agent Games
Chao Yu, Akash Velu, Eugene Vinitsky, Jiaxuan Gao, Yu Wang, Alexandre Bayen, Yi Wu*
Proceedings of NeurIPS  ·  2022CCF-A

News

百度百舸率先适配 RLinf v0.3

云厂商首家!百度百舸率先适配 RLinf v0.3,打通具身智能进化闭环

RLinf v0.3 发布

RLinf v0.3来了!从模型生态到真机部署五大能力跃升,无问芯穹与清华大学联合打造

RLinf上新STEAM算法

pi*0.6 RECAP没解决的问题,被STEAM解决了

RLinf参展中国国际供应链促进博览会

RLinf参展中国国际供应链促进博览会

RLinf 上线华为云

华为云、昇腾联合RLinf,共筑基于昇腾算力的具身智能开发生态

Outreach

CCF ADL173 期《具身智能机器人》论坛

活动详情

WAIC国地中心人形机器人与具身智能创新发展论坛报告

活动详情

RSS 2026 RL4VLA Workshop

活动时间:2026年7月12日至18日。

第8届智源大会强化学习论坛报告

活动详情

2026华为云INSPIRE创想者大会技术报告

活动详情

Sponsors