About

I work on generative models for human motion, camera behavior, and dynamic visual content.

I am a first-year Ph.D. student in Control Science and Engineering at Tsinghua University, supervised by Prof. Xiu Li. Before Tsinghua, I received my B.Eng. from Chongqing University, where I ranked first in my class (1/323).

My research focuses on controllable models that connect language, music, human motion, 3D states, camera trajectories, and video.

Previously, I was a Research Intern at Tencent Games, Shenzhen, working on joint human-camera reconstruction.

I am open to research discussions, collaborations, and research internship opportunities.

News

  • CameraOperator was accepted at ACM MM 2026.
  • InfiniteDance was accepted at ECCV 2026.
  • WATCH was accepted.
  • AToM was published at CVPR 2025.

Research

I study data and model design for dynamic visual worlds, with an emphasis on physically plausible human motion, camera-subject interaction, and controllable generation beyond curated settings.

My recent work spans large-scale 3D dance generation, joint human-camera reconstruction, camera trajectory generation, and controllable video generation.

Selected Publications

* denotes co-first authors. My name is underlined.

InfiniteDance paper preview

InfiniteDance: Scalable 3D Dance Generation Towards in-the-wild Generalization

Ronghui Li*, Zhongyuan Hu*, Li Siyao, Youliang Zhang, Haozhe Xie, Mingyuan Zhang, Jie Guo, Xiu Li, Ziwei Liu

ECCV 2026, accepted

Co-first author and core contributor; 100+ hours of multimodal 3D dance data.

CameraOperator method overview from the paper

Camera Operator: Object-Grounded Camera Trajectory Generation from Text and 3D Bounding Box Sequences

Zhongyuan Hu, Yue Ma, Jiangming Wang, Ronghui Li, Xiu Li

ACM MM 2026, accepted

First author and core contributor; BlockCam benchmark with 41K sequences.

WATCH paper preview

WATCH: World-aware Allied Trajectory and pose reconstruction for Camera and Human

Qijun Ying*, Zhongyuan Hu*, Rui Zhang, Ronghui Li, Yu Lu, Zijiao Zeng

Accepted

Co-first author and core contributor; joint human-camera reconstruction in the wild.

Controllable video generation survey preview

Controllable Video Generation: A Survey

Yue Ma*, Kunyu Feng*, Zhongyuan Hu*, Xinyu Wang*, Yucheng Wang, Mingzhe Zheng, Bingyuan Wang, Qinghe Wang, Xuanhua He, Hongfa Wang, Chenyang Zhu, Hongyu Liu, Yingqing He, Zeyu Wang, Zhifeng Li, Xiu Li, Sirui Han, Yike Guo, Wei Liu, Dan Xu, Linfeng Zhang, Qifeng Chen

ACM Computing Surveys, under review

Co-first author and core contributor; a taxonomy of controllable video generation.

AToM paper preview

AToM: Aligning Text-to-Motion Model at Event-Level with GPT-4Vision Reward

Haonan Han*, Xiangzuo Wu*, Huan Liao*, Zunnan Xu, Zhongyuan Hu, Ronghui Li, Yachao Zhang, Xiu Li

CVPR 2025

Event-level alignment for text-to-motion generation with GPT-4Vision reward.

Education

2025.09 - Present

Tsinghua University

Ph.D. in Control Science and Engineering, supervised by Prof. Xiu Li.

2021.09 - 2025.06

Chongqing University

B.Eng. in Computer Science and Technology. GPA 3.91/4.00; ranked 1/323.

Experience

2025.06 - 2025.12

Tencent Games

Research Intern, Shenzhen. Human motion capture, pose estimation, and camera estimation.