Embodied AI · Motion Generation · World Models

Xiangyue ZHANG (章湘粤 · Ian)

Ph.D. student in Mechano-Informatics at The University of Tokyo

Hi there 👋. I am Xiangyue ZHANG (章湘粤), a Ph.D. student in Mechano-Informatics at The University of Tokyo, supervised by Prof. Tatsuya Harada. I am supported by a full scholarship from The University of Tokyo Fellowship.

My research focuses on Embodied AI, 3D/2D motion generation, and world models, with a long-term interest in how intelligent agents perceive, move, and interact in physical environments.

Outside research, music is where I learn timing without equations. I sing for Xia Qian Yu (夏千嶼), an indie-rock band, and have performed at music festivals, bars, campus shows, and solo stages. Music keeps my research from becoming too mechanical: it asks me to listen before moving, to feel rhythm before explaining it, and to remember that expressive motion is something lived before it is modeled.

I am open to remote or on-site internship and visiting opportunities around embodied intelligence, motion generation, and human-centered AI systems.

Feel free to contact me by email if you’d like to discuss or collaborate. 欢迎优秀的本科/研究生联系科研合作! Email: x-zhang[AT]mi.t.u-tokyo.ac.jp

News

Jun 18, 2026
🎉 StreamTalk, OmniDance, and FlowerDance were accepted by ECCV 2026.
Mar 31, 2026
🎉 MACE-Dance was accepted by SIGGRAPH 2026.
Nov 8, 2025
🎉 GlobalDiff was accepted by AAAI 2026.
Jun 26, 2025
🎉 SemTalk was accepted by ICCV 2025.
Dec 22, 2024
🥂 Our band performed successfully at the Hua Young Music Festival! Cheers! Check More to see pictures.


Preprints

  1. World2Motion: Turning Video World Models into 3D Human Motion Generators video preview
    Preprint · 2026
    World2Motion: Turning Video World Models into 3D Human Motion Generators
    Fangyuan Tu, Xiangyue Zhang, Yiyi Cai, Yichen Peng, Kunhang Li, Bo Zheng, Zhixiang Wang, Kaipeng Zhang, Erwin Wu, Haoran Xie, Haiyang Liu
    arXiv preprint, 2026
  2. FloodDiffusion 2: Efficient and Path Controllable Streaming Motion Generation video preview
    Preprint · 2026
    FloodDiffusion 2: Efficient and Path Controllable Streaming Motion Generation
    Yiyi Cai, Yuhan Wu, Kunhang Li, Tu Fangyuan, Xiangyue Zhang, Qiaoge Li, Zhixiang Wang, Kaipeng Zhang, Haiyang Liu
    arXiv preprint, 2026

Selected Publications

  1. StreamTalk project preview
    🔥 ECCV 2026
    StreamTalk: Streaming Co-Speech Gesture Generation with Key-Pose Anchoring
    Xiangyue Zhang, Jianfang Li, Jiaxu Zhang, Kaixing Yang, and Steven Hoi
    European Conference on Computer Vision (ECCV), 2026
  2. GlobalDiff project preview
    🔥 AAAI 2026
    Mitigating Error Accumulation in Co-Speech Motion Generation via Global Rotation Diffusion and Multi-Level Constraints
    Xiangyue Zhang, Jianfang Li, Jianqiang Ren, and Jiaxu Zhang
    Annual AAAI Conference on Artificial Intelligence (AAAI), 2026
  3. SemTalk project preview
    🔥 ICCV 2025
    SemTalk: Holistic Co-speech Motion Generation with Frame-level Semantic Emphasis
    Xiangyue Zhang, Jianfang Li, Jiaxu Zhang, Ziqiang Dang, Jianqiang Ren, Liefeng Bo, and Zhigang Tu
    International Conference on Computer Vision (ICCV), 2025


Experience & Education

2025.12 - 2026.03Shenzhen

ByteDance

Research Intern, Intelligent Creation Team

Advisor: Youjiang Xu. Working on large streaming motion generation models.



Awards & Honors

2026 - 2029
JUN, 2026
Outstanding Graduate Award (Master's)
JAN, 2026
Wang Zhizhuo Scholarship for Innovative Talents (top 0.3%)
OCT, 2025
National Scholarship (top 3%)
JUN, 2023
Outstanding Graduate Award (Bachelor's)


Service

Reviewer
Reviewer Service: NeurIPS, AAAI, ACM MM, T-CSVT