Hi, I’m Jiwen. I’m currently at Xiaomi Robotics, working closely with my friend Yiran Qin.

My research currently focuses on robotics, with a particular emphasis on foundation models for robotics.

🤝 Open to collaborations on robotics and embodied foundation models. If you share these interests, let’s talk!

Contact me via 📬 Email / WeChat 🎓 Google Scholar Citations: 2697

Selected Works

(*: indicates equal contribution; #: indicates corresponding author)

ECCV 2026

MemLearner: Learning to Query Context Memory for Video World Models

Jiwen Yu, Jianxiong Gao, Jianhong Bai, Yiran Qin, Kaiyi Huang, Quande Liu, Xintao Wang#, Pengfei Wan, Kun Gai, Xihui Liu#

ECCV 2026

Paper | Project Page

SIGGRAPH Asia 2025

Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval

Jiwen Yu, Jianhong Bai, Yiran Qin, Quande Liu#, Xintao Wang, Pengfei Wan, Di Zhang, Xihui Liu#

SIGGRAPH Asia 2025

Paper | Project Page | Dataset

ICCV 2025

GameFactory: Creating New Games with Generative Interactive Videos

Jiwen Yu*, Yiran Qin*, Xintao Wang#, Pengfei Wan, Di Zhang, Xihui Liu#

ICCV 2025 Highlight

Paper | Project Page | GitHub | Dataset

Preprint
sym

Survey of Interactive Generative Video

Jiwen Yu*, Yiran Qin*, Haoxuan Che*, Quande Liu#, Xintao Wang, Pengfei Wan, Di Zhang, Kun Gai, Hao Chen, Xihui Liu#

Position: Interactive Generative Video as Next-Generation Game Engine

Jiwen Yu*, Yiran Qin*, Haoxuan Che, Quande Liu, Xintao Wang#, Pengfei Wan, Di Zhang, Xihui Liu#

Survey Paper | Position Paper

CVPRW 2026

MultiWorld: Scalable Multi-Agent Multi-View Video World Models

Haoyu Wu, Jiwen Yu, Yingtian Zou, Xihui Liu

CVPR 2026 Workshop MARS-EAI Best Paper Award

Paper | Project Page

ICML 2025
sym

WorldSimBench: Towards Video Generation Models as World Simulators

Yiran Qin*, Zhelun Shi*, Jiwen Yu, Xijun Wang, Enshen Zhou, Lijun Li, Zhenfei Yin, Xihui Liu, Lu Sheng, Jing Shao, Lei Bai, Wanli Ouyang, Ruimao Zhang

ICML 2025

Paper | Project Page

NeurIPS 2023
sym

CRoSS: Diffusion Model Makes Controllable, Robust and Secure Image Steganography

Jiwen Yu, Xuanyu Zhang, Youmin Xu, Jian Zhang#

NeurIPS 2023

Paper | GitHub

ICCV 2023
sym

FreeDoM: Training-Free Energy-Guided Conditional Diffusion Model

Jiwen Yu, Yinhuai Wang, Chen Zhao#, Bernard Ghanem, Jian Zhang#

ICCV 2023

Paper | GitHub

ICLR, 2023
sym

Zero-Shot Image Restoration Using Denoising Diffusion Null-Space Model

Yinhuai Wang*, Jiwen Yu*, Jian Zhang#

ICLR 2023 Spotlight

Paper | GitHub | Project Page

Education

M.S., Peking University, VILLA Lab

Advisor: Prof. Jian Zhang

Experiences

2026.09 - Now

Xiaomi Robotics, Beijing, China

2026.01 - 2026.06

Student Researcher at Anuttacon, Mountain View, CA, US

Advisor: Dr. Xin Tong

2024.09 - 2026.01

Student Researcher (Kuai Star) at Kling team, Shenzhen, China

Advisor: Dr. Xintao Wang

2023.04 - 2024.01

Student Researcher at Tencent AI Lab, Shenzhen, China

Advisor: Prof. Xiaodong Cun

Talks

  • Jul 2026 Interactive Video Generation towards World Models
  • Dec 2025 Controllable, Generalizable, and Memory-Enabled: Interactive Video World Models
  • Dec 2025 Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval
  • Oct 2025 Toward Higher-Level Intelligence in Interactive Generative Video for World Model
  • Jul 2025 Toward Higher-Level Intelligence of Interactive Generative Video

Academic Service and Honors

VideoWorldModel CVPR 2026 Workshop

I served as a Primary Organizer of the Video World Models workshop at CVPR 2026. The workshop has now concluded — you can revisit the program and accepted works on the workshop website.

View Recap →