I am an M.Sc. student in Computer Engineering at the National University of Singapore (NUS), where I conduct research at Show Lab under the supervision of Prof. Mike Zheng Shou.

My research focuses on video generation, world models, and robotics, with an emphasis on building generative models that can understand, simulate, and control interactive visual worlds. I am currently a Research Intern at Tencent IEG, working on algorithms for game world models. I received my bachelor’s degree from the School of Electronic Information at Wuhan University.

🎬 Video Generation 🌍 World Models 🤖 Robotics

🎓 Education & Experience

Research Intern
Tencent Interactive Entertainment Group (IEG)
Research on game world model algorithms · Present
M.Sc. in Computer Engineering
National University of Singapore · Show Lab
Advisor: Prof. Mike Zheng Shou · Present
Bachelor's Degree
School of Electronic Information, Wuhan University

✨ News

  • Sep. 2026Supervise What Survives was accepted to CoRL 2026! 🎉
  • Sep. 2026 — Released H3-World, an efficient framework that turns language understanding in a large video generator into temporally grounded world control.
  • Jun. 2026 — Released Supervise What Survives, our work on geometry-guided VLA adaptation from synthetic robot videos.
  • Oct. 2025LayerTracer was presented as an Oral paper at ICCV 2025.

📚 Publications

H3-World: Turning Language Understanding into World Control

Technical Report

Danze Chen, Zeqing Wang, Ziyue Lin, Xingyi Yang, Yeying Jin

Turns the language understanding already learned by a large video generator into precise, temporally grounded character and camera control with lightweight adaptation.

Supervise What Survives: Geometry-Guided VLA Adaptation from Synthetic Robot Videos

CoRL 2026

Danze Chen, Yanzhe Chen, Qiming Huang, Zhijun Cao, Chen Gao, Mike Zheng Shou

Introduces geometry-guided representation alignment for adapting vision-language-action models from synthetic robot videos while keeping low-level control grounded in real demonstrations.

WorldMind: Decoupled Game World Model for State-Aware NPC Behavior

CCF-A Conference Under Review

Zhiyang Deng, Boran Zhang, Danze Chen, Yeying Jin

Decouples state understanding, NPC decision-making, action control, and video generation, then reconnects them in a closed interaction loop for state-aware NPC behavior.

ReactiveGWM: Steering NPC in Reactive Game World Models

CCF-A Conference Under Review

Zeqing Wang, Danze Chen, Zhaohu Xing, Zizhao Tong, Yinhan Zhang, Xingyi Yang, Yeying Jin

Separates low-level player control from high-level NPC strategy, enabling controllable interactions and zero-shot strategy transfer across game world models.

LayerTracer text-to-SVG generation and layer-wise vectorization examples

LayerTracer: Cognitive-Aligned Layered SVG Synthesis via Diffusion Transformer

ICCV 2025 Oral

Yiren Song, Danze Chen, Mike Zheng Shou

Learns cognitively aligned layer-by-layer SVG construction with a diffusion transformer, producing editable vector graphics with meaningful semantic structure.

View the full publication list →

✉️ Contact

I am happy to connect with researchers and collaborators working on generative world models and embodied intelligence.