VIDEO GENERATION · WORLD MODELS

Yubo Huang.

I work on making video generation
interactive, continuous, and efficient.

Yubo Huang

I am a master’s student at the University of Science and Technology of China (USTC), advised by Prof. Enhong Chen at the State Key Laboratory of Cognitive Intelligence. I am currently a research intern on Tencent’s Game World Model Team.

I graduate in summer 2027 and am seeking job opportunities.

Research interests

My current focus is interactive video generation for world models and interactive games. In the longer term, I want to unlock emergent capabilities from native long-video training through compact visual representations and better training infrastructure.

Streaming generationDiffusion distillationEfficient training & inference
01

Selected projects

02

Selected publications

* Equal contribution.

Generated talking avatar in a recording studio
ECCV 2026Spotlight Oral · 1.57%2k+ stars

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Yubo Huang, Hailong Guo, Fangtai Wu, Weiqiang Wang, Shifeng Zhang, Shijie Huang, Qijun Gan, Lin Liu, Sirui Zhao, Enhong Chen, Jiaming Liu, Steven Hoi

Two characters sharing a conversation in a generated video
CVPR 2026Findings

Bind-Your-Avatar: Multi-Character-Talking Video Generation with Dynamic 3D-mask-based Embedding Router

Yubo Huang, Weiqiang Wang, Sirui Zhao, Tong Xu, Lin Liu, Enhong Chen

Self-Forcing training plot showing dynamic collapse while visual quality remains stable
ACM MM 2026Poster

DynaForcing: Overcoming Dynamic Collapse in Self-Forcing Distillation for Streaming Avatar Generation

Yubo Huang*, Sirui Zhao*, Xinchen Yao, Zhengye Zhang, Jinyang Huang, Feng-Qi Cui, Shiwei Wu, Enhong Chen

OmniPersona agent and user taking turns listening and speaking with voice and visual references
Under review

OmniPersona: Persona-Embodied Agents for Multimodal Conversation Interaction

Weiqiang Wang*, Yubo Huang*, Jinnan Chen, Yi Zhang, Qianyi Wu, Yiren Song, Boying Li, Sanghoon Lee, Qiuhong Ke, Jianfei Cai

CollectionLoRA gallery of image effects combined in one LoRA
ECCV 2026Poster

CollectionLoRA: Collecting 50 Effects in 1 LoRA via Multi-teacher On-Policy Distillation

F. Wu, H. Guo, S. Huang, J. Song, Yubo Huang, M. Liu, Z. Wang, Y. Yu, J. Liu, et al.

EditBridge paradigm and examples of faithful detail reconstruction and texture preservation
NeurIPS 2026Poster

EDITBRIDGE: Towards Faithful and Efficient Ultra-High-Resolution Image Editing

Jiayi Song, Shijie Huang, Fangtai Wu, Yubo Huang, Z. Tan, S. Liu, Jiaming Liu, Ruihua Huang

LiveAnimate driving pose and generated dancer side by side
Under review

LiveAnimate: Stable Long-Form Streaming Human Animation in Real-Time

Yuxuan Zhang, Haozhong Xiong, Yubo Huang, Jiayi Song, Jinpeng Yu, Haofan Wang, Jiaming Liu, Ruihua Huang, Liwei Wang

Other publications
03

Experience

Tencent

Jul. 2026 — Present

Game World Model Team · Qingyun Program

I work on continued training and streaming post-training for an interactive video generation project.

Advisors: Dr. Jiaming Liu & Dr. Jingwei Huang

Anuttacon

Jan. 2026 — Jul. 2026

LPM Team

I explored token compression to speed up LPM1.0 sampling, as a direction complementary to step distillation. I proposed cross-compression-ratio trajectory distillation, achieving 5.3× per-step and 1.8× end-to-end acceleration.

Advisors: Dr. Ailing Zeng & Dr. Xin Tong

Alibaba

Aug. 2025 — Jan. 2026

Z-Image Team

I worked on the core R&D of LiveAvatar—model training, inference optimization, audio-video data pipelines, post-training, and the streaming-distillation framework. I also helped bring it to Qwen’s consumer-facing interactive-video product.

Advisors: Dr. Jiaming Liu & Prof. Steven Hoi