Hongyu Liu

PhD Candidate

hongyu/hongyu-forest-portrait.webp

Computer Vision · Computer Graphics · Generative AI

I am a Ph.D. student in Computer Science and Engineering at HKUST, advised by Prof. Qifeng Chen.

My research spans video generation and generative world models, human-centered multimodal spatial intelligence, and image generation and editing. I currently focus on post-training world models and embodied intelligence.

Outside the lab, I love fishing and hope to explore China’s rivers, lakes, and coastlines, one fishing trip at a time.

News

Oct 09, 2026 Two papers have been accepted to NeurIPS 2026!
Oct 06, 2026 We released WorldSonus, a streaming video-to-audio model for synchronized, controllable spatial sound.
Sep 03, 2026 We released HelixWorld, a real-time interactive audio-visual world model with synchronized video and spatial sound.
Jul 24, 2026 LiveLight: Real-time Streaming Video Relighting with Interactive Control is accepted to ACM Transactions on Graphics (TOG) 2026.
Jul 09, 2026 We released OPSD-V, an on-policy self-distillation project for post-training few-step autoregressive video generators.

Recent Highlights

MeiGen · Technical Report

OPSD-V

On-policy self-distillation for post-training few-step autoregressive video generators, reducing long-horizon degradation.

Noiz AI · Technical Report

HelixWorld

Real-time interactive world generation with synchronized video and spatial sound that follows your viewpoint.

NVIDIA · NeurIPS 2026

DyaPlex

Full-duplex speech-motion model for dyadic interaction, coupling streaming speech and body motion for responsive digital agents.

Highlighted Research Topics

Three connected directions: video generation and generative world models, human-centered multimodal spatial intelligence, and image generation and editing.

All publications