Core Contributor
LingBot-World-Infinity
A real-time world model for open-ended interaction through camera control, actions, and text.
Core Contributor
A real-time world model for open-ended interaction through camera control, actions, and text.
First author and lead researcher
Interactive image editing from free-form strokes, color, edge, and layered visual cues.
MagicQuill V1 became one of Hugging Face's most visited editing projects and earned broad open-source adoption. V2 extends the same interaction language with layered visual cues.
First- and co-first-author papers.
Bringing text to life through video diffusion priors.
Interactive image editing from free-form strokes, color, and edge cues.
3.7K stars · Nearly 8M demo visitsPrecise interactive image editing with layered visual cues.
The second generation of the MagicQuill research lineReal-time autoregressive generation for multi-shot video narratives.
Causal video generation · Multi-shot storytellingI am a Ph.D. student at HKUST, advised by Prof. Qifeng Chen, and a research intern at Ant Group. My work focuses on controllable generation, interactive systems, and real-time visual models.
I enjoy working across the full stack of a research project, from model training and post-training to interfaces and deployment. Recent projects include MagicQuill, CausalCine, and LingBot-World-Infinity.
Experience
Research Intern
Interactive image editing, world models, and real-time generative systems.Machine Learning Engineer Intern
Machine learning engineering.Education
Ph.D. in Computer Science and Engineering
Visual Intelligence Lab · Advised by Prof. Qifeng ChenB.Sc. in Data Science and Technology
Graduated in the top 5%.