Yushu Wu
Software Engineer @ Google
I am a Software Engineer at Google, where I work on on-device generative media as part of the On-Device Machine Learning team. My role spans applied ML exploration, product prototyping, and production deployment. Before joining Google, I received my Ph.D. in Computer Engineering from Northeastern University, advised by Prof. Yanzhi Wang.
I build efficient generative video systems that make large-scale diffusion models practical for real-world deployment. My research focuses on reducing latency, memory footprint, and inference cost while preserving high visual fidelity.
My broader research interests span image and video generation, world models, and world action models. I am particularly interested in model–system co-design that makes these models practical in on-device and latency-constrained environments.
Research
-
Efficient Generative Models
D3HR · ICML2025,SnapGen-V · CVPR'25,Streamlined Inference · NeurIPS'24,SF-V · NeurIPS'24,S2DiT · CVPR'26,OmniMem · Preprint. -
Model Compression & Post-Training Optimization
H3AE · Preprint,ToP-ViM · NeurIPS'24,ToR-SSM · EMNLP'24. -
Hardware-Aware Co-Design & Edge Deployment
LOTUS · DAC'24,MOC · ICCAD'23,DACO · DATE'24,Mobile-SR · ECCV'22.
Recent News
Preprint
Selected Publications
Google Scholar for all publications. * means equal contribution.
Work experiences
Software Engineer, On-Device Machine Learning, Core ML, Google
Jul 2026 - Present, Sunnyvale, CAWorking on on-device generative media
Research Intern, Epic Games, Inc.
Jan 2026 - June 2026, Boston, MA (Remote)Worked on World Model
Research Intern, Creative Vision, Snap Inc.
May 2024 - Dec 2025, Santa Monica, CAWorked on efficient text-to-video diffusion
Research Intern, Bose Corporation
Jan 2023 - August 2023, Framingham, MAWorked on audio model quantization for NPU deployment