Recent papers, releases, and milestones.
2026 · 08 We release VibeGame , an agentic framework for game development.
2026 · 08 We release AURORA-LM , a diffusion language model for continuous-latent text modeling.
2026 · 07 We release ACE-Data-0 , a human-centric ambient capture engine for embodied AI.
2026 · 07 We release Apple-PI , a benchmark for evaluating physical understanding in video models.
2026 · 06 HSImul3R , our work on real-to-sim-to-real modeling, is accepted by ECCV 2026.
2026 · 06 AniFeats , our work on human animation, is accepted by TVCG 2026.
2026 · 03 GS-VTON , our work on 3D virtual try-on, is accepted by IJCV 2026.
2026 · 01 IGGT , our work on instance-grounded 3D semantic reconstruction, is accepted by ICLR 2026. Congratulations to Hao Li!
2025 · 09 CrowdMoGen , our work on collective motion generation, is accepted by IJCV 2025.
2025 · 09 Wukong , our work on 3D morphing, is accepted by NeurIPS 2025. Congratulations to Minghao!
2025 · 07 We release a survey on reconstructing 4D spatial intelligence .
2025 · 06 FreeMorph , our training-free approach to 2D image morphing, is accepted by ICCV 2025.
2025 · 05 We release CrowdMoGen for zero-shot collective 3D human motion generation.
2025 · 03 One paper is accepted by SIGGRAPH 2025. Congratulations to Minghao!
2025 · 02 Two papers, ArtiFade and AudCast , are accepted by CVPR 2025.
2025 · 01 AvatarGO is accepted by ICLR 2025.
2024 · 10 We release AvatarGO for zero-shot 4D human-object interaction generation and animation.
2024 · 10 We release GS-VTON for controllable 3D virtual try-on.
2024 · 09 We release ArtiFade for generating high-quality subjects from blemished images.
2024 · 08 I begin a new chapter at NTU, working with Prof. Ziwei Liu.
2024 · 06 We post a survey on 3D human avatar modeling .
2024 · 04 I successfully defend my Ph.D.
2024 · 02 DreamAvatar is accepted by CVPR 2024.
2023 · 09 HeadSculpt is accepted by NeurIPS 2023.
2023 · 08 We release Guide3D for transferring multi-view generated images to 3D avatars.
2023 · 06 We release HeadSculpt for editable 3D head generation from text.
2023 · 04 We release DreamAvatar for controllable 3D avatar generation from text.
2023 · 03 MA-NeRF is accepted by ICME 2023.
2023 · 02 SeSDF is accepted by CVPR 2023.
2022 · 03 JIFF is accepted by CVPR 2022.
Show all updates
SELECTED WORK
Publications
* Equal contribution · # Corresponding author
Under submission Agentic game modeling
VibeGame: Prompt-to-Game Development with AI-Native Engine and Self-Evolving Adversarial Agent Team
Wenbo Hu* , Ken Li* , Jiazhe Wei* , Yukang Cao , Weiyi Hong, Jiayi Dai , Chenjun Bai, Jiajun Liang, Yucheng Liao, Ruichuan An, Zeyu Lou, Haofan Wang, Wende Tan, Liucheng Guo, Yueming Lyu, Ziwei Liu , Chenyang Si#
Under submission, 2026
Technical Report Diffusion language modeling
AURORA-LM: Autoencoding Unified Representation for Continuous-Latent Diffusion Language Modeling
Jiajun Liang* , Yucheng Liao* , Yukang Cao* , Jiazhe Wei , Ken Li , Wende Tan, Jiankun Zhang, ZY Cui, Jingkang Yang , Liucheng Guo, Shiqi Yang, B. Yang, Caifeng Shan , Ziwei Liu , Chenyang Si#
Technical Report, 2026
Technical Report Ambient Capture Engine for Embodied AI
ACE-Data-0: Human-Centric Ambient Capture as Embodied Data Engine
Yukang Cao* , Haozhe Xie* , Beichen Wen* , Runmao Yao , Yinghao Liu, Yue Huang, Zhichao Liao, Yunxiang Wang, Haiheng Liu, Xingshun Tian, Dawei Su, Long Zhuo , Dacheng Tao , Xiaogang Wang , Liang Pan# , Ziwei Liu#
Technical Report, 2026
Under submission Physics benchmark for video models
Apple-PI: Benchmarking Thinking with Video Towards Law-Grounded Physical Intelligence
Runmao Yao* , Kairui Hu* , Yukang Cao , Ruisi Wang , Shulin Tian , Ziang Cao , Weichen Fan , Ziqi Huang , Yuhao Dong , Hao Li , Zhaoxi Chen , Zhongang Cai , Lei Yang , Ziwei Liu#
Under submission, 2026
TVCG 2026 Video-based human animation
AniFeats: Animate 3D Feature Meshes for Character Video Generation
Beijia Lu , Zekai Gu , Zhiyang Dou , Haotian Yuan , Yuan Liu , Chenyang Si , Yukang Cao , Yuming Jiang , Wenping Wang , Ziwei Liu#
IEEE Transactions on Visualization and Computer Graphics, 2026
IJCV 2026 3D virtual try-on
GS-VTON: Controllable 3D Virtual Try-on with Gaussian Splatting
Yukang Cao* , Masoud Hadi* , Liang Pan# , Ziwei Liu#
International Journal of Computer Vision, 2026
ICLR 2026 3D reconstruction
IGGT: Instance-Grounded Geometry Transformer for Semantic 3D Reconstruction
Hao Li* , Zhengyu Zou* , Fangfu Liu* , Xuanyang Zhang , Fangzhou Hong , Yukang Cao , Yushi Lan , Manyuan Zhang , Gang Yu , Dingwen Zhang# , Ziwei Liu
International Conference on Learning Representations, 2026
NeurIPS 2025 3D morphing
Wukong’s 72 Transformations: High-fidelity 3D Morphing via Flow Models
Minghao Yin , Yukang Cao , Kai Han#
Neural Information Processing Systems, 2025
Preprint 2025 Survey
Reconstructing 4D Spatial Intelligence: A Survey
Yukang Cao , Jiahao Lu , Zhisheng Huang , Zhuowen Shen , Chengfeng Zhao , Fangzhou Hong , Zhaoxi Chen , Xin Li , Wenping Wang , Yuan Liu# , Ziwei Liu#
Preprint, 2025
ICCV 2025 Image morphing
FreeMorph: Tuning-Free Generalized Image Morphing with Diffusion Model
Yukang Cao , Chenyang Si , Jinghao Wang , Ziwei Liu#
International Conference on Computer Vision, 2025
SIGGRAPH 2025 4D generation
Splat4D: Diffusion-Enhanced 4D Gaussian Splatting for Temporally and Spatially Consistent Content Creation
Minghao Yin , Yukang Cao , Songyou Peng , Kai Han#
ACM SIGGRAPH, 2025
CVPR 2025 Human video
AudCast: Audio-Driven Human Video Generation by Cascaded Diffusion Transformers
Jiazhi Guan , Kaisiyuan Wang, Zhiliang Xu, Quanwei Yang, Yasheng Sun, Shengyi He, Borong Liang, Yukang Cao , Yingying Li, Haocheng Feng, Errui Ding, Jingdong Wang, Youjian Zhao# , Hang Zhou# , Ziwei Liu#
IEEE Conference on Computer Vision and Pattern Recognition, 2025
ICLR 2025 Human-object interaction
AvatarGO: Zero-shot 4D Human-Object Interaction Generation and Animation
Yukang Cao , Liang Pan# , Kai Han , Kwan-Yee K. Wong , Ziwei Liu#
International Conference on Learning Representations, 2025
CVPR 2025 Subject generation
ArtiFade: Learning to Generate High-quality Subject from Blemished Images
Shuya Yang* , Shaozhe Hao* , Yukang Cao# , Kwan-Yee K. Wong#
IEEE Conference on Computer Vision and Pattern Recognition, 2025
Preprint 2024 Survey
A Survey on 3D Human Avatar Modeling — From Reconstruction to Generation
Ruihe Wang* , Yukang Cao*# , Kai Han , Kwan-Yee K. Wong
Preprint, 2024
CVPR 2024 Avatar generation
DreamAvatar: Text-and-Shape Guided 3D Human Avatar Generation via Diffusion Models
Yukang Cao* , Yan-Pei Cao* , Kai Han# , Ying Shan, Kwan-Yee K. Wong
IEEE Conference on Computer Vision and Pattern Recognition, 2024
CVPR 2023 Human reconstruction
SeSDF: Self-evolved Signed Distance Field for Implicit 3D Clothed Human Reconstruction
Yukang Cao , Kai Han , Kwan-Yee K. Wong
IEEE Conference on Computer Vision and Pattern Recognition, 2023
CVPR 2022 · Oral Human reconstruction
JIFF: Jointly-aligned Implicit Face Function for High Quality Single View Clothed Human Reconstruction
Yukang Cao , Guanying Chen , Kai Han , Wenqi Yang , Kwan-Yee K. Wong
IEEE Conference on Computer Vision and Pattern Recognition, 2022
PROFESSIONAL BACKGROUND
Employment
Research collaborations across academia and industry.
Primary appointments
Aug 2024–Present Research Fellow · MMLab@NTU Worked with Prof. Ziwei Liu .
Sep 2020–Jun 2024
Research internships
Mar–Jun 2024
Dec 2022–Jun 2023 Research Intern · Tencent PCG, ARC Lab Worked with Dr. Yan-Pei Cao .
ACADEMIC BACKGROUND
Education
Doctoral and undergraduate education in computer science and engineering.
Degrees
2020–2024 Ph.D. in Computer Science · The University of Hong Kong Advised by Prof. Kwan-Yee K. Wong .
2016–2020 Bachelor of Engineering · Zhejiang University Graduated with a class rank of 3/66.
Early research experience
Jul–Nov 2019 Research Student · The University of Texas at Austin Worked with Dr. Chandrajit Bajaj .
Sep–Dec 2019 Research Student · Peking University Worked with Dr. Dehua Xiong .
Jul–Oct 2018 Research Student · Tsinghua University Worked with Dr. Zuoqiang Shi .