Publications
Symbols (†) denotes student co-authors.
Beyond Pixel Histories: World Models with Persistent 3D State
Samuel Garcin, Thomas Walker, Steven McDonagh, Tim Pearce, Hakan Bilen, Tianyu He, Kaixin Wang, Jiang Bian
International Conference on Machine Learning (ICML) 2026
LIVE: Long-horizon Interactive Video World Modeling
Junchao Huang†, Ziyang Ye†, Xinting Hu, Tianyu He, Guiyu Zhang†, Shaoshuai Shi, Jiang Bian, Li Jiang
International Conference on Machine Learning (ICML) 2026
Luminark: Training-free, Probabilistically-Certified Watermarking for General Vision Generative Models
Jiayi Xu†, Zhang Zhang†, Yuanrui Zhang†, Ruitao Chen†, Yixian Xu†, Tianyu He, Di He
arXiv preprint arXiv:2601.01085 2026
Quotient-Space Diffusion Models
Yixian Xu†, Yusong Wang†, Shengjie Luo†, Kaiyuan Gao†, Tianyu He, Di He, Chang Liu
International Conference on Learning Representations (ICLR) Oral 2026
Geometry Forcing: Marrying Video Diffusion and 3D Representation for Consistent World Modeling
Haoyu Wu†, Diankun Wu†, Tianyu He, Junliang Guo, Yang Ye†, Yueqi Duan, Jiang Bian
International Conference on Learning Representations (ICLR) 2026
Fast Autoregressive Video Generation with Diagonal Decoding
Yang Ye†, Junliang Guo, Haoyu Wu†, Tianyu He, Tim Pearce, Tabish Rashid, Katja Hofmann, Jiang Bian
Findings of the Computer Vision and Pattern Recognition Conference (CVPR) 2026
Memory Forcing: Spatio-Temporal Memory for Consistent Scene Generation on Minecraft
Junchao Huang†, Xinting Hu, Boyao Han†, Shaoshuai Shi, Zhuotao Tian, Tianyu He, Li Jiang
arXiv preprint arXiv:2510.03198 2025
Reinforcement Learning with Inverse Rewards for World Model Post-training
Yang Ye†, Tianyu He, Shuo Yang†, Jiang Bian
arXiv preprint arXiv:2509.23958 2025
3D-LLaVA: Towards Generalist 3D LMMs with Omni Superpoint Transformer
Jiajun Deng, Tianyu He, Li Jiang, Tianyu Wang, Feras Dayoub, Ian Reid
Proceedings of the Computer Vision and Pattern Recognition Conference (CVPR) 2025
VidTwin: Video VAE with Decoupled Structure and Dynamics
Yuchi Wang†, Junliang Guo, Xinyi Xie†, Tianyu He, Xu Sun, Jiang Bian
Proceedings of the Computer Vision and Pattern Recognition Conference (CVPR) 2025
Video in-context Learning: Autoregressive Transformers are Zero-Shot Video Imitators
Wentao Zhang†, Junliang Guo, Tianyu He, Li Zhao, Linli Xu, Jiang Bian
International Conference on Learning Representations (ICLR) 2025
InstructAvatar: Text-Guided Emotion and Motion Control for Avatar Generation
Yuchi Wang†, Junliang Guo, Jianhong Bai†, Runyi Yu†, Tianyu He, Xu Tan, Xu Sun, Jiang Bian
Proceedings of the AAAI Conference on Artificial Intelligence (AAAI) 2025
UniEdit: A Unified Tuning-Free Framework for Video Motion and Appearance Editing
Jianhong Bai†, Tianyu He, Yuchi Wang, Junliang Guo, Haoji Hu, Zuozhu Liu, Jiang Bian
Proceedings of the ACM International Conference on Multimedia (ACM MM) 2025
VidTok: A Versatile and Open-Source Video Tokenizer
Anni Tang†, Tianyu He, Junliang Guo, Xinle Cheng†, Li Song, Jiang Bian
arXiv preprint arXiv:2412.13061 2024
IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI
Xiaoyu Chen†, Junliang Guo, Tianyu He, Chuheng Zhang, Pushi Zhang, Derek Cathera Yang, Li Zhao, Jiang Bian
arXiv preprint arXiv:2411.00785 2024
End-to-End Rate-Distortion Optimized 3D Gaussian Representation
Henan Wang†, Hanxin Zhu†, Tianyu He, Runsen Feng†, Jiajun Deng, Jiang Bian, Zhibo Chen
The European Conference on Computer Vision (ECCV) 2024
Is Vanilla MLP in Neural Radiance Field Enough for Few-shot View Synthesis?
Hanxin Zhu†, Tianyu He, Xin Li, Bingchen Li†, Zhibo Chen
Proceedings of the Computer Vision and Pattern Recognition Conference (CVPR) 2024
GAIA: Zero-Shot Talking Avatar Generation
Tianyu He, Junliang Guo, Runyi Yu†, Yuchi Wang†, Jialiang Zhu†, Kaikai An†, Leyi Li†, Xu Tan, Chunyu Wang, Han Hu, HsiangTao Wu, Sheng Zhao, Jiang Bian
International Conference on Learning Representations (ICLR) 2024
DAE-Talker: High Fidelity Speech-Driven Talking Face Generation with Diffusion Autoencoder
Chenpeng Du†, Qi Chen†, Tianyu He, Xu Tan, Xie Chen, Kai Yu, Sheng Zhao, Jiang Bian
Proceedings of the ACM International Conference on Multimedia (ACM MM) 2023
HiFace: High-Fidelity 3D Face Reconstruction by Learning Static and Dynamic Details
Zenghao Chai†, Tianke Zhang†, Tianyu He, Xu Tan, Tadas Baltruvsaitis, HsiangTao Wu, Runnan Li, Sheng Zhao, Chun Yuan, Jiang Bian
Proceedings of the International Conference on Computer Vision (ICCV) 2023