Yuang Ai

Hi everyone, welcome to my website!

I'm Yuang Ai, a first-year PhD student at MMLab, the Chinese University of Hong Kong, supervised by Prof. Xiangyu Yue. I am also a research intern at ByteDance Seed, advised by Yi Jiang.

My recent research interests primarily focus on topics with significant real-world applications, like foundation generative models, unified multimodal models, etc.

I am open to any discussion or collaboration. If you are interested, please feel free to contact me via email.

Email  /  Google Scholar  /  Github

profile photo
Selected Publications
BitDance preview BitDance: Scaling Autoregressive Generative Models with Binary Tokens
Yuang Ai*, Jiaming Han*, Shaobin Zhuang*, Weijia Mao, Xuefeng Hu, Ziyan Yang, Zhenheng Yang, Yali Wang, Huaibo Huang, Xiangyu Yue, and Hao Chen.
Preprint, 2026
paper / code GitHub stars
UniWeTok preview UniWeTok: An Unified Binary Tokenizer with Codebook Size 2128 for Unified Multimodal Large Language Model
Shaobin Zhuang*, Yuang Ai*, Jiaming Han*, Weijia Mao, Xiaohui Li, Fangyikang Wang, Xiao Wang, Yan Li, Shanchuan Lin, Kun Xu, Zhenheng Yang, Huaibo Huang, Xiangyu Yue, Hao Chen, and Yali Wang.
Preprint, 2026
paper / code
DiCo preview DiCo: Revitalizing ConvNets for Scalable and Efficient Diffusion Modeling
Yuang Ai, Qihang Fan, Xuefeng Hu, Zhenheng Yang, Ran He, and Huaibo Huang.
NeurIPS Spotlight, 2025
paper / code
LAformer architecture preview Fast and Accurate Image Restoration and Generation with Rank Enhanced Linear Attention
Yuang Ai, Huaibo Huang, Tao Wu, Qihang Fan, and Ran He.
ECCV, 2026
paper / code
DreamClear preview DreamClear: High-Capacity Real-World Image Restoration with Privacy-Safe Dataset Curation
Yuang Ai, Xiaoqiang Zhou, Huaibo Huang, Xiaotian Han, Zhengyu Chen, Quanzeng You, and Hongxia Yang.
NeurIPS, 2024
paper / code GitHub stars
SODA preview Uncertainty-Aware Source-Free Adaptive Image Super-Resolution with Wavelet Augmentation Transformer
Yuang Ai, Xiaoqiang Zhou, Huaibo Huang, Lei Zhang, and Ran He.
CVPR, 2024
paper / code
MPerceiver preview Multimodal Prompt Perceiver: Empower Adaptiveness, Generalizability and Fidelity for All-in-One Image Restoration
Yuang Ai, Xiaoqiang Zhou, Huaibo Huang, Jiexiang Wang, and Ran He.
CVPR, 2024
paper / code
Experiences
ByteDance Seed
2026.07 - Present
World Models & Video Generation
ByteDance TikTok
2025.02 - 2026.07
Foundation Generative Models & Unified Multimodal Models
ByteDance Seed
2024.03 - 2025.02
Vision-Language Models & Applications of Generative Models
Honours and Awards
Academic Service
Teaching