|
Yuang Ai
Hi everyone, welcome to my website!
I'm Yuang Ai, a first-year PhD student at MMLab, the Chinese University of Hong Kong, supervised by Prof. Xiangyu Yue. I am also a research intern at ByteDance Seed, advised by Yi Jiang.
My recent research interests primarily focus on topics with significant real-world applications, like foundation generative models, unified multimodal models, etc.
I am open to any discussion or collaboration. If you are interested, please feel free to contact me via email.
Email  / 
Google Scholar  / 
Github
|
|
|
BitDance: Scaling Autoregressive Generative Models with Binary Tokens
Yuang Ai*, Jiaming Han*, Shaobin Zhuang*, Weijia Mao, Xuefeng Hu, Ziyan Yang, Zhenheng Yang, Yali Wang, Huaibo Huang, Xiangyu Yue, and Hao Chen.
Preprint, 2026
paper /
code
|
|
UniWeTok: An Unified Binary Tokenizer with Codebook Size 2128 for Unified Multimodal Large Language Model
Shaobin Zhuang*, Yuang Ai*, Jiaming Han*, Weijia Mao, Xiaohui Li, Fangyikang Wang, Xiao Wang, Yan Li, Shanchuan Lin, Kun Xu, Zhenheng Yang, Huaibo Huang, Xiangyu Yue, Hao Chen, and Yali Wang.
Preprint, 2026
paper /
code
|
|
DiCo: Revitalizing ConvNets for Scalable and Efficient Diffusion Modeling
Yuang Ai, Qihang Fan, Xuefeng Hu, Zhenheng Yang, Ran He, and Huaibo Huang.
NeurIPS Spotlight, 2025
paper /
code
|
|
Fast and Accurate Image Restoration and Generation with Rank Enhanced Linear Attention
Yuang Ai, Huaibo Huang, Tao Wu, Qihang Fan, and Ran He.
ECCV, 2026
paper /
code
|
|
DreamClear: High-Capacity Real-World Image Restoration with Privacy-Safe Dataset Curation
Yuang Ai, Xiaoqiang Zhou, Huaibo Huang, Xiaotian Han, Zhengyu Chen, Quanzeng You, and Hongxia Yang.
NeurIPS, 2024
paper /
code
|
|
Uncertainty-Aware Source-Free Adaptive Image Super-Resolution with Wavelet Augmentation Transformer
Yuang Ai, Xiaoqiang Zhou, Huaibo Huang, Lei Zhang, and Ran He.
CVPR, 2024
paper /
code
|
|
Multimodal Prompt Perceiver: Empower Adaptiveness, Generalizability and Fidelity for All-in-One Image Restoration
Yuang Ai, Xiaoqiang Zhou, Huaibo Huang, Jiexiang Wang, and Ran He.
CVPR, 2024
paper /
code
|
|
ByteDance Seed
2026.07 - Present
World Models & Video Generation
|
|
ByteDance TikTok
2025.02 - 2026.07
Foundation Generative Models & Unified Multimodal Models
|
|
ByteDance Seed
2024.03 - 2025.02
Vision-Language Models & Applications of Generative Models
|
- 2024.12: National scholarship
- 2024.12: NeurIPS 2024 Top Reviewer (main track)
- 2023.07: Outstanding Graduate named by Beijing and by BIT
- 2023.05: Second-Place Winner in the NTIRE 2023 Challenge on Image Super-Resolution
- 2022.12: Second-Place Winner in the NTIRE 2023 Challenge on 360deg Omnidirectional Image and Video Super-Resolution
- 2022.12: National scholarship
|
- Reviewer for TPAMI, TIP, CVPR, ICCV, NeurIPS, ICLR, ICML.
- Programme Committee for AAAI.
|
- Deep Learning Methods and Applications, Teaching Assistant - Fall 2024
|
|