我是邵世通(Sutton),香港科技大学(广州)数据科学与分析学域博士生,导师为 Zeke Xie 教授。我的研究聚焦高效生成建模,主要关注视频扩散模型、模型蒸馏、高效采样与可落地的生成系统。
我的工作覆盖从方法设计到产业部署的完整链路。近期研究与工程成果已应用于 First-Intelligence 和 Hedra 的真实视频生成产品,包括 HYVideo-1.5 与 Character-C3。
994+ Google Scholar 引用
English homepage →
研究方向
高效生成建模
我从采样步数、模型规模、实时推理、VAE 加速与稀疏注意力等多个维度研究视频扩散模型的效率。代表项目包括 MagicDistillation、FastLightGen、AMD、PISA 与 LIVEditor。
采样优化与生成质量
我研究初始噪声设计和采样轨迹优化如何在降低推理成本的同时提升生成质量,代表工作包括 IV-Mixed Sampler、Golden Noise 与 CoRe²。
从研究到产品
我致力于将研究方法转化为可部署的视频生成系统,在模型训练、系统优化、在线服务与产品集成方面具有完整实践经验。
精选论文
2026
-
LIVEditor-14B: Lightning Unified Video Editing via In-Context Sparse Attention
Shitong Shao, Zikai Zhou, Haopeng Li, and 4 more authors
In ICMLFirst author , 2026
-
Improved and Accelerated Text-to-Image Generation With Collect, Reflect, and Refine
Shitong Shao, Zikai Zhou, Dian Xie, and 5 more authors
In IEEE TPAMI, 2026
-
FastLightGen: Fast and Light Video Generation with Fewer Steps and Parameters
Shitong Shao, Yufei Gu, and Zeke Xie
In CVPR, 2026
-
Accelerating Diffusion Model Training under Minimal Budgets: A Condensation-Based Perspective
Rui Huang, Shitong Shao, Zikai Zhou, and 6 more authors
In CVPRD²C · * Equal contribution , 2026
2025
-
Golden Noise for Diffusion Models: A Learning Framework
Zikai Zhou, Shitong Shao, Lichen Bai, and 4 more authors
In ICCV, 2025
-
MagicDistillation: Weak-to-Strong Video Distillation for Large-Scale Few-Step Synthesis
Shitong Shao, Hongwei Yi, Hanzhong Guo, and 5 more authors
In arXiv preprint, 2025
-
IV-Mixed Sampler: Leveraging Image Diffusion Models for Enhanced Video Synthesis
Shitong Shao, Zikai Zhou, Lichen Bai, and 2 more authors
In ICLR, 2025
2024
-
Generalized Large-Scale Data Condensation via Various Backbone and Statistical Matching
Shitong Shao, Zeyuan Yin, Muxin Zhou, and 2 more authors
In CVPR Highlight, 2024
-
Elucidating the Design Space of Dataset Condensation
Shitong Shao, Zikai Zhou, and Zhiqiang Shen
In NeurIPSEDC , 2024
2023
-
Teaching What You Should Teach: A Data-Based Distillation Method
Shitong Shao, Huanran Chen, Zhen Huang, and 3 more authors
In IJCAI OralTST , 2023
2022
-
What Role Does Data Augmentation Play in Knowledge Distillation?
Wei Li, Shitong Shao, Weiyan Liu, and 3 more authors
In ACCV OralCCD , 2022
查看全部 56 篇论文 →
研究与产业经历
研究实习生 · 字节跳动 Seed
2026 年 4 月 22 日至今
研究方向:面向高效生成建模的 GAN 蒸馏。
研究科学家实习生 · First-Intelligence
2025 年 10 月 – 2026 年 4 月 1 日
负责 HYVideo-1.5 的实时生成与高效部署研究,包括四步视频蒸馏和 4.5× 加速的 VAE。
研究科学家实习生 · Hedra
2024 年 10 月 – 2025 年 7 月
构建并产品化 Character-C3 的核心蒸馏流程,用于少步数说话视频生成。
研究实习生 · MBZUAI,申志强研究组
2023 年 7 月 – 2024 年 3 月
研究大规模数据浓缩与数据优化,成果包括 G-VBSM(CVPR Highlight)与 EDC(NeurIPS)。
早期研究与工程实习
曾在 OPPO、上海人工智能实验室、北京理工大学与一流科技参与数据浓缩、AI 编译器、知识蒸馏与模型工程研究。
查看完整简历 →