Shitong Shao

生成式人工智能研究者 · 香港科技大学(广州)博士生

shitong-shao-anime.png

香港科技大学(广州)

中国广州

我是邵世通(Sutton),香港科技大学(广州)数据科学与分析学域博士生,导师为 Zeke Xie 教授。我的研究聚焦高效生成建模,主要关注视频扩散模型、模型蒸馏、高效采样与可落地的生成系统。

我的工作覆盖从方法设计到产业部署的完整链路。近期研究与工程成果已应用于 First-IntelligenceHedra 的真实视频生成产品,包括 HYVideo-1.5 与 Character-C3。

994+ Google Scholar 引用

研究方向

高效生成建模

我从采样步数、模型规模、实时推理、VAE 加速与稀疏注意力等多个维度研究视频扩散模型的效率。代表项目包括 MagicDistillation、FastLightGen、AMD、PISALIVEditor

采样优化与生成质量

我研究初始噪声设计和采样轨迹优化如何在降低推理成本的同时提升生成质量,代表工作包括 IV-Mixed Sampler、Golden NoiseCoRe²

从研究到产品

我致力于将研究方法转化为可部署的视频生成系统,在模型训练、系统优化、在线服务与产品集成方面具有完整实践经验。

精选论文

2026

  1. ICML
    LIVEditor-14B: Lightning Unified Video Editing via In-Context Sparse Attention
    Shitong Shao, Zikai Zhou, Haopeng Li, and 4 more authors
    In ICMLFirst author , 2026
  2. IEEE TPAMI
    Improved and Accelerated Text-to-Image Generation With Collect, Reflect, and Refine
    Shitong Shao, Zikai Zhou, Dian Xie, and 5 more authors
    In IEEE TPAMI, 2026
  3. CVPR
    FastLightGen: Fast and Light Video Generation with Fewer Steps and Parameters
    Shitong Shao, Yufei Gu, and Zeke Xie
    In CVPR, 2026
  4. CVPR
    Accelerating Diffusion Model Training under Minimal Budgets: A Condensation-Based Perspective
    Rui Huang, Shitong Shao, Zikai Zhou, and 6 more authors
    In CVPRD²C · * Equal contribution , 2026

2025

  1. ICCV
    Golden Noise for Diffusion Models: A Learning Framework
    Zikai Zhou, Shitong Shao, Lichen Bai, and 4 more authors
    In ICCV, 2025
  2. arXiv preprint
    MagicDistillation: Weak-to-Strong Video Distillation for Large-Scale Few-Step Synthesis
    Shitong Shao, Hongwei Yi, Hanzhong Guo, and 5 more authors
    In arXiv preprint, 2025
  3. ICLR
    IV-Mixed Sampler: Leveraging Image Diffusion Models for Enhanced Video Synthesis
    Shitong Shao, Zikai Zhou, Lichen Bai, and 2 more authors
    In ICLR, 2025

2024

  1. CVPR Highlight
    Generalized Large-Scale Data Condensation via Various Backbone and Statistical Matching
    Shitong Shao, Zeyuan Yin, Muxin Zhou, and 2 more authors
    In CVPR Highlight, 2024
  2. NeurIPS
    Elucidating the Design Space of Dataset Condensation
    Shitong Shao, Zikai Zhou, and Zhiqiang Shen
    In NeurIPSEDC , 2024

2023

  1. IJCAI Oral
    Teaching What You Should Teach: A Data-Based Distillation Method
    Shitong Shao, Huanran Chen, Zhen Huang, and 3 more authors
    In IJCAI OralTST , 2023

2022

  1. ACCV Oral
    What Role Does Data Augmentation Play in Knowledge Distillation?
    Wei Li, Shitong Shao, Weiyan Liu, and 3 more authors
    In ACCV OralCCD , 2022

查看全部 56 篇论文 →

研究与产业经历

研究实习生 · 字节跳动 Seed
2026 年 4 月 22 日至今
研究方向:面向高效生成建模的 GAN 蒸馏。

研究科学家实习生 · First-Intelligence
2025 年 10 月 – 2026 年 4 月 1 日
负责 HYVideo-1.5 的实时生成与高效部署研究,包括四步视频蒸馏和 4.5× 加速的 VAE。

研究科学家实习生 · Hedra
2024 年 10 月 – 2025 年 7 月
构建并产品化 Character-C3 的核心蒸馏流程,用于少步数说话视频生成。

研究实习生 · MBZUAI,申志强研究组
2023 年 7 月 – 2024 年 3 月
研究大规模数据浓缩与数据优化,成果包括 G-VBSM(CVPR Highlight)与 EDC(NeurIPS)。

早期研究与工程实习
曾在 OPPO、上海人工智能实验室、北京理工大学与一流科技参与数据浓缩、AI 编译器、知识蒸馏与模型工程研究。

查看完整简历 →