💭 About me

I am a Ph.D. student in Computer Science and Technology at Shanghai Jiao Tong University, advised by Prof. Xiaohong Liu and jointly trained at Shanghai Innovation Institute. Previously, I received my B.Eng. in Software Engineering from Central South University. My research interests include AIGC, video generation, video editing, and multimodal generation.

📖 Education

  • 2026.09 - Present: Jointly Trained Ph.D. Student, Shanghai Innovation Institute.
  • 2026.09 - Present: Ph.D. Student, Computer Science and Technology, Shanghai Jiao Tong University.
  • 2022.09 - 2026.06: B.Eng., Software Engineering, Central South University.

🔥 News

  • [06/2026] One paper has been accepted by European Conference on Computer Vision (ECCV) 2026.
  • [06/2026] One paper has been accepted by IEEE Transactions on Circuits and Systems for Video Technology (TCSVT).
  • [05/2026] Two papers have been submitted to NeurIPS 2026. Wish us luck!
  • [05/2026] One paper has been accepted by ICML 2026.
  • [03/2026] Excited to release A2Edit, a reference-image-guided framework for image restoration and editing with powerful generative capabilities. It has been submitted to ECCV 2026.
  • [02/2026] Excited to release LayerT2V, a unified multi-layer video generation framework. We build the first large-scale multi-layer video dataset. It has been submitted to ICML 2026.
  • [02/2026] One paper has been accepted by CVPR 2026.
  • [12/2025] Excited to release FlowDirector, a training- and inversion-free video editing framework that demonstrates strong performance. It has been submitted to CVPR 2026.
  • [05/2025] One paper has been submitted to TCSVT.

📝 Publications

* equal contribution, † corresponding author

AMATok preview

From Redundancy to Efficiency: Adaptive Video Tokenization via Appearance-Motion Decoupling and Dynamic Token Allocation

Submitted to Conference on Neural Information Processing Systems (NeurIPS) 2026

RewardFlow preview

RewardFlow: Learning to Align Flow Models with Reward Distributions

Submitted to Conference on Neural Information Processing Systems (NeurIPS) 2026

A2-Edit preview

A2-Edit: Precise Reference-Guided Image Editing of Arbitrary Objects and Ambiguous Masks

European Conference on Computer Vision (ECCV) 2026

[paper] [code] [project page]

LayerT2V preview

LayerT2V: A Unified Multi-Layer Video Generation Framework

International Conference on Machine Learning (ICML) 2026

[paper] [code] [project page]

FlowDirector preview

FlowDirector: Training-Free Flow Steering for Precise Text-to-Video Editing

IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2026

[paper] [code] [project page]

SceneVLP preview

SceneVLP: Augmenting Vision-Language Pre-Training Models for Video Action Recognition with Video Scene Graphs

IEEE Transactions on Circuits and Systems for Video Technology (TCSVT)