Jinpeng Yu

Senior Researcher & Content Generation Algorithm Lead
Work Email: yujinpeng.yjp (at) alibaba-inc.com

I’m always open to research collaborations and happy to mentor motivated students.
If you’re interested in the following areas, feel free to reach out!

Real-Time Controllable Video Generation · Content Creation Agent Systems
Biography

Jinpeng Yu (于劲鹏) is currently the Content Generation Algorithm Lead for the Qwen App at Alibaba Group, where he focuses on multimodal AI agent systems and generative models for image and video creation. Previously, he was a Senior Researcher on the AIGC team at Xiaohongshu, where he worked on controllable image generation and multimodal understanding. He received his bachelor's degree with honors from Harbin Institute of Technology (HIT) in 2021. Subsequently, he earned his master's degree from SIST, ShanghaiTech University in 2024 under the supervision of Prof. Shenghua Gao (高盛华, HKU).

Selected Publications (First Author or Project Lead; View Full List)

* indicates equal contribution. † indicates project lead.

Video Generation

UniSwap audio-visual identity swapping overview
Yuxuan Zhang, Haozhong Xiong, Jiayi Song, Jinpeng Yu, Yang Shi, Jiaming Liu, Ruihua Huang, and Liwei Wang.
Preprint, August 2026

A streaming framework that jointly replaces visual identity and vocal timbre in talking videos while preserving motion, scene content, and audio-video synchronization.

LiveAnimate long-form streaming human animation overview
Yuxuan Zhang, Haozhong Xiong, Yubo Huang, Jiayi Song, Jinpeng Yu, Jiaming Liu, Ruihua Huang, and Liwei Wang.
Preprint, August 2026

Real-time, causal human animation from a reference image and streaming pose controls, designed to remain stable over long-form generation.

AVA-Encoder agent-native video representation learning overview
Chuyue Li, Jinpeng Yu, Haozhe Wang, Xueyun Tian, Zhijing Zhang, Bingnan Li, Shuqi Gu, Kan Ren, Jiaming Liu, and Ruihua Huang.
Preprint, August 2026

An agentic video auto-encoding framework that transforms films into structured knowledge graphs and learns agent-native video representations through reconstruction-driven optimization.

Rethinking classifier-free guidance in on-policy diffusion distillation overview
Bingnan Li, Haozhe Wang, Haozhong Xiong, Fangtai Wu, Jinpeng Yu, Yang Shi, Jiaming Liu, and Ruihua Huang.
Preprint, July 2026

Introduces branch-aware Positive-Direction Matching to address negative-branch asymmetry and improve guidance-scale robustness in on-policy diffusion distillation.

Image Generation

Visually aligned image-editing follow-up suggestions overview
Zhijing Zhang, Jinpeng Yu, Xin Song, Bingnan Li, Chuyue Li, Changhui Du, Xiaolin Fang, Jiaming Liu, and Ruihua Huang.
Preprint, August 2026

A three-stage multimodal framework for recommending useful, diverse, and visually consistent follow-up edits in image-creation conversations.

SearchGen agentic visual generation overview
Haozhe Wang, Weijia Feng, Jinpeng Yu, Che Liu, Ping Nie, Fangzhen Lin, Jiaming Liu, Ruihua Huang, Jimmy Lin, Wenhu Chen, and Cong Wei.
Preprint, July 2026 (#3 Paper of the Day)

An agentic visual generation framework that retrieves external visual knowledge for knowledge-intensive prompts and decides when search is necessary.

SSR-Encoder subject-driven generation overview
Yuxuan Zhang, Jiaming Liu, Yiren Song, Rui Wang, Jinpeng Yu, Hao Tang, Huaxia Li, Xu Tang, Yao Hu, Han Pan, and Zhongliang Jing.
CVPR, 2024

Encodes selective subject representation from reference images for subject-driven image generation.

3D Generation

Pseudo-plane regularized SDF architecture
Jinpeng Yu, Jing Li, Ruoyu Wang, and Shenghua Gao.
IJCV, 2024

Pseudo-plane regularization for high-fidelity SDF-based reconstruction of texture-less indoor scenes.

GeoFormer point cloud completion overview
Jinpeng Yu, Binbin Huang, Yuxuan Zhang, Huaxia Li, Xu Tang, and Shenghua Gao.
ACM MM, 2024

A tri-plane integrated transformer for accurate and detailed point cloud completion.

Services
  • Journal Reviewer of: IJCV
  • Conference Reviewer of: CVPR, ICCV, ECCV, ACM MM, NeurIPS, ICLR
  • Selected Honors
  • Academic Excellence Scholarship, ShanghaiTech (2022, 2023, 2024)
  • Outstanding Graduate, HIT (2021, top 1%)
  • National Scholarship, Ministry of Education, People's Republic of China (2019, 2020, top 1%)
  • Passions (Discovering, Recording, and Creating Beauty)
  • Traveling Around, Photography (Nikon Z6 and my powerful iPhone), Hiking (Mount Siguniang, 4,668m), Taekwondo (8th Kup)
  • Language Enthusiast, Sports Enthusiast (Skiing, Swimming, Badminton, Table Tennis)