I am an incoming Ph.D. student in Computer Science at Peking University, supervised by Prof. Hao Tang. I also collaborate closely with Prof. Li Yi and Prof. Hao Zhao. I am currently a research intern at Kling Group, Kuaishou Technology, working on image generation foundation models. My current research interests include 3D computer vision, diffusion models, and embodied AI.

Previously, I received my M.S. degree from the Institute of Computing Technology, Chinese Academy of Sciences (ICT, CAS), where I was supervised by Prof. Xing Hu. I received my B.S. degree in Computer Science from the University of Chinese Academy of Sciences (UCAS).

CV / Google Scholar / GitHub / Email

News.

  • 2026.08: AdvFD is released on arXiv.
  • 2026.06: One paper accepted to ECCV 2026.
  • 2026.03: One paper accepted to CVPR 2026.
  • 2025.03: One paper accepted to CVPR 2025.
  • 2024.07: Two papers accepted to ECCV 2024, and one paper accepted as a MICCAI 2024 Spotlight.

Publications

arXiv 2026
AdvFD

AdvFD: Boosting Visual Generation via Adversarial Fréchet Distance Loss

Mingju Gao*, Jingkai Zhou*, Kun Gai, Changqian Yu, Hao Tang.

arXiv 2026.

ECCV 2026
PhysRAG

PhysRAG: Enhancing Physics-Awareness in Video Generation via Retrieval-Augmented Generation

Kexu Cheng*, Zicheng Liu*, Mingju Gao*, Chunhe Song, Hao Tang.

ECCV 2026. (Equal contribution)

CVPR 2026
PAM

PAM: A Pose-Appearance-Motion Engine for Sim-to-Real HOI Video Generation

Mingju Gao*, Kaisen Yang*, Huan-ang Gao, Bohan Li, Ao Ding, Wenyi Li, Yangcheng Yu, Jinkun Liu, Shaocong Xu, Yike Niu, Haohan Chi, Hao Chen, Hao Tang, Yu Zhang, Li Yi, Hao Zhao

CVPR 2026.

CVPR 2025

PartRM: Modeling Part-Level Dynamics with Large Cross-State Reconstruction Model

Mingju Gao*, Yike Pan*, Huan-ang Gao*, Zongzheng Zhang, Wenyi Li, Hao Dong, Hao Tang, Li Yi, Hao Zhao.

CVPR 2025.

ECCV 2024
SCP-Diff

SCP-Diff: Spatial-Categorical Joint Prior for Diffusion Based Semantic Image Synthesis

Huan-ang Gao*, Mingju Gao*, Jiaju Li, Wenyi Li, Rong Zhi, Hao Tang, Hao Zhao.

ECCV 2024. (Equal contribution)

ECCV 2024
Training-free model merging

Training-free model merging for multi-target domain adaptation

Wenyi Li*, Huan-ang Gao*, Mingju Gao, Beiwen Tian, Rong Zhi, Hao Zhao.

ECCV 2024.

MICCAI 2024 Spotlight
FairDiff

FairDiff: Fair Segmentation with Point-Image Diffusion

Wenyi Li*, Haoran Xu*, Guiyu Zhang*, Huan-ang Gao, Mingju Gao, Mengyu Wang, Hao Zhao.

MICCAI 2024 Spotlight.

Education

Experience

Professional Service

  • Conference Reviewer: NIPS 2026, CVPR 2026, CVPR 2025, ICCV 2025.