Xiaogang Xu

aka Theo
>

Ensure that artificial general intelligence benefits all of humanity.

— OpenAI Mission · the north star I work toward
~$
“A great man can bend in adversity and rise when his moment comes.”
Xiaogang Xu
Research Journey
Research journey timeline

Biography

I obtained my Ph.D. degree in the Department of Computer Science and Engineering at the Chinese University of Hong Kong, supervised by Prof. Jiaya Jia and Prof. Bei Yu. Before that, I obtained my B.E. degree in Information Engineering at the College of Information Science and Electronic Engineering, Zhejiang University.

During my Ph.D. life, I have spent wonderful times collaborating with, among others:

In the industrial community, I conduct research and development on AGI.

Zhejiang Lab
I primarily led the construction of a 10,000-card GPU cluster and spearheaded the training of domain-specific large science models, in close collaboration with Academician Jian Wang, Prof. Hujun Bao, and Prof. Zhe Liu.
Huawei
My research and development focused on training unified inference models for generation and understanding, as well as advancing AIGC models including text-to-image and text-to-video. I successfully drove the deployment of these models into core Huawei products, and established strong academic collaborations with Prof. Hanwang Zhang from NTU Singapore.
MiroMind
I dedicated myself to the post-training of agentic foundation models and the development of self-evolving harnesses, working in close partnership with Prof. Shuicheng Yan from NUS Singapore.
Several specific topics of our current research interests and focus:
01Multi-modality data (image, video, 3D, etc.) generation & manipulation via AIGC
02Multi-modality large model and agent → AGI
03Generative Computational Photography: large model and efficiency optimization
04Security and alignment for large models / AGI
05Embodied AI (VLA, VA, etc.) and World Model
I carry out development and research in the industrial and academic communities at the same time. If you are interested in collaborating with me, please feel free to contact me through the email.
Selected Collaborations
Zhejiang University
University of Oxford
CUHK
Adobe
MIT
Huawei
HKUST
Alibaba
Max Planck Institute for Informatics
Meta
Peking University
Microsoft
National University of Singapore
ByteDance
Tsinghua University
Tencent
The University of Hong Kong
Snapchat
Nanyang Technological University
SmartMore
Sun Yat-sen University
SenseTime
Zhejiang Lab
MiroMind
Collaborator

News

Talks & Presentations

Talks & Presentations in 2024

  • Give a talk in Southeast University with the topic of "AIGC-based Image Restoration".
    Oct. 2024.
  • Give a talk in Zhejiang Lab with the topic of "Improve Model Robustness under Extreme Dark Environments".
    Sep. 2024.
  • Co-host a workshop in ChinaSys with the topic of "AI System Building for Large Models".
    June. 2024.
  • Give a talk at Nanjing University of Aeronautics and Astronautics about "Transferrable Adversarial Attacks" .
    June. 2024.
  • Oral presentation about future media technology at Huawei STW conference (Shenzhen).
    May. 2024.
  • Invited poster presentation at VALSE2024 on "Boosting Image Restoration via Priors from Pre-trained Models".
    May. 2024.
  • Invited talk at China3DV with the topic of "Efficient 3D Modeling for Data with Real-world Degradations".
    Apr. 2024.
  • Give a talk at Nankai University with the topic of "AIGC for Computational Photography in the RAW Domain" .
    Apr. 2024.
  • Invited talk to Responsible AI team at ByteDance, on "Responsible LLM and AIGC".
    Apr. 2024.
  • Selected into "Young Talent Nurturing Project at Zhejiang Lab (之江青年人才托举)" for Large Models (大模型).
    Mar. 2024.
  • Organizer at GAMES Webinar on "Multi-view Synthesis and 3D Shape Completion via Diffusion Models", [News].
    Mar. 2024.
  • Presentation at [Shining 3D] with topic of "High-quality 3D Reconstruction and Generation".
    Jan. 2024.

Talks & Presentations in 2023

  • Give a talk at Alibaba International Digital Commerce (AIDC), "Intelligent Generation and Restoration".
    Dec. 2023.
  • Presentation for Huawei Central Media Research Institute, "Multi-Modality Low-Light Data Enhancement".
    Oct. 2023.
  • Invited talk at VALSE Webinar on "LLIE via Structure Modeling and Guidance", [News].
    Sep. 2023.
  • Organize a nationwide academic meeting at Hangzhou with topic of "Intelligent Computing and Security".
    Sep. 2023.
  • Give a talk at Zhejiang University, "Reliable Artificial Intelligence for Image Generation and Manipulation".
    July. 2023.
  • Presentation for CVLab at ETH, "Multi-Modality Restoration".
    July. 2023.
  • Invited to give a talk at HKUST, "Effective Generative Models for Real-World Manipulation and Restoration".
    July. 2023.
  • Give a talk to Alibaba DAMO Academy, with topic of "Real-world Generation for 2D and 3D Data".
    May. 2023.
  • Awarded with "Science Fund Program for Excellent Young Scientists at Zhejiang Lab (之江优秀青年科学基金)".
    Mar. 2023.
  • AI TIME Personal Talk: "Deep Parametric 3D Filters for Multiple Degradations Restoration".
    Mar. 2023.
  • AI TIME ECCV 2022: "Multi‑Task Learning via Transformer and Cross‑Task Reasoning".
    Dec. 2022.

Research Summary

Multi-modality Generation
Multi-modality Restoration
Multi-modality Understanding

Technical Report

Full Reports in 2026

Full Reports in 2025

Full Reports in 2024

Full Reports in 2023

Full Reports in 2022

Publications [Google Scholar]

Publications in 2026

Publications in 2025

Publications in 2024

Publications in 2023

Publications in 2022

Publications in 2021

Publications in 2020

Publications in 2019

Publications in 2018

Publications in 2017

Experiences

Work Experiences

  • MiroMind, Singapore
    Sep. 2025 – May 2026

    AI Scientist
    Topic: Agent Models and Harness,
    Future Prediction, Agentic RL, Deep Research, etc.
  • Huawei 2012 Lab, Hangzhou, China
    Mar. 2024 – Sep. 2025

    Top Minds, Staff AIGC Scientist
    Topic: Lead AIGC, MLLM, World Model Teams,
    Mobile application products with Computational photography
  • Zhejiang Lab, Hangzhou, China
    Jan. 2023 – Mar. 2024

    Research Scientist, AIGC leader
    Topic: Computer Vision Models,
    Diffusion, Autoregressive model, and Alignment

Postdoctoral and Visiting Experiences

Intern Experiences

Professional Activities

Conference Reviewer

  • IEEE Conference on Computer Vision and Pattern Recognition (CVPR'18-26, CCF-A)
  • IEEE International Conference on Computer Vision (ICCV'19-25, CCF-A)
  • European Conference on Computer Vision (ECCV'20-26, CCF-B)
  • SIGGRAPH and SIGGRAPH Asia (23-25, CCF-A)
  • Neural Information Processing Systems (NeurIPS'19-25, CCF-A)
  • International Conference on Learning Representations (ICLR'20-26)
  • International Conference on Machine Learning (ICML'22-25, CCF-A)
  • AAAI Conference on Artificial Intelligence (AAAI'20-26, CCF-A), Program Committee
  • IEEE Winter Conference on Applications of Computer Vision (WACV'21-24)
  • Asian Conference on Computer Vision (ACCV'22, CCF-C)
  • European Conference on Artificial Intelligence (ECAI'24-25, CCF-B)
  • Chinese Conference on Pattern Recognition and Computer Vision (PRCV'24, CCF-C)

Journal Reviewer

  • IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI, CCF-A)
  • IEEE Transactions on Circuits and Systems for Video Technology (TCSVT, CCF-B)
  • IEEE Transactions on Visualization and Computer Graphics (TVCG, CCF-A)
  • IEEE Transactions on Multimedia (TMM, CCF-B)
  • IEEE Transactions on Instrumentation and Measurement (TIM, JCR-Q1)
  • IEEE Transactions on Neural Networks and Learning Systems (TNNLS, CCF-B)
  • IEEE Signal Processing Letters (SPL, CCF-C)
  • Computer Vision and Image Understanding (CVIU, CCF-B)
  • Neural Processing Letters (CCF-C)
  • International Journal of Computer Vision (IJCV, CCF-A)
  • International Journal of Human-Computer Interaction (IJHC, CCF-B)
  • Neurocomputing (CCF-C)
  • Pattern Recognition (PR, CCF-B)
  • Knowledge-Based Systems (KBS, CCF-C)
  • Neural Networks (CCF-B)
  • IET Image Processing (CCF-C)
  • Journal of Computer-Aided Design & Computer Graphics (计算机辅助设计与图形学学报, CCF中文A类)
  • ACM Transactions on Multimedia Computing Communications and Applications (TOMM, CCF-B)

Area Chair

  • International Conference on Machine Learning (ICML'26, CCF-A)
  • Association for Computational Linguistics (ACL'26, CCF-A)
  • Conference and Workshop on Neural Information Processing Systems (NeurIPS'26, CCF-A)
  • AAAI Conference on Artificial Intelligence (AAAI'27, CCF-A), Senior Program Committee

Honors & Awards

Honors & Awards in 2024

  • Best Demo Honorable Mention in CVPR 2024, for Depth Anything
    2024
  • Zhejiang Lab Elite Scientist Sponsorship Program, 30 from 3000+ researchers in Zhejiang Lab
    2024
  • 2024
  • Large Model Safety Risk Guardrail Theory and Key Technologies (浙江省自然科学基金重大项目, 1,000,000 CNY)
    2024

Honors & Awards in 2023

  • Cadre member in Zhejiang KunPeng Project (鲲鹏计划, the highest honored research funding in Zhejiang Province)
    2023
  • Science Fund Program for Excellent Young Scientists at Zhejiang Lab (之江优秀青年科学基金, 1,000,000 CNY)
    2023

Honors & Awards in 2018-2022

Honors & Awards in 2015-2017

  • National Scholarship, Ministry of Education of P.R. China
    2015
  • Title of Outstanding Students, ZJU
    2015/16/17
  • The scholarship for excellence in research and innovation, ZJU
    2016/17
  • Zhejiang Provincial Government Scholarship
    2016
  • China Undergraduate Mathematical Contest in Modeling, National second prize
    2016
  • Mathematical Contest in Modeling (Honorable Mention), COMAP (U.S.A)
    2016

Patents

Teaching

Life Beyond Research 生活

Research is what I do; these are what keep me whole. 研究之外,我更向往生活本身。

Running 跑步

Long distances teach patience better than any paper deadline — one steady step at a time.
长跑教会我的耐心,比任何 deadline 都多。

HALF MARATHON ×5+ FULL MARATHON ×2
Hiking 爬山

Every summit is a reminder that the view is worth the climb.
山顶的风景,总值得那段上坡路。

Reading 读书

Borrowing other lives and other centuries, a few pages at a time.
读书,是借别人的一生,换自己的开阔。

Seeking the Way 悟道

Sitting with the questions that have no benchmark and no leaderboard.
有些问题没有 benchmark,只能慢慢想、慢慢悟。

「路虽远,行则将至;事虽难,做则必成。」

© Xiaogang Xu | Last updated: 01/08/2026