Skip to content
Xiang Wen

Xiang Wen

Tencent Hunyuan · Foundation Model Department

  • Interactive virtual worlds
  • Video generation
  • Diffusion models
  • Computer graphics
  • 3D vision
  • TPAMI / ICMLJournal and conference
  • SIGGRAPH Asia / ACM MMGraphics and multimedia
  • GDC / GTCInvited talks
  • SkyReels-V12.7K stars

About

I am working toward real-time interactive virtual worlds: environments generated as you move through them, that change because of what you did, rather than finished before you arrive.

That problem sits across generative AI, diffusion models, computer graphics, and 3D vision. I have shipped work in all four, and the hard parts tend to fall in the seams between them. Right now that means post-training video generation models at Tencent Hunyuan: generation quality, a unified architecture handling text, image, multi-reference, and editing in one model, and the inference efficiency that real-time interaction ultimately depends on.

Before Tencent I built SkyReels AI, an AI short-drama platform, from scratch as its business lead and CTO, led game AI at ByteDance, and worked on game agents and content generation at NetEase Fuxi AI Lab. Over the same years I finished a PhD in artificial intelligence at Zhejiang University.

Read more →

News

  1. 2026.03Received the Business Application Pioneer Award at Tencent for leading the algorithm work on AniMatrix.Award
  2. 2026.03HY-WU open-sourced at Tencent Hunyuan, a framework that generates instance-conditioned LoRA adapters on the fly for text-guided image editing.Open sourcearXiv
  3. 2025.12Completed my PhD in artificial intelligence at Zhejiang University.
  4. 2025.08VolGen accepted to TPAMI.TPAMI
  5. 2025.07Joined the Tencent Hunyuan foundation model department, working on post-training for video generation.
  1. 2025.02Open-sourced SkyReels-V1, a human-centric video foundation model, releasing both text-to-video and image-to-video weights along with the inference framework.Open source
  2. 2025.02Released SkyReels-A1, expressive portrait animation on video diffusion transformers, with code.arXivOpen source
  3. 2024.12Released Video Diffusion Transformers are In-Context Learners.arXiv
  4. 2024.09Talk at the Apsara Conference on AI workflows for producing drama online, from script generation to storyboarding.Apsara
  5. 2024.08Released SkyScript-100M, a script and shooting-script dataset for short drama.arXiv
  6. 2023.07Joined Kunlun Tech to build SkyReels AI, an AI short-drama platform, from scratch.
  7. 2023.03Talk at the GDC Machine Learning Summit on GPT-3 powered text to lifelike speech and animation for NPCs.GDC
  8. 2020.09Talk at the Beijing International Game Innovation Conference on AI-assisted game content production.BIGC
  9. 2019.12Talk at NVIDIA GTC China on video-guided intelligent choreography, shipped in Justice Online.GTC
Show all (14)Show less

Work

Selected papers

All papers →
Show 11 moreShow less

Now

Working on post-training for anime video generation at Tencent Hunyuan, focused on unified generation architectures and inference acceleration. On the side, building a pipeline that turns papers into narrated videos.

Read more →