Linqing Wang 王林青
Tencent Hunyuan · Multimodal Understanding & AIGC Generation
Hello! I work on large-scale multimodal foundation models at Tencent Hunyuan, focusing on omni-unified multimodal interaction, text-to-image and video generation, unified understanding–generation models, and post-training for diffusion systems.
I was previously involved in robotics and bionics research during graduate study at Beihang University (BUAA). Feel free to reach out if you would like to chat or discuss research ideas.
news
- Apr 2026 Released HY-SOAR, a reward-free post-training method for diffusion models beyond SFT and RL — no reward models, preference labels, or negative samples.
- Apr 2026 Released RvR, enlarging modification space boosts image refinement in unified multimodal models.
- Apr 2026 Released HiVG, a 3B model beating GPT-5 and Gemini 2.5 on image-to-SVG via hierarchical SVG tokenization.
- Jan 2026 Released TAGRPO for image-to-video GRPO with direct trajectory alignment. ICML 2026
- Nov 2025 Released JarvisEvo, a self-evolving photo editing agent with synergistic editor–evaluator optimization. CVPR 2026
- Oct 2025 Released HunyuanImage 3.0, the first open-source unified understanding and generation image model, ranking #1 on LMArena.
- Sep 2025 Released PromptEnhancer, powering Hunyuan Image 2.1 to rank #1 among open-source models on Artificial Analysis Arena. CVPR 2026
- Dec 2024 Released HunyuanVideo; core contributor on portrait image-to-video.
- Jul 2024 Released Kolors; pioneered noise-schedule optimization for direct 1K image output.
selected works
-
CVPR Workshops 2026
ChatUMM: Robust Context Tracking for Conversational Interleaved Generation
IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops, 2026
-
ICML 2026
TAGRPO: Boosting GRPO on Image-to-Video Generation with Direct Trajectory Alignment
International Conference on Machine Learning, 2026
-
Open Source 2024
Kolors: Effective Training of Diffusion Model for Photorealistic Text-to-Image Synthesis
Bilingual photorealistic T2I; noise-schedule optimization for direct 1K output.
Full list and citation counts on Google Scholar.