Scholar
Haozhe Liu
Google Scholar ID: QX51P54AAAAJ
KAUST
Computer Vision
Reinforcement Learning
Multimodal
Image/Video Generation
Follow
Homepage
↗
Google Scholar
↗
Citations & Impact
All-time
Citations
949
H-index
18
i10-index
21
Publications
20
Co-authors
29
list available
Contact
Email
haozhe.liu@kaust.edu.sa
CV
Open ↗
GitHub
Open ↗
Publications
22 items
SoL-Refiner: Speed-of-Light One-Step Refinement for High-Resolution Video
2026
Cited
0
Sol-H3: Recursive Self-Improvement for MiniMax-H3 Inference Acceleration on Sol-Engine across Cloud and Edge
2026
Cited
0
SoL-Pi: Recursively Scaling Auto-Research Loops for Efficient Agent Harness
2026
Cited
0
P2Voxel: Pyramid Pivot Voxelization for 3D Mesh Tokenization
2026
Cited
0
TrajWiki: Source-Grounded Memory Trajectories for Long-Horizon Dialogue Agents
2026
Cited
0
Sol-Attn: Accelerating Video Generation Inference via On-the-Fly Attention Sparsification
2026
Cited
0
SANA-Video 2.0: Hybrid Linear Attention with Attention Residuals for Efficient Video Generation
2026
Cited
0
Sol Video Inference Engine: Agent-Native Full-Stack Acceleration Framework for Efficient Video Generation
2026
Cited
0
Load more
Resume
Academic Achievements
- Published over 10 papers in top-tier venues, with more than 700 citations on Google Scholar
- Notable projects: BoxDiff, NLSOM, TGATE, MarDini
- Co-developed the pretraining of a 7B commercial text-to-image model at Meta, scaling to O(1B) training samples
- Papers accepted by TMLR, CVPR, ICCV, ECCV, MICCAI, AAAI
- NLSOM recognized as the best paper at NeurIPS 2023 workshop
Research Experience
- Research Scientist Intern at Meta AI (London), focusing on efficient video generation (Summer 2024)
- Intern at Tencent (Shenzhen), focusing on GAN-based image synthesis and medical imaging
- Research Scientist Intern at Meta (MPK), focusing on GenAI-related topics (Summer 2025)
Education
- Degree: Ph.D. Candidate
- University: KAUST
- Advisor: Prof. Juergen Schmidhuber
- Time: August 2022 - Present
- Field: Not specified
Background
- Research Interests: Multimodal generative models, particularly video and image generation
- Long-term goal: Developing a Physical AI Model (world model)
- Brief: Currently pursuing a Ph.D. at KAUST, supervised by Prof. Juergen Schmidhuber. Previously interned at Meta AI and Tencent.
Miscellany
- Open to future collaborations, whether it's co-founding a venture or pursuing full-time opportunities in industry or academia
- Personal interests not specified
Co-authors
17 total
Yefeng Zheng
Professor, Westlake University, Hangzhou, China, IEEE Fellow, AIMBE Fellow
Jinheng Xie
National University of Singapore
Linlin Shen
Shenzhen University
Bernard Ghanem
Professor, King Abdullah University of Science and Technology
Mike Z. SHOU
National U. of Singapore; Facebook AI; Columbia University
Haoqian Wu
Tencent Youtu Lab, Shenzhen University
Juergen Schmidhuber
King Abdullah University of Science and Technology / The Swiss AI Lab, IDSIA / University of Lugano
Raghavendra Ramachandra
Professor, Norwegian University of Science and Technology (NTNU), Norway