Scholar
Xiang An
Google Scholar ID: 1ckaPgwAAAAJ
DeepGlint
Computer Vision
Follow
Homepage
↗
Google Scholar
↗
Citations & Impact
All-time
Citations
891
H-index
13
i10-index
13
Publications
20
Co-authors
14
list available
Contact
Email
xiangan@deepglint.com
GitHub
Open ↗
Publications
19 items
HaPRL: Human-Anchored Process Reinforcement Learning for Visual Search Agent
2026
Cited
0
EviViT: Evidence-Adaptive Vision Transformers for Fine-Grained Perception
2026
Cited
0
Spatial-OPSD: Self-Improving Spatial Reasoning via Label-Free Self-Distillation
2026
Cited
0
StreamOPD: A Post-Training Recipe with Spatio-Temporal Cue Gating for Streaming Video Understanding
2026
Cited
0
Mage-VL: An Efficient Codec-Native Streaming Multimodal Foundation Model
2026
Cited
0
LLaVA-OneVision-2: Towards Next-Generation Perceptual Intelligence
2026
Cited
0
4DThinker: Thinking with 4D Imagery for Dynamic Spatial Understanding
2026
Cited
0
OneVision-Encoder: Codec-Aligned Sparsity as a Foundational Principle for Multimodal Intelligence
2026
Cited
0
Load more
Resume
Research Experience
Contributed to the development of LLaVA-OneVision-1.5, a fully open framework for democratized multimodal training
Enhanced OCR and localization capabilities of Large Pretrained Vision Transformers (ViT) for ICCV2025 with the RICE method
Trained large-scale vision models using weak supervision on LAION400M and COYO700M for ECCV2024 with the MLCD approach
Proposed UNICOM: Universal and Compact Representation Learning for Image Retrieval, presented at ICLR2023
Developed Partial FC: an efficient and robust distributed hybrid parallel algorithm for large-scale face recognition trained on WebFace260M
Co-authors
13 total
Jiankang Deng
Imperial College London
Kaicheng Yang
DeepGlint
Jia Guo
insightface.ai
Yongle Zhao
DeepGlint
Jing Yang
University of Cambridge
Xuhan Zhu
UCAS
Ying Fu
Beijing Institute of Technology
Tiancheng Gu
The University of Sydney