Scholar
Hisham Cholakkal
Google Scholar ID: bZ3YBRcAAAAJ
Mohamed bin Zayed University of Artificial Intelligence (MBZUAI)
Computer Vision
Large Multimodal Models
LLM
Healthcare Foundation Model
Conversational Assistant
Follow
Homepage
↗
Google Scholar
↗
Citations & Impact
All-time
Citations
4,714
H-index
32
i10-index
57
Publications
20
Co-authors
22
list available
Publications
46 items
WorldGuide: Goal-Directed Video World Model for Procedural Task Execution
2026
Cited
0
Omni-Embed-Mini: Binding Modalities Without Forgetting via Dense Distillation
2026
Cited
0
Hard Vision, Easy Vision: What GPT-6 Astra Reveals Across Computer Vision
2026
Cited
0
Small yet Assistive: Spatially-Aware Post-Training for Low Vision
2026
Cited
0
Ground3D-LMM: Fine-Grained 3D Point Grounding and Spatial Reasoning with LMM
2026
Cited
0
Paying More Attention to Visual Tokens in Self-Evolving Large Multimodal Models
2026
Cited
0
Ask, Solve, Generate: Self-Evolving Unified Multimodal Understanding and Generation via Self-Consistency Rewards
2026
Cited
0
SafeDiffusion-R1: Online Reward Steering for Safe Diffusion Post-Training
2026
Cited
0
Load more
Co-authors
15 total
Rao Muhammad Anwer
Mohamed bin Zayed University of Artificial Intelligence (MBZUAI)
Fahad Shahbaz Khan
MBZUAI, Linköping University Sweden
Salman Khan
MBZUAI, Australian National University
Ling Shao, Fellow of IEEE/IAPR
Terminus; Founder of IIAI/MBZUAI
Yanwei Pang
Tianjin University
Sanath Narayan
Technology Innovation Institute, Abu Dhabi
Mubarak Shah
Trustee Chair Professor of Computer Science, University of Central Florida
Ming-Hsuan Yang
University of California at Merced; Google DeepMind