Scholar
Haoning Wu
Google Scholar ID: ia4M9mMAAAAJ
Shanghai Jiao Tong University
Computer Vision
Multi-modal Learning
Generative Models
Follow
Homepage
↗
Google Scholar
↗
Citations & Impact
All-time
Citations
194
H-index
6
i10-index
5
Publications
11
Co-authors
0
Contact
No contact links provided.
Publications
35 items
ST-WAM: Semantic-Temporal World Action Model for Robust Manipulation under Visual Distribution Shifts
2026
Cited
0
PerceptionBench: Evaluating Atomic Visual Perception in Multimodal Large Language Models
2026
Cited
0
BasketEvent: Understanding Who Did What and When in Basketball Videos
2026
Cited
0
VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding
2026
Cited
0
ABot-M0.5: Unified Mobility-and-Manipulation World Action Model
2026
Cited
0
PhysEditWorld: A Large-Scale Dataset Toward Physics-Editable World Models
2026
Cited
0
Count Anything at Any Granularity
2026
Cited
0
Improving Human Image Animation via Semantic Representation Alignment
2026
Cited
0
Load more
Resume (English only)
Co-authors
0 total
Co-authors: 0 (list not available)