Scholar
Yimu Wang
Google Scholar ID: TV2vnN8AAAAJ
University of Waterloo
Multi-modal Learning
Follow
Google Scholar
↗
Citations & Impact
All-time
Citations
363
H-index
11
i10-index
11
Publications
20
Co-authors
0
Contact
No contact links provided.
Publications
11 items
LitTraceQA: A Benchmark for Multi-Stage Grounding and Verification in Scientific Question Answering
2026
Cited
0
MASS: Multiplayer World Models with Authoritative Shared State
2026
Cited
0
Where Does the Answer Come From? Benchmarking View-Level Visual Evidence Identification in Multi-View MLLMs for Autonomous Driving
2026
Cited
0
VISTAQA: Benchmarking Joint Visual Question Answering and Pixel-Level Evidence
2026
Cited
0
UNIFORM: Unifying Knowledge from Large-scale and Diverse Pre-trained Models
2025
Cited
0
HAWAII: Hierarchical Visual Knowledge Transfer for Efficient Vision-Language Models
2025
Cited
0
Survey of Video Diffusion Models: Foundations, Implementations, and Applications
2025
Cited
0
LEO-MINI: An Efficient Multimodal Large Language Model using Conditional Token Reduction and Mixture of Multi-Modal Experts
2025
Cited
0
Load more
Resume (English only)
Co-authors
0 total
Co-authors: 0 (list not available)