Scholar
Yuhan Zhu
Google Scholar ID: ydgR3LgAAAAJ
Nanjing University, Shanghai AI Lab
Computer Vision
Vision-Language Models
Video Understanding
Follow
Homepage
↗
Google Scholar
↗
Citations & Impact
All-time
Citations
328
H-index
7
i10-index
6
Publications
9
Co-authors
7
list available
Publications
20 items
OneStreamer: Unifying Perception, Memory, and Proactive Response in Streaming Video Interaction
2026
Cited
0
Benchmarking and Enhancing Skill-Level Memory for Partially Observable Robotic Manipulation
2026
Cited
0
WorldTS: World Modeling for Multimodal Covariate-aware Time Series Forecasting
2026
Cited
0
WPBench: A Comprehensive Benchmark for Wind Power Forecasting
2026
Cited
0
MUSE: Dependency-Aware Adaptation of a Frozen Vision Backbone for Multivariate Time Series Forecasting
2026
Cited
0
TimeLens2: Generalist Video Temporal Grounding with Multimodal LLMs
2026
Cited
0
VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding
2026
Cited
0
VBGS-SLAM: Variational Bayesian Gaussian Splatting Simultaneous Localization and Mapping
2026
Cited
0
Load more
Co-authors
6 total
Limin Wang
Nanjing University
Guozhen Zhang
Nanjing University
Haocheng Shen
AI Lab, vivo
Xinhao Li
Nanjing University
Jing Tan
The Chinese University of Hong Kong
Zhiyu Zhao
ByteDance, Nanjing University