Scholar
Wenhui Tan
Google Scholar ID: cELItK0AAAAJ
Renmin University of China
Multimodal
LLM Reasoning
Embodied AI
Follow
Homepage
↗
Google Scholar
↗
Citations & Impact
All-time
Citations
117
H-index
5
i10-index
3
Publications
8
Co-authors
11
list available
Publications
15 items
Can Computation from Earlier Problems Help LLMs Solve New Ones?
2026
Cited
0
Informed Masking: Structure-Aware Perturbation for Reinforcement Learning in Diffusion Large Language Models
2026
Cited
0
Screenshots or Tools? Eliciting Tool Use and Managing Multimodal Context in Hybrid GUI-MCP Computer-Use Agents
2026
Cited
0
AVOC: Enhancing Hour-Level Audio-Video Understanding in Omni-Modal LLMs via Retrieval-Inspired Token Compression
2026
Cited
0
Pair-In, Pair-Out: Latent Multi-Token Prediction for Efficient LLMs
2026
Cited
0
MSJoE: Jointly Evolving MLLM and Sampler for Efficient Long-Form Video Understanding
2026
Cited
0
BFS-PO: Best-First Search for Large Reasoning Models
2026
Cited
0
Video-OPD: Efficient Post-Training of Multimodal Large Language Models for Temporal Video Grounding via On-Policy Distillation
2026
Cited
0
Load more
Co-authors
11 total
Ruihua Song
Renmin University of China
Chuhao jin
PhD student of AI, Renmin University of China
Bei Liu
Microsoft Research
Jianlong Fu
Microsoft Research
Jiange Yang
Nanjing University
Peng Cao
Northeastern University
Jian Luan
Toshiba, Microsoft, Xiaomi
Ju Jianzhong
Xiaomi