Scholar
Belinda Zeng
Google Scholar ID: B8g_FcUAAAAJ
Amazon
NLP
computer vision
multi-modal
large language model
Follow
Google Scholar
↗
Citations & Impact
All-time
Citations
927
H-index
12
i10-index
14
Publications
20
Co-authors
0
Publications
6 items
Native Action-Prior Learning from Videos for World Action Models
2026
Cited
0
Video, Ergo Genero: Unifying Video Tasks via Spatiotemporal Analogy
2026
Cited
0
TUA-Bench: A Benchmark for General-Purpose Terminal-Use Agents
2026
Cited
0
WavFlow: Audio Generation in Waveform Space
2026
Cited
0
Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation
2026
Cited
0
Learning LLM Preference over Intra-Dialogue Pairs: A Framework for Utterance-level Understandings
2025
Cited
0