Scholar
Zhifeng Kong
Google Scholar ID: jAOD1dsAAAAJ
Senior Research Scientist, NVIDIA
Deep Generative Models
Diffusion Models
Audio Foundation Models
Audio LM
Trustworthy ML
Follow
Homepage
↗
Google Scholar
↗
Citations & Impact
All-time
Citations
3,053
H-index
15
i10-index
16
Publications
20
Co-authors
22
list available
Publications
15 items
RMS-AQA: A Two-Stage Spatial Audio Question Answering Benchmark for Real-World Domestic Environments
2026
Cited
0
Audio-Visual Flamingo: Open Audio-Visual Intelligence for Long and Complex Videos
2026
Cited
0
Unified Audio Intelligence Without Regressing on Text Intelligence
2026
Cited
0
Benchmarking Single-Factor Physical Video-to-Audio Generation
2026
Cited
0
Audio Flamingo Next: Next-Generation Open Audio-Language Models for Speech, Sound, and Music
2026
Cited
0
Music Flamingo: Scaling Music Understanding in Audio Language Models
2025
Cited
0
UALM: Unified Audio Language Model for Understanding, Generation and Reasoning
2025
Cited
0
Audio Flamingo Sound-CoT Technical Report: Improving Chain-of-Thought Reasoning in Sound Understanding
2025
Cited
0
Load more
Co-authors
15 total
Bryan Catanzaro
NVIDIA
Wei Ping
Distinguished Research Scientist, NVIDIA
Rafael Valle
NVIDIA, UC Berkeley, CNMAT
Kamalika Chaudhuri
FAIR @ Meta
Arushi Goel
Research Scientist, NVIDIA
Dahua Lin
The Chinese University of Hong Kong
Zhaoyang Lyu
PhD of Information Engineering, The Chinese University of Hong Kong
Sang-gil Lee
NVIDIA