Scholar
Sang-gil Lee
Google Scholar ID: P93s2UQAAAAJ
NVIDIA
Deep Generative Model
Audio Synthesis
Language Model
Follow
Homepage
↗
Google Scholar
↗
Citations & Impact
All-time
Citations
1,043
H-index
11
i10-index
12
Publications
20
Co-authors
14
list available
Publications
10 items
Audio-Visual Flamingo: Open Audio-Visual Intelligence for Long and Complex Videos
2026
Cited
0
Unified Audio Intelligence Without Regressing on Text Intelligence
2026
Cited
0
Benchmarking Single-Factor Physical Video-to-Audio Generation
2026
Cited
0
Audio Flamingo Next: Next-Generation Open Audio-Language Models for Speech, Sound, and Music
2026
Cited
0
Music Flamingo: Scaling Music Understanding in Audio Language Models
2025
Cited
0
UALM: Unified Audio Language Model for Understanding, Generation and Reasoning
2025
Cited
0
Audio Flamingo 3: Advancing Audio Intelligence with Fully Open Large Audio Language Models
2025
Cited
0
UniWav: Towards Unified Pre-training for Speech Representation Learning and Generation
2025
Cited
0
Load more
Co-authors
12 total
Sungroh Yoon
Professor, Electrical and Computer Engineering & Artificial Intelligence, Seoul National University
Bryan Catanzaro
NVIDIA
Heeseung Kim
Seoul National University
Wei Ping
Distinguished Research Scientist, NVIDIA
Rafael Valle
NVIDIA, UC Berkeley, CNMAT
Chaehun Shin
Seoul National University
Zhifeng Kong
Senior Research Scientist, NVIDIA
Boris Ginsburg
NVIDIA