Scholar
Haizhou Li
Google Scholar ID: z8_x7C8AAAAJ
The Chinese University of Hong Kong, Shenzhen (CUHK-Shenzhen), China; NUS, Singapore
Automatic Speech Recognition
Speaker Recognition
Language Recognition
Voice Conversion
Machine Translation
Follow
Homepage
↗
Google Scholar
↗
Citations & Impact
All-time
Citations
22,881
H-index
71
i10-index
402
Publications
20
Co-authors
72
list available
Publications
103 items
SpikingVLA: Asynchronous Spiking Vision-Language-Action Models
2026
Cited
0
HuatuoGPT-3: RL-Only Domain Adaptation from Base Models
2026
Cited
0
Multimodal Target Speaker Extraction: Towards Unified Speaker Cues Across Modalities
2026
Cited
0
Dialogue-Based Streaming Audio-Visual Target Speaker Extraction with Predictive Dialogue Information
2026
Cited
0
Exploring a Single Autoregressive LLM for Unified Target Speech Extraction across Synchronous and Asynchronous Cues
2026
Cited
0
Spoken Language Models that Think Aloud
2026
Cited
0
GenTraceBench: A Benchmark for Tracing Audio Deepfakes Across Pre- and Post-training Stages
2026
Cited
0
PIVOT: Physics-Grounded Verification for AI-Generated Audio-Video Detection
2026
Cited
0
Load more
Co-authors
16 total
Eng-Siong Chng
Nanyang Technological University
Kong Aik Lee
The Hong Kong Polytechnic University, Hong Kong
Xiong Xiao
Principal Applied scientist, Microsoft
Berrak Sisman
Assistant Professor (ECE & DSAI), Johns Hopkins University
Min Zhang
Professor of Computer Science, Soochow University
Minghui Dong
Institute for Infocomm Research, A*Star, Singapore
Tomi Kinnunen
Professor, University of Eastern Finland
Zhizheng Wu
The Chinese University of Hong Kong, Shenzhen (CUHK-Shenzhen), Mel Lab