Scholar
Boyi Kang
Google Scholar ID: cvZ0uNwAAAAJ
The Hong Kong University of Science and Technology
Multimodal Intelligence
Audio Processing
Follow
Google Scholar
↗
Citations & Impact
All-time
Citations
12
H-index
1
i10-index
1
Publications
3
Co-authors
0
Publications
5 items
COT-TTS: Audio Context-Aware Text-to-Speech with Chain-of-Thought Reasoning
2026
Cited
0
ISCSLP 2026 CoT-TTS Challenge: Chain-of-Thought Reasoning for Context-Aware Text-to-Speech
2026
Cited
0
NVBench: A Benchmark for Speech Synthesis with Non-Verbal Vocalizations
2026
Cited
0
MeanFlowSE: One-Step Generative Speech Enhancement via MeanFlow
2025
Cited
0
LLaSE-G1: Incentivizing Generalization Capability for LLaMA-based Speech Enhancement
2025
Cited
0