Scholar
Zhehuai Chen
Google Scholar ID: AZrMB-AAAAAJ
NVIDIA
Speech Recognition
Speech Synthesis
LLM
Follow
Homepage
↗
Google Scholar
↗
Citations & Impact
All-time
Citations
2,224
H-index
22
i10-index
49
Publications
20
Co-authors
35
list available
Publications
23 items
All In Good Time: Causality-Aware Framework for LLM-Based Simultaneous Speech-to-Speech Translation
2026
Cited
0
NemotronLabs VoiceChat: An Open Full-duplex Speech-to-Speech Model with Tool Calling Capabilities
2026
Cited
0
A frontend-backend architecture for tool calls in full-duplex speech models
2026
Cited
0
Enabling Streaming User Transcription in Full-Duplex Speech-to-Speech Models
2026
Cited
0
JarvisBench: Always-on Intelligence Between Humans and Agents
2026
Cited
0
VoiceChat-TTS: A Low-Latency Continuous Speech Synthesis Model for Interactive Agents
2026
Cited
0
Voice Memory for Agentic Speech Recognition
2026
Cited
0
Just A Rather Very Intelligent Spoken Agent
2026
Cited
0
Load more
Co-authors
17 total
Bhuvana Ramabhadran
Director/Principal Research Scientist, Google DeepMind
Andrew Rosenberg
Google DeepMind
Boris Ginsburg
NVIDIA
Gary Wang
Google
Kai Yu(俞凯)
Shanghai Jiao Tong University
Yonghui Wu
Head of Research, ByteDance Seed
Ankur Bapna
Google Deepmind
Piotr Żelasko
Principal Research Scientist @ Nvidia