Scholar
Ziyang Ma
Google Scholar ID: 4RZnXGMAAAAJ
Shanghai Jiao Tong University
Speech and Language Processing
Textless NLP
Self-supervised Learning
Multimedia
Follow
Homepage
↗
Google Scholar
↗
Citations & Impact
All-time
Citations
1,894
H-index
24
i10-index
36
Publications
20
Co-authors
30
list available
Publications
17 items
WorldSonus: Bringing Sound to Worlds
2026
Cited
0
TinyAudio: Compact and Efficient Text-to-Audio Generation for Low-Resource Deployment
2026
Cited
0
Omni Demand Understanding: A Benchmark for Contextual User-Intent Inference in Multimodal Interaction
2026
Cited
0
EmoS: A Theory-Grounded Framework for Evaluating and Aligning Emotional Intelligence in Spoken Language Models
2026
Cited
0
GROW: Group-Relative Advantage-Weighted On-Policy Reinforcement Learning of Autoregressive-Diffusion Text-to-Speech model
2026
Cited
0
Native Active Perception as Reasoning for Omni-Modal Understanding
2026
Cited
0
Speech Meets ELF: Audio Conditional Continuous-Target Diffusion for Speech Recognition and Translation
2026
Cited
0
MMAE: A Massive Multitask Audio Editing Benchmark
2026
Cited
0
Load more
Co-authors
20 total
Xie Chen
Shanghai Jiao Tong University <- Microsoft <- Cambridge University
ShiLiang Zhang
Unknown affiliation
Kai Yu(俞凯)
Shanghai Jiao Tong University
Zhisheng Zheng
The University of Texas at Austin
Yifan Yang
Shanghai Jiao Tong University, Tencent, Microsoft, Xiaomi
gao zhifu
Tongyi Lab, Alibaba Group
Guanrou Yang
Shanghai Jiao Tong University
Zhihao Du
Alibaba