Scholar
Shuang Qiu
Google Scholar ID: -Z7fY00AAAAJ
City University of Hong Kong
Reinforcement Learning
Agentic AI
Large Language Models
Embodied AI
Follow
Homepage
↗
Google Scholar
↗
Citations & Impact
All-time
Citations
1,125
H-index
17
i10-index
22
Publications
20
Co-authors
0
Contact
GitHub
Open ↗
Publications
29 items
Test-Time Scaling via Budgeted Multi-Attribute Verification
2026
Cited
0
Train Where the Quantized Model Goes: On-Policy Distillation for Low-Bit Reasoning
2026
Cited
0
Efficient Online Lexicographic Generalized Low-Rank Matrix Bandits
2026
Cited
0
SPOT: Sparse Probing and Outcome Calibration for On-Policy Distillation
2026
Cited
0
Cost-Aware Multi-Objective Bandits: Theory and Application to Budgeted LLM Configuration Evaluation
2026
Cited
0
DemoPSD: Disagreement-Modulated Policy Self-Distillation
2026
Cited
0
ReasonAlloc: Hierarchical Decoding-Time KV Cache Budget Allocation for Reasoning Models
2026
Cited
0
Where to Refine, When to Stop: Rethinking Redundancy via Latent Discrepancy for Efficient Visual Autoregressive Generation
2026
Cited
0
Load more