AgoraResearch hub
ExploreLibraryProfile
Account
Sign In
Shuang Qiu
Scholar

Shuang Qiu

Google Scholar ID: -Z7fY00AAAAJ
City University of Hong Kong
Reinforcement LearningAgentic AILarge Language ModelsEmbodied AI
Homepage↗Google Scholar↗
Citations & Impact
All-time
Citations
1,125
 
H-index
17
 
i10-index
22
 
Publications
20
 
Co-authors
0
 
Contact
GitHubOpen ↗
Publications
27 items
Efficient Online Lexicographic Generalized Low-Rank Matrix Bandits
2026
Cited
0
SPOT: Sparse Probing and Outcome Calibration for On-Policy Distillation
2026
Cited
0
Cost-Aware Multi-Objective Bandits: Theory and Application to Budgeted LLM Configuration Evaluation
2026
Cited
0
DemoPSD: Disagreement-Modulated Policy Self-Distillation
2026
Cited
0
ReasonAlloc: Hierarchical Decoding-Time KV Cache Budget Allocation for Reasoning Models
2026
Cited
0
Where to Refine, When to Stop: Rethinking Redundancy via Latent Discrepancy for Efficient Visual Autoregressive Generation
2026
Cited
0
Unifying Value Alignment and Assignment in Cross-Domain Offline Reinforcement Learning with Heterogeneous Datasets
2026
Cited
0
Reference-Sampled Boltzmann Projection for KL-Regularized RLVR: Target-Matched Weighted SFT, Finite One-Shot Gaps, and Policy Mirror Descent
2026
Cited
0
Resume (English only)
Co-authors
0 total
Co-authors: 0 (list not available)