Scholar
Qi Wang
Google Scholar ID: GsgwHM4AAAAJ
Tsinghua University
Operation research
Reinforcement learning
Follow
Google Scholar
↗
Citations & Impact
All-time
Citations
63
H-index
3
i10-index
3
Publications
3
Co-authors
0
Publications
6 items
You Only Edit Once: Incentivizing In-Context Capability of LLMs via Local Demonstration Refinement
2026
Cited
0
TRACE: A Unified Rollout Budget Allocation Framework for Efficient Agentic Reinforcement Learning
2026
Cited
0
RLVR without Ineffective Samples: Group Prioritized Off-Policy Optimization for LLM Reasoning
2026
Cited
0
Listwise Policy Optimization: Group-based RLVR as Target-Projection on the LLM Response Simplex
2026
Cited
0
Dynamics-Predictive Sampling for Active RL Finetuning of Large Reasoning Models
2026
Cited
0
Small Generalizable Prompt Predictive Models Can Steer Efficient RL Post-Training of Large Reasoning Models
2026
Cited
0