NewXResearch hub
ExploreLibraryProfile
Account
Sign In
Yantao Liu
Scholar

Yantao Liu

Google Scholar ID: CKieAy4AAAAJ
Qwen, Alibaba
Reinforcement LearningReward ModelingLarge Language Models
Homepage↗Google Scholar↗
Citations & Impact
All-time
Citations
4,296
 
H-index
10
 
i10-index
10
 
Publications
15
 
Co-authors
0
 
Publications
8 items
JPO: Juris Policy Optimization for Structured Legal Reasoning in Criminal Judgment Prediction
2026
Cited
0
Qwen-AgentWorld: Language World Models for General Agents
2026
Cited
0
OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language World Models
2026
Cited
0
Outcome Accuracy is Not Enough: Aligning the Reasoning Process of Reward Models
2026
Cited
0
StockBench: Can LLM Agents Trade Stocks Profitably In Real-world Markets?
2025
Cited
0
Are Reasoning Models More Prone to Hallucination?
2025
Cited
0
Toward Generalizable Evaluation in the LLM Era: A Survey Beyond Benchmarks
2025
Cited
0
Pairwise RM: Perform Best-of-N Sampling with Knockout Tournament
2025
Cited
0