Scholar
Yongjin Yang
Google Scholar ID: qGVZm3sAAAAJ
University of Toronto
RL
LLM Agent
Generative Models
Alignment
Follow
Homepage
↗
Google Scholar
↗
Citations & Impact
All-time
Citations
143
H-index
6
i10-index
4
Publications
16
Co-authors
34
list available
Publications
17 items
IdeaScientist: Orchestrating Agents for Grounded Scientific Ideation
2026
Cited
0
Toward Skill-Native LLMs: Skill Entropy for Benchmarking and Training Long-Horizon Reasoning
2026
Cited
0
LLM-as-a-Tutor: Policy-Aware Prompt Adaptation for Non-Verifiable RL
2026
Cited
0
Transferability for General Reasoning: An Automated Curriculum for Multi-Domain RLVR
2026
Cited
0
MixSD: Mixed Contextual Self-Distillation for Knowledge Injection
2026
Cited
0
Self-Distillation Zero: Self-Revision Turns Binary Rewards into Dense Supervision
2026
Cited
0
UniSAFE: A Comprehensive Benchmark for Safety Evaluation of Unified Multimodal Models
2026
Cited
0
Entropy-Aware On-Policy Distillation of Language Models
2026
Cited
0
Load more
Co-authors
16 total
Se Young Yun
KAIST
Yujin Kim
PhD student at KAIST AI
Namgyu Ho
PhD student at KAIST
Haneul Yoo
KAIST
Joonkee Kim
LG AI Research
James Thorne
KAIST
Tergel Munkhbat
KAIST
Seo Hyun Kim
KAIST