Scholar
Leo Gao
Google Scholar ID: r6mBY50AAAAJ
EleutherAI
AI Alignment
Language Modelling
Deep Learning
Follow
Homepage
↗
Google Scholar
↗
Citations & Impact
All-time
Citations
34,729
H-index
19
i10-index
24
Publications
20
Co-authors
13
list available
Publications
2 items
Weight-sparse transformers have interpretable circuits
2025
Cited
0
Monitoring Reasoning Models for Misbehavior and the Risks of Promoting Obfuscation
2025
Cited
1
Co-authors
11 total
Stella Biderman
EleutherAI
Jason Phang
New York University
Jeffrey Wu
Anthropic AI, OpenAI
Jan Leike
Anthropic
John Schulman
Thinking Machines
Jacob Hilton
Alignment Research Center
Collin Burns
Researcher, Anthropic
Pavel Izmailov
Anthropic; NYU