Scholar
Stephen Casper
Google Scholar ID: zaF8UJcAAAAJ
PhD student, MIT
AI safety
AI responsibility
red-teaming
robustness
auditing
Follow
Homepage
↗
Google Scholar
↗
Citations & Impact
All-time
Citations
2,834
H-index
22
i10-index
35
Publications
20
Co-authors
39
list available
Publications
33 items
From Constitutions to Control: Interpretable Rewards for Aligning Language Models
2026
Cited
0
Embedded Assessments for Frontier AI
2026
Cited
0
Underwriting the Agent Economy: The Blueprint for an AI Insurance Stack
2026
Cited
0
Open Weight AI Models Require Proportional Evaluation Approaches
2026
Cited
0
Open Problems in Frontier AI Risk Management
2026
Cited
0
International AI Safety Report 2026
2026
Cited
6
The 2025 AI Agent Index: Documenting Technical and Safety Features of Deployed Agentic AI Systems
2026
Cited
0
TamperBench: Systematically Stress-Testing LLM Safety Under Fine-Tuning and Tampering
2026
Cited
1
Load more
Co-authors
16 total
Dylan Hadfield-Menell
Massachusetts Institute of Technology
Gabriel Kreiman
Professor, Harvard Medical School and Children's Hospital
Daniel Filan
PhD Student, UC Berkeley
Andrew Critch
UC Berkeley, Department of Electrical Engineering and Computer Sciences
Shlomi Hod
PhD Candidate, Boston University
David Scott Krueger
Assistant Professor, University of Montreal, Mila
Anson Ho
Epoch AI
Cody Wild
Google DeepMind