NewXResearch hub
ExploreLibraryProfile
Account
Sign In
Thilo Hagendorff
Scholar

Thilo Hagendorff

Google Scholar ID: GlhI8lQAAAAJ
Research Group Leader, University of Stuttgart
AI SafetyAI EthicsMachine PsychologyLarge Language Models
Homepage↗Google Scholar↗
Citations & Impact
All-time
Citations
5,324
 
H-index
30
 
i10-index
43
 
Publications
20
 
Co-authors
0
 
Contact
CVOpen ↗GitHubOpen ↗
Publications
12 items
Deception by Omission: Language Models Knowingly Hide Their Mistakes
2026
Cited
0
How Much Do LLM-as-a-Judge Design Choices Matter? A Systematic Comparison of Prompt Designs, Rating Scales, and Models
2026
Cited
0
Shutdown Sabotage Propensities in Multi-Agent Systems
2026
Cited
0
Evaluation Awareness in Language Models Has Limited Effect on Behaviour
2026
Cited
0
"Dark Triad"Model Organisms of Misalignment: Narrow Fine-Tuning Mirrors Human Antisocial Behavior
2026
Cited
0
Emergently Misaligned Language Models Show Behavioral Self-Awareness That Shifts With Subsequent Realignment
2026
Cited
0
Speciesism in AI: Evaluating Discrimination Against Animals in Large Language Models
2025
Cited
0
Large Reasoning Models Are Autonomous Jailbreak Agents
2025
Cited
0