Scholar
Phil Blandfort
Google Scholar ID: NCtMAPMAAAAJ
Independent Researcher
AI
Machine Learning
AI Safety
Alignment
Follow
Google Scholar
↗
Citations & Impact
All-time
Citations
131
H-index
5
i10-index
4
Publications
14
Co-authors
18
list available
Publications
5 items
An Investigation of Model Coherence: Narrow Finetunes Contradict Themselves Under Resampling
2026
Cited
0
Strangers to Themselves: What Language Models Say About Themselves Is Generic
2026
Cited
0
Moral Preferences of LLMs Under Directed Contextual Influence
2026
Cited
0
Red-teaming Activation Probes using Prompted LLMs
2025
Cited
0
Detecting High-Stakes Interactions with Activation Probes
2025
Cited
0
Co-authors
15 total
Desmond Upton Patton
University of Pennsylvania
Andreas Dengel
Professor of Computer Science, University of Kaiserslautern & Executive Director, DFKI
Tushar Karayil
Researcher DFKI
Jörn Hees
H-BRS, DFKI
Rossano Schifanella
University of Turin
Shih-Fu Chang
Professor of Electrical Engineering and Computer Science, Columbia University
Damian Borth
Professor of Artificial Intelligence & Machine Learning, University of St. Gallen
Kathleen McKeown
Professor of Computer Science and Director, Data Science Institute, Columbia University