Faithful Where It Can Be Checked: Auditing a Reflection Agent Against Its System Prompt in a Randomized Trial

πŸ“… 2026-09-16
πŸ“ˆ Citations: 0
✨ Influential: 0
πŸ“„ PDF
πŸ€– AI Summary
η ”η©Άι€šθΏ‡ε―Ήζ―”θ―•ιͺŒε’Œε―Ήθ―ηΌ–η οΌŒε‘ηŽ°θŒδΈšεζ€δ»£η†εœ¨ζ‰§θ‘Œη³»η»Ÿζη€Ίζ—Άε­˜εœ¨δΈδΈ€θ‡΄οΌŒε°€ε…Άζ˜―ε―Ήε†³η­–ηš„θ¦ζ±‚ε’žεŠ δΊ†ε‚δΈŽθ€…ηš„η–‘θ™‘γ€‚
πŸ“ Abstract
Conversational agents are increasingly used to guide reflection. A recent randomized trial compared a GPT-4o career reflection agent with the same program in a static journaling survey. Agent participants ended less committed to their career plans and more doubtful. We coded all 17,930 turns from its two studies, checked our coding against human coders and linked conversations to the trial's surveys. The rules the agent followed were the easy-to-check ones, like a reply length cap. Told not to flatter, it praised participants in half of its turns; told to challenge gently, it almost never did, and such a break leaves no visible trace. The behavior tied to the worse outcome was the demand to decide: the survey posed each decision once, while the agent asked again when participants hesitated, and those pressed most ended most doubtful. Our findings inform reflection agent design and the writing of checkable instructions.
Problem

Research questions and friction points this paper is trying to address.

Reflection Agent
System Prompt
Randomized Trial
Career Reflection
Innovation

Methods, ideas, or system contributions that make the work stand out.

Reflection Agent
System Prompt
Randomized Trial
Decision Making
Instruction Design
πŸ”Ž Similar Papers
No similar papers found.
S
Subigya K. Nepal
University of Virginia, USA
S
Serena Soh
Stanford University, USA
N
Noah Vinoya
Stanford University, USA
S
SoHyun Park
NAVER Cloud, Republic of Korea
M
Mahnaz Roshanaei
Stanford University, USA
G
Gabriella Harari
Stanford University, USA