PerceptUI: LLM Agents as Human-Aligned Synthetic Users for UI/UX Evaluation

📅 2026-06-04
📈 Citations: 0
✨ Influential: 0
📄 PDF
🤖 AI Summary
This work addresses the high cost, lengthy timelines, and limited fidelity of traditional UI/UX evaluation methods—such as user studies and A/B testing—in simulating authentic user feedback. To overcome these limitations, the authors propose PerceptUI, a novel framework that enables fine-grained, persona-based UI/UX feedback generation for the first time. Leveraging multimodal large language models, PerceptUI integrates persona-conditioned prompting, contrastive reflection-based fine-tuning, and a failure-trajectory-driven prompt evolution mechanism to distill rational justifications from human decision-making processes, thereby enhancing the model’s introspective capabilities. Experimental results demonstrate that PerceptUI achieves human-level authenticity in generated feedback across multiple domains and datasets, generalizes effectively to unseen interface issues and user personas, and supports the synthesis of population-level response distributions.
📝 Abstract
User interface (UI) and user experience (UX) evaluation is central to product development, yet reliable feedback still relies on recruiting human participants or running online A/B tests, making early-stage iteration slow and costly. In light of this, recent work has explored Multimodal Large Language Models as proxy evaluators. However, existing approaches either produce surface-level critiques or a judgment that reflects the model's own biases rather than the genuine response of a particular user. We introduce PerceptUI, a framework for persona-conditioned UI/UX evaluation that predicts how a specific user would answer interface-related questions and produces natural-language rationales. PerceptUI is trained in two stages: (i) contrastive reflection fine-tuning distills teacher-generated rationales by extracting lessons from human decisions, and (ii) a reflective prompt-evolution step from the model's own failure traces. Across multiple domains and datasets, PerceptUI achieves human-level realism, generalizes to unseen questions and personas, and yields population-level response distributions.
Problem

Research questions and friction points this paper is trying to address.

UI/UX evaluation
human-aligned feedback
synthetic users
persona-conditioned evaluation
user experience
Innovation

Methods, ideas, or system contributions that make the work stand out.

persona-conditioned evaluation
multimodal LLM agents
contrastive reflection fine-tuning
reflective prompt evolution
synthetic user modeling
N
Nicolas Bougie
Woven by Toyota
X
Xiaotong Ye
Woven by Toyota
G
Gian Maria Marconi
Woven by Toyota
Narimasa Watanabe
Narimasa Watanabe
Woven by Toyota