Structure Tax: How Structured Output affects LLMs Performance

📅 2026-10-08
📈 Citations: 0
✨ Influential: 0
📄 PDF
🤖 AI Summary
This study addresses the "structure tax" problem, wherein enforcing structured outputs in large language models degrades accuracy, by conducting a systematic evaluation across multiple models and datasets. Methodologically, it employs Centered Kernel Alignment (CKA) to analyze representation separability in intermediate Transformer layers, alongside confidence calibration and hidden-layer geometric measurements. The findings reveal that accuracy degradation stems from schema design rather than structural constraints per se, proposing a paradigm shift from "whether to structure" to "how to structure." Experiments demonstrate that a reasoning-prioritized field ordering strategy matches or surpasses free-text performance while significantly improving model calibration.
📝 Abstract
Deploying large language models in production often requires constraining outputs to structured formats such as JSON or XML, and prior work treats the resulting accuracy loss as an inherent `structure tax'. We re-examine this claim by evaluating a battery of models, datasets and schemas, measuring task accuracy, confidence calibration, and hidden-state geometry. The tax turns out to depend on schema design rather than on structure per se: reasoning-first field ordering matches or exceeds free-form accuracy, while answer-first ordering causes steep drops, particularly in smaller models. Format sensitivity scales inversely with a task's own structural constraints, and schemas that preserve reasoning order also improve calibration with CKA showing greater separability between correct and incorrect representations in middle transformer layers. Our findings indicate that properly designed structured formats can match or exceed free-form performance, reframing the critical question from `whether to structure' to `how to structure' for optimal reasoning preservation.
Problem

Research questions and friction points this paper is trying to address.

Large Language Models
Structured Output
Structure Tax
Schema Design
Reasoning Preservation
Innovation

Methods, ideas, or system contributions that make the work stand out.

Structured Output
Schema Design
Reasoning-First Ordering
Confidence Calibration
Representation Geometry
🔎 Similar Papers
No similar papers found.
💼 Related Jobs
No related jobs found.
V
Vineet Kumar
PayPal Artificial Intelligence, PayPal, Bengaluru India
K
Kanishka
PayPal Artificial Intelligence, PayPal, Bengaluru India
B
Bhuvanesh Mandora
PayPal Artificial Intelligence, PayPal, Bengaluru India