Natural-Language Agent Harnesses

📅 2026-03-26
📈 Citations: 0
✨ Influential: 0
📄 PDF
🤖 AI Summary
This work addresses the critical yet often overlooked dependence of agent performance on harness engineering, where control logic is typically entangled within code, hindering transferability, reuse, and systematic study. To overcome this limitation, we propose— for the first time—externalizing the high-level control logic of harnesses into editable, executable natural language specifications, supported by a unified Intelligent Harness Runtime (IHR) architecture. The IHR introduces explicit contracts, persistent artifacts, and lightweight adapters to enable modular harness design and cross-task transfer. We validate our approach on programming and computer-use benchmarks, demonstrating its effectiveness through comprehensive experiments. Ablation studies and successful transfers from code-based to text-based harnesses further illustrate the framework’s flexibility and feasibility.

Technology Category

Natural Language Processing: Code Generation / Program Synthesis from Natural LanguageIntelligent Robots: Behavior Learning & ControlCognitive Modeling & Cognitive Systems: Agent Architectures

Application Category

Responsible Web: Machine-in-the-loop, human agency and autonomySemantics and Knowledge: Data modeling to support human-machine intelligence, including LLMs agents, intelligent system behavior, explanations, and user-friendly interactionsEconomics, Online Markets and Human Computation: Cost models of using LLMs in production systems
📝 Abstract
Agent performance increasingly depends on \emph{harness engineering}, yet harness design is usually buried in controller code and runtime-specific conventions, making it hard to transfer, compare, and study as a scientific object. We ask whether the high-level control logic of an agent harness can instead be externalized as a portable executable artifact. We introduce \textbf{Natural-Language Agent Harnesses} (NLAHs), which express harness behavior in editable natural language, and \textbf{Intelligent Harness Runtime} (IHR), a shared runtime that executes these harnesses through explicit contracts, durable artifacts, and lightweight adapters. Across coding and computer-use benchmarks, we conduct controlled evaluations of operational viability, module ablation, and code-to-text harness migration.
Problem

Research questions and friction points this paper is trying to address.

agent harness
harness engineering
natural language
portable artifact
executable specification
Innovation

Methods, ideas, or system contributions that make the work stand out.

Natural-Language Agent Harness
Intelligent Harness Runtime
harness engineering
executable artifact
agent control logic
🔎 Similar Papers
No similar papers found.
L
Linyue Pan
Shenzhen International Graduate School, Tsinghua University
L
Lexiao Zou
Harbin Institute of Technology (Shenzhen)
Shuo Guo
Shuo Guo
University of Minnesota, Twin Cities
Wireless Sensor NetworksVehicular NetworksDelay Tolerant NetworksContent Oriented NetworksOverlay Networks
J
Jingchen Ni
Shenzhen International Graduate School, Tsinghua University
H
Hai-Tao Zheng
Shenzhen International Graduate School, Tsinghua University