Callability Is Not Operability: Controlled Interface Interventions for LLM Agents

📅 2026-08-23
📈 Citations: 0
✨ Influential: 0
📄 PDF
🤖 AI Summary
研究解决了工具调用与操作性之间的差距问题,通过Agent-First Tooling机制和AFT-Bench框架来确保代理在不确定性下安全继续执行。
📝 Abstract
A tool call can be perfectly valid yet still leave an autonomous agent unable to determine what to do next. For example, if an external effect commits but its response is lost, committed and uncommitted states may become indistinguishable to the agent even though they require different continuation actions. We study this gap between callability and operability: whether a tool interface exposes the action-relevant state and semantics needed for an agent to continue safely under operational uncertainty. We operationalize tool operability through Agent-First Tooling (AFT), a set of interface mechanisms spanning selective capability discovery, execution lifecycle and recovery, explicit external-effect semantics, machine-readable results, and postcondition verification. We introduce AFT-Bench, a controlled interface-intervention framework that holds the task, backend, initial state, injected failure, agent, and language model fixed while varying the interface exposed to the agent.
Problem

Research questions and friction points this paper is trying to address.

callability
operability
autonomous agent
tool interface
operational uncertainty
Innovation

Methods, ideas, or system contributions that make the work stand out.

Agent-First Tooling
operational uncertainty
interface mechanisms
AFT-Bench