
YZ
Yukun Zhang, Kemu Xu, Yishen Chen
· 1 min read
ResearcharXiv cs.AI
The Organization of Inference: Information, Resource Constraints, and AI Production
arXiv:2609.20449v1 Announce Type: new
Abstract: The economic value of inference depends on how capacity and task information are distributed across stages of AI production. We study these organizational margins using controlled workflow experiments on externally verified software-engineering tasks. In two matched resource panels, direct execution records the same success rate of 59.6 percent at logical-token ceilings of 12,000 and 24,000, while success under information-constrained planning rises from 36.2 to 51.2 percent. The planning disadvantage narrows by 15.0 percentage points (95 percent task-cluster bootstrap interval: 4.2 to 25.8). A strict read-only planning campaign varies whether the planner sees the task issue. At 12,000 tokens, issue access raises success by about 16 percentage points over issue-hidden planning. Compared with direct execution, task-informed planning is about 10 points lower at 12,000 tokens; at 24,000 tokens, it shows a 29.6-point advantage. In the resource panels, direct execution uses substantially less than either ceiling, while the planning workflow's binding rate falls from 46.2 to 0.8 percent and downstream execution accounts for 89.9 percent of the increase in total use. Scale determines the capacity available to a system; workflow and information structure shape the productive value
Original source
This story was published by arXiv cs.AI and written by Yukun Zhang, Kemu Xu, Yishen Chen. SyncAI.news shows a preview; the complete article is on the publisher's site.
Read the full story on arxiv.org


