developer tools
Long Tasks Favored a Fresh Context Over a Bigger Instruction Pile
A study found subagent execution stronger than loading reusable skill packages directly when each skill exposed a clear input-output contract.

Summary
A study found subagent execution stronger than loading reusable skill packages directly when each skill exposed a clear input-output contract.
The researchers compare two ways of reusing procedural knowledge: putting a skill package into the main agent’s growing context, or invoking that package in a fresh subagent context. On long-horizon tasks, the subagent approach performed better when the package stated a clear input-output contract and encoded actionable procedure. The improvement came with extra coordination tokens, so the result is a trade-off rather than a universal argument for delegation. The paper’s central claim is architectural: how knowledge is invoked can matter as much as what the knowledge contains.
Why it matters
A study found subagent execution stronger than loading reusable skill packages directly when each skill exposed a clear input-output contract.
Limits and context
No additional limitation was separately recorded.
Key claims
A study found subagent execution stronger than loading reusable skill packages directly when each skill exposed a clear input-output contract.
Evidence: source-2026-09-10-003
Sources
- arXiv preprint 2609.09233arXiv · primary research
Corrections
No corrections have been recorded for this story.