How we made GitHub Copilot CLI more selective about delegation (opens in new tab)
GitHub improved Copilot CLI by making subagent delegation more selective rather than treating delegation as inherently beneficial. The new orchestration policy keeps narrow tasks with the main agent, delegates broad or independent work, and encourages parallel execution instead of waiting. After full production rollout, it reduced tool failures by 23% and improved high-percentile wait times without reducing quality.
The Cost of Unnecessary Delegation
- Subagents help with complex investigations, large repositories, and parallel work, but every handoff adds tool calls, coordination, and latency.
- Copilot sometimes delegated simple, well-scoped tasks that the main agent could complete directly.
- Common problems included:
- Repeated or overlapping repository searches.
- Subagents re-discovering context already available to the main agent.
- Sequential delegation that left the main agent idle.
- Stale paths, incorrect relative paths, and workspace mismatches.
- The result was slower execution and more tool failures for tasks that should have required only a few steps.
How the Problem Was Identified
- GitHub used LLMs to analyze complete agent trajectories rather than manually reviewing sessions.
- The analysis found that delegation was frequently used for narrow, obvious, or fully described tasks.
- This led to a clear target:
- Keep focused discovery-and-edit work with the main agent.
- Reserve subagents for broad exploration, cross-cutting tasks, or genuinely independent work.
A More Selective Orchestration Policy
- Copilot now starts with the narrowest effective workflow:
- Find and read the relevant file.
- Make the targeted change.
- Verify the result.
- Delegation becomes appropriate when additional context, uncertainty, or parallel execution creates real value.
- Subagents are treated as a parallelism mechanism, not a reason for the main agent to pause.
- Handoffs should clearly specify:
- The user’s request.
- What the main agent already knows.
- Which work the subagent owns.
- What result the subagent should return.
Evaluation and Production Results
- GitHub tested the change with generated regression cases and existing benchmarks before rollout.
- Staff and public A/B tests measured reliability, responsiveness, subagent workload, and quality.
- Production results showed:
- 23% fewer tool failures per session.
- 27% fewer search-tool failures.
- 18% fewer edit-tool failures.
- 5% lower P95 wait time.
- 3% lower P75 wait time.
- No quality regression.
- The improvements came mainly from avoiding unnecessary subagent paths and reducing orchestration overhead, not from making individual model calls faster.
Copilot CLI users can access the improvement by running /update and upgrading to version 1.0.42 or later. The broader recommendation is to delegate selectively: use the main agent for focused tasks and subagents only when independent context or parallel work provides meaningful leverage.