A team builds a code-audit agent with the Claude Agent SDK. The main agent plans the audit and writes the final report, and it delegates repository exploration (reading files, grepping for patterns) to a subagent that never edits anything. Exploration generates most of the token volume, and the team wants to cut cost and latency without weakening the main agent's reasoning. What is the best approach?