Every agent cost investigation eventually starts with the same miserable question: which part of the prompt got expensive? The invoice says tokens. The trace says a model call happened. The runtime says the agent was “working.” None of that tells you whether the bill came from duplicated system instructions, a