Jul 24
•
Articles
• 2 min read
Claims such as “90% less output” or “50% less context” are useful diagnostics, but they do not answer the question that matters for a coding agent: did the complete, verified task become cheaper?
I ran a small matched experiment to see how different...