RESEARCHInvestigateNEXT 12 MONTHS
Prompt-Induced Waste in Large Reasoning Models: A Preregistered Two-Harness Benchmark of Coding Agents
arXiv cs.CL — Computation and Language
Factual evidence
What the source reports
A preregistered benchmark of six reasoning models reveals that prompt wording significantly drives token waste and cost in agentic coding tasks.
Open source