Executable harnesses now determine agent cost, reliability, and test-time gains
Agent performance is increasingly determined by the executable control layer around the model. TTHE improves fixed large language models (LLMs) by editing their harnesses from unlabeled traces, while…