Tag
#agent-harness
Reports tagged agent-harness.
Reports
5 reportsSep 2026
09-20 agent-harness How an AI Agent Should Decide Which Decisions Deserve Credit or Blame A final score does not explain a trajectory. Reliable improvement reconstructs the decision, tests fair alternatives, measures uncertainty, and applies the smallest reversible repair. 09-19 agent-harness When an AI agent’s past becomes useful An agent learns from experience only when past outcomes change later choices through scoped, testable, revisable guidance that survives variation without hidden regressions or disproportionate cost. 09-13 agent-harness When an Agent Fails, Should You Change the Model or the System? The useful distinction is not weak intelligence versus bad engineering. It is which intervention removes the failures that matter—and whether the gain survives realistic costs and checks. 09-13 agent-harness When can an AI agent say the task is complete? Completion is a judgment about the requested result—not a synonym for stopping, passing a check, or sounding confident. 09-07 agent-harness Can a Harness Make a Small Model Match a Large One? A harness is a bigger lever than most engineers assume, and the weaker model is the more exposed to it. Whether a small model can substitute is decided by the task, not the scaffolding.