← Back to issue2 / 3 · Week of Aug 10, 2026

LongHorizon-Harness keeps agent state only after an audit

LongHorizon-Harness splits a long task across a manager, executor, and auditor. The executor starts each round with fresh context, and only results checked independently in the environment enter persistent task state. Why it matters: Long tasks break when an agent's memory records guesses as progress. This design makes verification a condition for carrying work into the next round.

Try this: Run one low-risk task with and without an audit gate. Compare the final files, the state that survives a reset, and the checks recorded for each run.

Source
AMAP-ML / LongHorizon-Harness
View source →

Get the field brief every week.

One lead signal, three quick hits, one thing to try, one concept decoded - and the rest of the week on the wire. For people who want to know what matters and what to do next.

Subscribe free →
Free weekly·No spam·Unsubscribe anytime