← Back to issue4 / 12 · Week of Aug 10, 2026

LongHorizon-Harness keeps agent state only after an audit

LongHorizon-Harness splits a long task across a manager, executor, and auditor. The executor starts each round with fresh context, and only results checked independently in the environment enter persistent task state. Why it matters: Long tasks break when an agent's memory records guesses as progress. This design makes verification a condition for carrying work into the next round.

Try this: Run one low-risk task with and without an audit gate. Compare the final files, the state that survives a reset, and the checks recorded for each run.

Source
AMAP-ML / LongHorizon-Harness
View source →

Get the field brief every week.

Important AI developments, useful explanations, and practical resources in one weekly read. Context to understand what matters, with links to the original sources and deeper reading.

Subscribe free →
Free weekly·No spam·Unsubscribe anytime