MEMORY

THE COLONY
LEARNS.

Missions leave trails. Trails that led to verified work strengthen; everything else fades. Reinforcement learned before reproducible evidence would reward persuasive prose rather than demonstrated work — so learning is gated on the canonical, evidence-backed outcome of each mission.

HOW TRAILS ARE EARNED

POSITIVE — ONLY FROM VERIFIED WORK

A role earns a positive trail when its output was consumed downstream, required evidence passed, and the mission ended completed_verified — the only positive learning signal. The Tester catching a real failure and the Soldier correctly blocking a dangerous patch are also positive signals for them.

NON-POSITIVE OUTCOMES

Completed-but-unverified, partial, timed-out, and failed missions never reinforce positively — the episode is stored, without reward. A typed failure attributable to a role's output records a negative trail for that role and task type.

FAIR ATTRIBUTION

Cancellation is neutral, and a role is never punished for not running. Provider outages update provider reliability, and environmental tool failures update tool reliability — not the worker's skill.

DECAY TOWARD NEUTRAL

Trails fade toward neutral over time. A route that stops proving itself stops being preferred; old success is not a permanent verdict.

WHAT IS ACTUALLY STORED

CURRENT LIMITS, STATED PLAINLY

Sources: PLAN.md (Stage E, gap table), ANT_EXECUTION.md (outcome semantics), and ADR-004 (artifact store).

NEXT — WHERE THE COLONY IS GOING