Executive summary
Hiro's autonomous improvement workflow now treats real interaction failures as higher-priority work in the same ranked queue used for external upgrade ideas.
The previously diagnosed Los Angeles event-discovery failure and its failed correction were harvested as separate private incidents, reproduced with deterministic local contracts, repaired, and recorded as implemented outcomes.
Automatic Self-Improvement V2 production-chat probes were retired. External discovery remains active, while candidate construction, evaluation, and probation wait for a quiet production window and never use the live chat endpoint as a test harness.
Future candidates use one eight-hour probation contract with checkpoints at 0, 60, 240, and 480 minutes. Routine candidates retain standing authority; the dashboard reports the ranked idea, origin lane, next action, affected files, test gates, deadline, and terminal result.
A live acceptance test initially exposed an additional public-event routing defect: the word 'event' sent a public discovery request into the private-calendar resolver, and correction escalation then produced plausible but unsupported recommendations. That defect was repaired during this session rather than accepted as a passing test.
The final live acceptance returned concrete August activities directly from current web-search evidence with source URLs. Its universal quality record contained one web-search tool receipt, three evidence records, a passing final gate, and no repair request.
A concurrent worker later attempted to overwrite an authorized implemented outcome with a stale rejection. Queue transitions now use expected-state compare-and-set protection, and the affected incident was reconciled with an append-only audit event.