Executive summary
The two recent promotions were valid code improvements, but they were not evidence that Hiro had achieved a self-sustaining autonomous improvement loop. Both emerged from long-lived candidates that were repeatedly made-next, rebuilt after platform changes, and carried through invalid harness or governor outcomes during direct supervision.
The arithmetic candidate had been in the queue since August 17. It was explicitly made-next three times, crossed multiple builder and harness revisions, passed candidate tests several times, and was finally promoted on August 22 after a governor full-suite failure was diagnosed as infrastructure-related and retried.
The prompt-injection candidate was also made-next three times. Its successful path required a sequence of platform corrections for editing, missing task context, a double-applied response boundary, production reachability, an invalid regression expectation, and a governor manifest mismatch. Twelve repository commits occurred during the concentrated supervision interval before its final promotion.
After supervision stopped, Hiro reverted to breadth-first queue churn. In the current post-repair window, 28 ideas received one investigation each and only three ideas reached three investigations. None reached candidate_tests_passed. The missing component is therefore an encoded supervisory control loop that holds focus, clusters systemic failures, repairs the builder or harness, and resumes the same candidate with accumulated causal feedback.