Executive summary
Hiro and its local Qwen 3.8 model were already running when restart verification began. Qwen was loaded with the intended 8,192-token context and a single parallel slot, while Hiro exposed its HTTP application and auxiliary local listeners.
The production health endpoint reported an OK service, a connected LLM, and qwen/qwen3.8-27b. A real non-streaming chat request returned the exact requested readiness token in 6.9 seconds.
The complete regression suite passed 635 tests in 171.79 seconds. The checked-out revision did not change during the run, so the result applies cleanly to c004c1e rather than a moving target.
The benchmark dashboard rendered the risk-based 15-minute low-risk and 60-minute moderate-risk schedules, status-card filtering, the independent model ledger, and the active ranked queue without browser warnings or errors.
The autonomous loop remained live throughout verification. It advanced from one travel-planning candidate to a writing-assistance candidate, leaving one active item, 13 waiting, 21 retrying, three implemented, and 146 rejected with the infrastructure circuit breaker closed at zero failures.