Executive summary
Hiro qualified exactly one already-authorized hosted path, OpenAI GPT-5.6 Sol, for the narrow claim-extraction boundary. Qwen 3.8 remained Hiro's local central model and local independent validator; production routing was not changed.
The qualification reused the latest frozen two-stage grounded extraction contracts, exact-span parser, canonical claim assembler, six gold claims, six-source controls, independent validator, thresholds, and stop rules. Only the extraction transport changed from local llama.cpp to the hosted Responses endpoint.
Hosted service operation was clean during the permitted gold stage: nine of nine requests completed on their first attempt with zero API errors, timeouts, retries, or terminal failures. Mean latency was 2.472121 seconds and p95 latency was 4.289065 seconds.
The unchanged gold gate failed on precision. All six expected evidence fragments were covered, but the model produced ten independently validated grounded claim units rather than exactly six, including one false claim-bearing span. Fabricated fields, unsupported surviving claims, parser failures, request failures, and validator failures were all zero.
The required stop rule was applied. The frozen six-source corpus, repeated service-reliability workload, frozen twenty-source corpus, production integration, and Phase 3F were not run. Final disposition: HOSTED EXTRACTION NOT QUALIFIED — QUALITY.