Executive summary
The exact Bartowski Mistral Small 3.2 24B Instruct 2506 Q6_K GGUF became available and was qualified as the sole newly authorized claim-extraction candidate.
The 19,345,944,704-byte artifact was independently hashed as 3b1f9516b3446859f145f114152b260388253b4f911528bfe7545a79a09a8874 before execution and again by the qualification runner before its immutable manifest was created.
Mistral used the unchanged llama.cpp 2.31.2 configuration, full-GPU-offload request, flash attention, 8,192-token context, one slot, frozen extraction contract, supervisor policy, and startup boundary.
The candidate never became API-ready within the frozen 300-second startup policy. It therefore failed at MODEL_START_TO_API_READY with normalized reason MODEL_RELOAD_FAILED before any extraction request was submitted.
Because no request ran, this result contains zero requests, zero attempts, zero generation hangs, zero retries, and zero completed cycles. It is a startup failure and is not mislabeled as the hard-generation-hang pattern seen in prior llama.cpp candidates.
Quality testing was not authorized. No production routing changed, the frozen twenty-source corpus was not run, and no other model was tested.
The failed startup exposed a cleanup gap for processes still loading before API readiness. Exact pre-listen cleanup was added and used to terminate only the verified candidate process, after which central Qwen was restored and passed real inference health.
Final disposition for this bounded candidate: NO QUALIFIED EXTRACTION MODEL — RUNTIME BOTTLENECK.