Executive summary
Hiro's rank-one open-problem challenge is now connected to the existing autonomous candidate pipeline through a durable single-item queue. The bridge freezes an ImprovementSpec, constructs at most one patch in an external worktree, runs the existing public, held-out, regression, latency, and invariant gates, and stops before integration.
Qwen 3.8 27B was loaded locally and used for the first bounded cycle. Four candidate-construction runs exercised the system: the harness discovered and repaired an inconclusive baseline, missing repair feedback, a weak status-trusting solution, and an evaluator-attribution problem. No candidate was promoted, integrated, or applied to Hiro's active branch.
The strongest candidate passed its targeted tests and one replicated global evaluation, but the same public epistemics-category regression appeared in its original evaluation and a second replication. A predeclared two-clean-replication rule therefore stopped it before the independent adversarial challenge. The result is negative but useful evidence: uncertainty failed closed and the next research target is evaluator variance and causal attribution.