Executive summary
The stale-replay repair restored active, causally complete candidate construction, but it has not yet restored a healthy promotion funnel. In approximately three hours after restart, the queue began 36 investigations, confirmed 34 potentials, scheduled 30 candidate revisions, marked three artifacts blocked, and reproduced two cases as already fixed. No candidate reached paired evaluation or promotion.
Counting distinct ideas in the same interval gives 30 investigated and 28 confirmed. Twenty-seven distinct ideas received at least one construction retry. Candidate validation failure appeared in 28 retry events; three also reported that targeted tests already passed on the untouched baseline.
The honest current promotion expectation is therefore effectively zero until the construction-to-evaluation transition improves. The system is busy and Qwen is healthy, but activity at five-to-six-minute intervals is not equivalent to improvement throughput.
For a functioning version of this architecture, a reasonable operating target is one small, attributable promotion per 12 to 24 hours of active runtime, with larger evaluator or harness upgrades occurring one to three times per week. A six-hour interval with no candidate reaching paired evaluation should be treated as a pipeline incident, not as normal idea selectivity.