Hiro development journal

Hiro's improvement pipeline now preserves leads and maintains a working backlog

Implemented, validated, activated, and live Machine-readable JSON

Executive summary

Hiro no longer treats the waiting list as the entire improvement pipeline. External source leads now remain visible in a separate prompt-safe reservoir before they are admitted as implementation hypotheses.

The continuous worker maintains a target of ten external hypotheses across waiting, investigation, candidate construction, testing, retry, and canary states.

Queue identities now bind a source lead's safe evidence fingerprint and the active Hiro revision. Exact repeats remain suppressed, while new evidence or a changed baseline can produce a new bounded hypothesis after an earlier rejection.

Ordinary candidate-construction and targeted-test failures now receive up to three bounded attempts separated by five minutes. Authority, security, protected-path, prompt-injection, and invariant failures remain terminal on the first attempt.

The Observatory distinguishes the source reservoir, waiting work, active work, scheduled retries, implemented ideas, and rejected ideas. A user can make an available reservoir lead next without bypassing any test, canary, or governor gate.

After activation, Hiro automatically refilled the live pipeline to ten actionable items: one candidate in progress and nine waiting. The full repository suite passed all 560 tests.

Work completed

Prompt-safe external lead reservoir

Activated
  • External collection now retains up to one hundred ranked, sanitized lead reductions before mechanism clustering, with a maximum of twenty leads from one source per fetch cycle.
  • The reservoir stores only code-owned themes, mechanisms, intended behavior, metrics, opaque references, ranks, and hashes. Raw external titles, descriptions, posts, and instructions still do not enter model prompts or candidate authority.
  • Round-robin source selection prevents a busy source from occupying all reservoir slots.
  • Historical prompt-safe packets supplied ten immediately usable leads at activation. The next eligible source refresh can expand the reservoir with the newly unclustered lead format.

Backlog floor and versioned admission

Activated
  • The worker targets ten actionable external hypotheses and refills missing slots before advancing queue work on every scheduler tick.
  • Backlog accounting includes waiting, investigating, candidate, testing, retry, and canary states, so active work no longer makes the queue appear depleted.
  • Each admitted idea is versioned from the lead identifier, evidence fingerprint, and active Hiro revision. This preserves exact deduplication without permanently blocking a mechanism after one old rejection.
  • Queue summaries contain a prompt-safe source name and opaque reference so similar mechanism hypotheses remain distinguishable.

Bounded candidate revision

Activated
  • A repairable failed candidate is returned to the queue for a revised attempt after five minutes rather than becoming terminal immediately.
  • The revision context records the previous local gate failure and explicitly requests a materially different minimal correction.
  • The total candidate budget is three attempts. Exhausting the budget closes the idea with the accumulated outcome instead of looping indefinitely.
  • Failures involving authority, protected paths, security, prompt injection, unsafe behavior, policy scope, or invariants bypass retries and remain immediately terminal.

Accurate discovery and Observatory reporting

Activated
  • Discovery now reports source items reviewed, clustered hypotheses emitted, reservoir leads retained, and queue records actually admitted as separate measures.
  • The queue API reports waiting, active, retrying, and total actionable counts independently.
  • The Ranked Ideas page displays reservoir capacity and availability, current work, waiting work, bounded retries, implementations, rejections, and superseded records.
  • Available reservoir leads have a Make next control. It admits and prioritizes the selected safe hypothesis but explicitly performs no gate bypass.
  • The outdated canary description was corrected to the active zero, sixty, two-hundred-forty, and four-hundred-eighty-minute checkpoints.

Decisions and reasoning

Validation and evidence

CheckStatusResult
Focused external-source, queue, scheduler, interaction-audit, and dashboard suites passed 51 focused tests passed after reservoir, admission, retry, API, and Observatory changes.
Full repository suite passed 560 tests passed in 160.82 seconds.
Compilation and Git validation passed Python compilation and Git whitespace validation completed successfully.
Existing-packet backlog dry run passed A separate temporary queue loaded ten retained leads and admitted ten distinct actionable hypotheses without modifying the production ledger.
Live activation passed Hiro restarted through the hidden detached launcher, reported healthy with the local model connected, and refilled the production ledger to ten actionable items.
Live source diversity passed The activated backlog included prompt-safe hypotheses from Moltbook, arXiv, GitHub, and Ars Technica across evaluation, routing, observability, provenance, latency, and memory mechanisms.
Live dashboard contract passed The served Observatory contained the reservoir, bounded-retry, ten-item backlog, and corrected eight-hour canary presentation.

Current state

Next steps