Primary-source and existing-state review
Completed- Read https://github.com/Contrastive-LM/CLM, src/clm/embedder.py, src/clm/engine.py and https://huggingface.co/Contrastive-LM/CLM-v0.1-8B after following the supplied Reddit discussion.
- Reference heads use a frozen Qwen3-8B encoder with last-token pooling; the small head download does not include the encoder.
- Current serving documentation uses vLLM. The embedding client accepts an OpenAI-compatible endpoint, but that does not prove quantized llama.cpp embeddings preserve trained-head rankings or calibration.
- Inspected existing Hiro opportunity priority, continuous idea scoring and bounded external-lead selection. A learned scorer could be compared at these existing boundaries, without creating another research queue.
- Checked open handoffs; the existing ranked-autonomy handoff was not executed.