The fastest run so far (loop71, v146) built taskboard in about 38 minutes. Straight from the run records, here's what the local LLMs were doing between receiving the spec and the PASS.
Width is time. Color is the model (blue = Qwen3.8-Flash-Next, green = Gemma-4-26B-A4B-it, yellow = Qwen3.8-27B).
"+N min" is time since the start. The seconds on the right are how long the LLM took to answer.
All numbers come from loop71's run records (escalation-chain-run.json, llm_calls.jsonl, token-summary.json). The Claude API estimate uses Opus 5.5 list prices with the measured cache hits.