Autonomous subsystem improvement. Generated 2026-08-31T21:42:25+00:00. 99 snapshot(s), 1 experiment(s), 0 promotion(s).
Nothing yet. The Coach opens its first experiment once the muse has run and the mutation cooldown has elapsed.
muse.prompt: v1 challenging v0North star: calibration.brier with its sample count - the one number improvement is for, because a lucky week cannot move it and a well-calibrated view can. Guardrails (must not degrade, both directions): sizing.refused_rate and exit.uncorroborated_decisives (rule compliance - a seam losing the conditional payoff, or a stop firing on quote noise), the trade/decline balance behind attribution.attributable_rate (always-trading and never-trading are both drift, and this book has measured both), and coach.cost_usd_today with model.cal_age_days (cost and staleness). Everything else here is diagnostic.
Coach actions on the same series: experiment_opened v0 vs v1 (2026-08-29)
| lever | incumbent | since | state |
|---|---|---|---|
muse.prompt | v0 7809d229 | seed | running |
Variants live in data/state/levers/*.json. Set "paused": true to stop experimenting on a lever - it closes any open experiment and freezes the incumbent, which is how you hold behaviour still for a demo. Editing the incumbent text by hand is supported; the fingerprint is recomputed from the text on load. Nothing here can reach code, gate thresholds, sizing, or the constitution.