milk routes each prompt between a fast primary agent and a deep escalation agent — keeping the full conversation in sync across both. Start cheap. Go deep when you need it.
curl -fsSL https://raw.githubusercontent.com/scoutme/milk/main/install.sh | sh
Each prompt is classified and sent to the right agent without you changing tools.
When escalation fires, the primary conversation is reformatted as context — the escalation agent orients itself without a separate setup step.
A Percept store survives across sessions; key facts are reinforced, decay, and promote to long-term memory over time.
Monitors agent output for repeating patterns, warns in the status bar, and auto-interrupts when an agent gets stuck.
Run the same scenarios against different agents and compare LLM-judged quality, tokens, cache efficiency, and latency side-by-side.