Right tier, every time
Haiku, Sonnet, or Opus, picked by task type and budget — not by habit. Reasoning effort (low / medium / high) is routed independently, with an outcome-learning adapter that fails open to the static pick when it doesn't yet have enough evidence.