Routing each coding request to the cheapest model that can still get it right is the right instinct - most people just default to the strongest model for everything and eat the cost. Being able to use Claude in Codex and GPT in Claude Code depending on which plan still has quota is a real practical unlock, not just a cost gimmick.