Weave Router 2.0 - Subscription aware coding agent router
by•
Use Claude models in Codex and GPT models in Claude Code on your existing plans, routed to whichever has quota left. Weave Router 2.0 routes each coding agent request to the cheapest model that can get it right. On Terminal-Bench 4.0 and SWE-Atlas it matches GPT-6 Astra at half the cost, 2x faster. Powered by a new classifier that scores task complexity and cache-aware switching that only moves when savings beat the rebuild cost.


Replies
Kick
Congrats on the launch, guys!
Weave Engineering Intelligence
@andrew_roth2 thanks!
Congrats on the launch! This looks really useful. My Question: Is there a free tier to try it out first?
Love this. Once you're running more than one coding agent, subscription-aware routing stops being a nice-to-have. Excited to try how you balance quality and cost across models.
The cache-aware switching is the detail that makes this feel practical, since a cheaper model is not really cheaper if every switch rebuilds a large context. I also like that the model pool stays configurable.
Congrats on the launch! 🙌 🙌 was literally just weighting a local model to cut agent costs, and this is the smarter take.
quick question - when it routes down and gets it wrong, does it catch it med-task and bump back up, or do i only find out after a cheaper model quietly shipped a worse answer? and, is the routing personalized at all, does it learn from my own history or feedback over time, or is it one global classifier for everyone?
neat idea either way 😃