Launching today

Qwen3.8-Max
A New Bar for Coding and Cowork
4 followers
A New Bar for Coding and Cowork
4 followers
📢 Meet Qwen3.8-Max, our most capable model to date. A new bar for coding and work at 2.4T parameters: 🔵 Autonomous coding: 10+ days of self-evolving development, from empty folder to production without hand-holding, complete project trace in the GitHub:qwen-code-dev-bot/oh-my-cli. 🔵 Long-horizon mastery: System-level autonomous planning with closed-loop adaptive learning, driving 500+ turns of chip design optimization and 365 days of e-commerce strategy. Start building with Qwen3.8-Max! 🚀




The number that decides whether anyone can actually use the 1M window is the cache, not the $2 input. Full context is $2 a call before you get a single token back, and $0.25 cached is the only reason a long agent loop pencils out. So the spec I want published is the implicit cache TTL and what invalidates it. Implicit is lovely until you're forecasting a month of spend and can't explain why the hit rate moved.