A live spec sheet for the four frontier and open-weight LLMs everyone's choosing between right now. See coding benchmark scores, price per million tokens, and context window side by side - then re-rank instantly by what you actually care about: cheapest, best for coding, or largest context. No sign-up, no ads, just the numbers. Built for developers and teams picking a model for real work, not chasing leaderboard hype.
Hey Product Hunt 👋
I kept seeing the same question in dev communities: "Which AI model should I actually use?" - followed by five different opinions and zero side-by-side numbers.
So I built AI Model Compare: a straight comparison of GPT-5.6 Sol, Claude Opus 4.8, Gemini 3.1 Pro, and DeepSeek V4 Pro on the three things that actually decide which one fits your job - coding benchmark performance, price per million tokens, and context window size. Click a priority (cheapest, best for coding, largest context) and it re-ranks instantly.
No login, no tracking, just data. I'll be updating it as new models ship.
Would love feedback - especially if there's a metric you wish was on there that isn't.