Claude Sonnet 5 - AI that plans, acts, and gets work done
by•
Our most agentic Sonnet yet, with top-tier intelligence for coding and everyday professional work.
Replies
Best
The agentic framing is the interesting shift here, more than the benchmark bump. Curious about the guardrail side though - when it's chaining terminal and browser calls on its own, is there a built-in checkpoint before anything irreversible (deleting files, submitting a form, sending a message) or is that entirely left to whoever builds the wrapper app? That's the part that decides whether I'd actually trust it running unattended overnight.
Report
Already a fan! The research that it can do and the precision of the facts is what I like.
Report
Switched from ChatGPT to Claude a few months back and never looked back. Built my entire product on top of it, and what actually mattered was reasoning through the hard decisions when stuff broke. Raw code generation is the easy part. Curious how much of the "agentic" jump in Sonnet 5 shows up there specifically, not just clean-run coding benchmarks. That's where the real work lives.
Report
Switched from ChatGPT to Claude a few months back and never looked back. Built my entire product on top of it, and what actually mattered was reasoning through the hard decisions when stuff broke. Raw code generation is the best part. The way claude handles long context without losing nuance genuinely impresses me
Replies
The agentic framing is the interesting shift here, more than the benchmark bump. Curious about the guardrail side though - when it's chaining terminal and browser calls on its own, is there a built-in checkpoint before anything irreversible (deleting files, submitting a form, sending a message) or is that entirely left to whoever builds the wrapper app? That's the part that decides whether I'd actually trust it running unattended overnight.
Already a fan! The research that it can do and the precision of the facts is what I like.
Switched from ChatGPT to Claude a few months back and never looked back. Built my entire product on top of it, and what actually mattered was reasoning through the hard decisions when stuff broke. Raw code generation is the easy part. Curious how much of the "agentic" jump in Sonnet 5 shows up there specifically, not just clean-run coding benchmarks. That's where the real work lives.
Switched from ChatGPT to Claude a few months back and never looked back. Built my entire product on top of it, and what actually mattered was reasoning through the hard decisions when stuff broke. Raw code generation is the best part. The way claude handles long context without losing nuance genuinely impresses me