I made a bunch of promises in the last launch thread. Here's the scorecard.

Last launch ended with 91 comments. I answered every one, and in most of them I said "not yet, that's next."

Six weeks on, here's what actually happened.

Done: asked about Paddle, it now has 21 curated failure patterns instead of a bare spec import.

and both asked who keeps the sandbox honest over time. There's now a fidelity probe that fires the same request at the real provider and at our sandbox and diffs the result. It's already caught four real drifts in our own simulation.

flagged the favicon was still the Next.js default. Fixed.

Half done: , , and , out-of-order events. The permutation fuzzer is built and Paddle has the patterns. What isn't wired yet is the end-to-end path where you describe the ordering bug and it reproduces it on demand. Not going to say it's done when I just checked and it isn't.

Not done: , per-event latency jitter. , ambiguous-ack mid-response. , replay from real production sequences. and , version-range scoping. All still open.

What launches tonight: called it "memory of breakage," which was better than anything I'd written. said the reproduce-fix-rerun loop was the part they were doing by hand.

plugs into Cursor or Claude Code, reproduces the real failure on your own code, and once your agent fixes it, re-runs the check. You get a receipt.

If you were in that thread, this roadmap came from you. What am I still getting wrong?

26 views

Add a comment

Replies

Be the first to comment