Since our last update we have doubled the number of explainers made to 20k and we hope today's PH launch can bring an even bigger group of users testing Explain. We also have a Chrome Extension and plugins with ChatGPT and Claude to make it even easier to create Explainers anywhere.
Check out the launch and join the conversation here
Generation tools are low-stakes when they're wrong draft an email, get a bad draft, edit it. Judgment tools are different: grading, flagging, scoring. If the output looks confident and it's actually a guess, someone might act on it without knowing that.
Working on letsflw's exam-checker made this concrete objective questions are easy (exact match, no ambiguity), but for open-ended answers we made a call to surface the AI's mark as a suggestion to review, not a final score, because presenting it as settled felt dishonest.