Tests
Tests are hands-on verdicts only. They prove whether a tool, feature, model, prompt or workflow holds up against a real task, with what worked, what broke and what would change the view.
Claude vs ChatGPT vs Copilot: fixing a broken budget spreadsheet
I gave Claude, ChatGPT and Microsoft Copilot the same ten-tab programme budget and one loose prompt. All three fixed the error I planted. Only one left a record I would hand to a finance director.
Claude Fable 5 vs GPT-5.6: Project Handover Review
GPT-5.6 produced the better steering-ready handover. Claude Fable 5 was stronger at challenging disputed approvals and conflicting evidence.
ChatGPT vs Copilot in PowerPoint: the steering committee test
I gave ChatGPT and Copilot ten files of project mess and a locked committee deck. One edited the template perfectly. One signed my name to it. See the scores.
Claude Sonnet 5 test: one call, two clean outputs
I gave Claude Sonnet 5 one messy transcript and asked for two clean outputs in one prompt. Read what worked, what broke, and whether it holds up on day one.
Get each test when it lands.
Guides, templates, builds, tests and opinions for practical AI at work, at home and for play.