We test AI so you know what actually works.

Real-world tests, practical guides and side-by-side model comparisons — for people who actually use this stuff, not for people reading about it.

Scroll

Just published

All tests

Live scoreboard

Who has actually been tested

Every model that has sat a published test, and how long it took over them.

4 models · 12 appearances · bars are rounds entered
  1. Gemini55m 0s average
  2. ChatGPT47m 29s average
  3. Claude229m 40s average
  4. DeepSeek1time not recorded

Nothing has been called yet — every round is still open, and the bars count who turned up rather than who won.

See the full leaderboard

Popular right now

All tests

What are you trying to do with AI?

What are you trying to do with AI?

AI for developers

Claude Code, MCP and AI-assisted coding — the parts that only start to matter once you are past the demo.

Developers

The first technical write-ups are being tested before they are written.

Something not working?

Symptoms in the words people actually use, and the fixes that were tried rather than the ones that sound plausible.

All problems

No fixes published yet. This is where they will live.

MCP, without the fluff

Which server to reach for, where it actually helps, and how it behaves in a real workflow.

Explore MCP

Practical guides

All guides

Nothing here yet — the first ones are being run.

We publish the prompts, the outputs and the measurements.

Published tests
8
Model runs
23
Automated checks
91
Community votes
1
  • The same prompt, word for word.
  • Outputs published unedited.
  • Real completion times, measured in the app.
  • A method you can read and argue with.
Read the methodology