The leaderboard
Rankings
Ranked by average score across every battle a tool has fought — not raw win count, since a tool that's only fought once shouldn't outrank one that keeps showing up and keeps winning. A "win" means it took the verdict in that head-to-head; see how scoring works.
| # | Tool | Record | Avg score | Categories |
|---|---|---|---|---|
| 1 | Google AI Mode | 1W–0L | 9.2/10 | Search |
| 2 | ChatGPT | 1W–0L | 8.6/10 | Writing |
| 3 | OpenAI TTS | 1W–0L | 8.5/10 | Voice |
| 4 | FLUX | 1W–0L | 8.4/10 | Image |
| 5 | GitHub Copilot | 1W–0L | 8.4/10 | Coding |
| 6 | Suno | 1W–0L | 8.4/10 | Music |
| 7 | DALL·E 3 | 0W–1L | 8.3/10 | Image |
| 8 | ChatGPT Search | 0W–1L | 8.2/10 | Search |
| 9 | Cursor | 0W–1L | 8.0/10 | Coding |
| 10 | Runway | 1W–0L | 7.9/10 | Video |
| 11 | Perplexity | 0W–1L | 7.9/10 | Search |
| 12 | Windsurf | 0W–1L | 7.9/10 | Coding |
| 13 | Claude | 0W–1L | 7.8/10 | Writing |
| 14 | ElevenLabs | 0W–1L | 7.8/10 | Voice |
| 15 | Pika | 0W–1L | 7.8/10 | Video |
| 16 | Udio | 0W–1L | 7.8/10 | Music |
| 17 | Notion AI | 1W–0L | 7.7/10 | Writing |
| 18 | Midjourney | 0W–1L | 7.4/10 | Image |
| 19 | Obsidian + AI plugins | 0W–1L | 7.3/10 | Writing |
| 20 | Sora | 0W–1L | 5.8/10 | Video |
Scores are editorial judgment, not lab benchmarks or aggregated user ratings — no invented review counts or usage figures went into this table. Updated as new battles publish.