Researched 35d ago2021 chat · 32 image · 30 video · 72 voice · 6 music

The 18 LLM leaderboards we'd actually use.

Editor-curated picks for every job — coding agents, CLI workflows, writing, image, voice — benchmark- and pricing-aware, refreshed weekly.

Editorial tiersExcellentStrongSolid
  1. 1 · Eligibility
    A model is eligible for a board only if it's tagged with that use case. Editors pin a handful per board, with exactly one designated Editor's Choice.
  2. 2 · Editorial tiers
    Each pick is bucketed into one of three qualitative tiers — Excellent · Strong · Solid. No decimals, no composite score, just editorial judgment.
  3. 3 · Cross-check
    Picks are opinionated. /best → is the objective benchmark composite for the same capability.