We compare tools you can actually use in each plan, across writing, research, coding, and completing tasks. Each category has equal importance in the overall editorial call. A model announcement alone doesn’t win a category.
01 / Check the gate. Verify current model access, restrictions, and product availability for every tier. Private previews and business-only plans are excluded.
02 / Weigh the evidence. Comparable independent tests and published research lead. Official releases establish features. Articles and first-hand posts on X add context; likes are not scores.
03 / Make the call. Every category gets a practical recommendation. “Lean” means our best estimate when comparisons are incomplete or close; “Current pick” has stronger comparative support. We explain the tradeoff, show confidence, and revise the call as evidence changes. We never invent benchmark scores.