So... which AI are we using today?

Hourly watch enabledLast checked Sep 23, 3:26 PM EDT

YOUR PLAN
Highest individual plan. No extra usage purchases.

CHOOSE YOUR OBJECTIVE

AI CONDITIONS ADVISORY

FIELD BULLETIN
ALL-AROUND USE

ChatGPT

GPT-6 Astra · Work · Codex

Claude

Opus 5.5 · Fable · Cowork / Code

Gemini

3.8 Flash · Deep Think · Spark

Grok

Grok ecosystem · 4.7 in Build

Access unverified
CURRENT PICK: CLAUDE
HIGHEST PAID

CURRENT FIELD REPORT

Moderate confidence

Claude has the edge. For now.

Opus 5.5 gives Claude the strongest new evidence in professional work, with competitive coding results. That makes it our provisional all-around pick. Writing taste and research workflows remain close calls; this is not a clean sweep.

An evidence-based editorial assessment.

LAST CHECKED

Conditions by platform

Overall · Highest paid
01

ChatGPT

ChatGPT Pro · 20x

GPT-6 Astra · Work · Codex

Strong

Astra remains a close contender, especially for demanding research and connected workflows.

Plan access

Highest individual Pro allowance. Includes the broader Work/Codex ecosystem; extra usage purchases are excluded.

02

Claude

CURRENT PICK
Claude Max · 20x

Opus 5.5 · Fable · Cowork / Code

Leading

The newest independent results favor Opus 5.5 for professional work.

Plan access

Highest individual Max allowance. Fable has its own allocation; Opus is the basis of the newest benchmark evidence.

03

Gemini

Google AI Ultra · 20x

3.8 Flash · Deep Think · Spark

Strong

A broad Google workspace and research toolkit. A current four-way app comparison is still missing.

Plan access

The highest Ultra option. Advanced features have country and rollout restrictions; private model previews are excluded.

04

Grok

SuperGrok Heavy

Grok ecosystem · 4.7 in Build

Competitive

4.7 strengthens knowledge work and coding; precise Heavy access remains a coverage gap.

Plan access

Heavy is documented, but precise model and effort entitlements remain partially unverified. No private preview is counted.

Sources behind this overall assessment. Company claims are identified; access and independent results are checked separately.

01
Artificial Analysis Independent test

Opus 5.5: independent evaluation

Opus 5.5 leads the Intelligence Index and professional-work evaluations; terminal coding is level with Astra. These are model-and-harness tests, not a test of every consumer plan.

02
Artificial Analysis Independent test

Benchmarking Grok 4.7

Grok improves in professional work and coding. Its latest coding-agent results trail the leading Claude and Astra systems; higher token use complicates efficiency claims.

03
Anthropic Official

Introducing Claude Opus 5.5

The latest Opus release adds work and communication improvements. Vendor scores and early-access testimonials are treated as launch evidence, not independent replication.

04
Google Official

Gemini 3.8 Flash arrives

Google confirms 3.8 Flash for AI Pro and Ultra. The free and AI Plus plans are not assumed to include the same flagship access.

05
OpenAI Official

ChatGPT plans and model access

Go is the first paid tier; Pro includes Astra reasoning. Free has limited research and desktop Work access. Live rollout notes take precedence over older model rows.

06
Anthropic Official

Claude plan comparison

Pro includes Opus, Research and Claude Code. Max 20x increases capacity. Free includes Sonnet and Haiku, but not Opus or the dedicated Research feature.

07
Google Official

Google AI subscriptions

Free lists 3.6 Flash, limited 3.1 Pro and Deep Research. AI Plus expands usage; Ultra’s largest option adds more capacity and access to advanced features.

08
SpaceXAI Official

Grok subscriptions and usage limits

Paid Grok products share a weekly usage pool. The public FAQ describes SuperGrok and Heavy, but does not fully verify the latest model allocation for each plan.

09
Artificial Analysis X

Opus 5.5 launch evaluation on X

The indexed post announces the new Intelligence Index leader. Full post access was unavailable; the linked evaluation is the substantive evidence, not a second independent vote.

10
Matthew Groff Article

Astra vs Fable for knowledge work

A first-hand account finds both useful and highlights workflow and usage-limit tradeoffs. It predates Opus 5.5, so it provides context rather than a current winner.

11
Google Official

Gemini adds another wave of connected apps

The September 23 rollout adds productivity and creative connections, including Airtable, Linear, Adobe and Webflow. The announcement does not specify identical access for every plan or independently demonstrate task reliability.

HOW WE CALL IT

About this advisory.

We compare tools you can actually use in each plan, across writing, research, coding, and completing tasks. Each category has equal importance in the overall editorial call. A model announcement alone doesn’t win a category.

01 / Check the gate. Verify current model access, restrictions, and product availability for every tier. Private previews and business-only plans are excluded.

02 / Weigh the evidence. Comparable independent tests and published research lead. Official releases establish features. Articles and first-hand posts on X add context; likes are not scores.

03 / Make the call. Every category gets a practical recommendation. “Lean” means our best estimate when comparisons are incomplete or close; “Current pick” has stronger comparative support. We explain the tradeoff, show confidence, and revise the call as evidence changes. We never invent benchmark scores.

Ratings are editorial judgments, not standardized benchmark scores. “Unverified” means missing evidence, not poor performance. Colors identify platforms; they do not imply risk or rank.