Three flagship AI models launched within nine days of each other: Claude Opus 5.5 (September 22), Claude Sonnet 5.5 (September 28), and Gemini 4 Argon (September 30). Only two of them are usable today. Gemini 4 Argon is restricted to a small group of cybersecurity testers through Google’s Fairwind Program, with no public release date β Google itself frames the launch as an attempt to “rejoin the frontier conversation” after falling behind.
Nine days, three companies, three “flagship” announcements. If you only read the headlines, it looks like three new options to choose from this week. Read past the announcement posts and it’s really two options you can use today and one you’re being asked to take on faith.
Three Launches, Three Different Access Levels
| Model | Announced | Who can use it right now |
|---|---|---|
| Claude Opus 5.5 | Sep 22, 2026 | Everyone β live on the Claude API, Amazon Bedrock, Google Cloud, and Microsoft Foundry |
| Claude Sonnet 5.5 | Sep 28, 2026 | Everyone β same platforms as Opus 5.5 |
| Gemini 4 Argon | Sep 30, 2026 | Only trusted cybersecurity testers in Google’s Fairwind Program; no public date for paid API or Google AI Ultra access |
That gap matters more than any benchmark number. A model you can call today beats a model with a better slide deck and no access.
What Each One Actually Claims
Claude Opus 5.5: Anthropic says it matches Fable 5.1-level output on most work while costing roughly 40% less to run than Opus 5, with a 1-million-token context window. It’s priced at $4 input / $20 output per million tokens on current trackers.
Claude Sonnet 5.5: Anthropic kept pricing identical to Sonnet 5 ($2/$10 per million tokens) and says it generates output more than 30% faster while costing up to 30% less per task in its own testing β a rare case of a model upgrade with zero price increase.
Gemini 4 Argon: Google reports 77.9% on DeepSWE v1.1 (long-horizon software engineering), 91.7% on LVBench (long video understanding), and a #1 ranking on Zapier’s AutomationBench at 51.3%. The output limit jumped from 64,000 to 1 million tokens. On CWE-bench v1, a vulnerability-patching benchmark, Argon ties for first at 68% β notable since Google is initially releasing it specifically to help cybersecurity teams find and patch flaws.
All three sets of numbers are self-reported at launch. Anthropic’s Opus and Sonnet 5.5 claims have started showing up in independent trackers within days, the normal pattern we’ve seen with every other launch-week benchmark on this site. Argon’s numbers haven’t had that independent check yet, partly because almost nobody outside Google can run it.
Google Is Explicit About Catching Up
Google’s own framing around Argon is unusually candid: outlets covering the launch described it as Google’s attempt to rejoin the frontier conversation after being surpassed by OpenAI’s GPT-6 Astra and Anthropic’s Fable 5. Bloomberg separately reported internal concern at Google about whether Argon is actually competitive with Anthropic and OpenAI’s latest. That’s a different posture than most launch announcements, which rarely acknowledge falling behind at all.
The restricted rollout fits that picture. Google says it’s still tuning safeguards against cyber and CBRN misuse before wider release, with internal and external red teams testing the model and monitors watching its reasoning in real time. That’s a reasonable caution for a model this capable β but it also means “Gemini 4 Argon” is, for now, a model you read about rather than one you can compare side by side with Opus 5.5 yourself.
What This Means If You’re Choosing Today
- Need a model right now: it’s Opus 5.5 or Sonnet 5.5, not Argon β Argon isn’t an option yet regardless of its benchmark claims. See our existing Claude review and Claude vs. ChatGPT comparison for how the broader Claude lineup fits different tasks.
- Heavier reasoning or Fable-tier work on a budget: Opus 5.5’s 40% cost cut versus Opus 5 is the headline here, similar in spirit to Fable 5.1’s own access and pricing story β check current Claude plan access before assuming Fable-level output comes free with any tier.
- Everyday coding and document work: Sonnet 5.5 at unchanged pricing is the easy upgrade with no new cost to weigh.
- Waiting on Argon: there’s no cost to waiting. Independent benchmarks and real public pricing won’t exist until Google actually ships it past the Fairwind testers, so there’s nothing to evaluate yet beyond Google’s own claims.
Frequently Asked Questions
Can I use Gemini 4 Argon right now?
No, unless you’re part of Google’s Fairwind Program for trusted cybersecurity testers. Google has given no public release date for paid API customers, Google AI Ultra subscribers, or general availability.
Is Claude Opus 5.5 better than Claude Opus 5?
Anthropic says Opus 5.5 matches Fable 5.1-level output on most work while costing about 40% less to run than Opus 5, with a 1-million-token context window. These are Anthropic’s own launch claims; independent verification typically follows within days to weeks of release.
Did Claude Sonnet 5.5 get more expensive than Sonnet 5?
No. Anthropic kept the same $2/$10 per million token pricing while claiming over 30% faster output and up to 30% lower cost per task in its own testing.
Why is Google restricting access to Gemini 4 Argon?
Google says it’s still strengthening safeguards against cyber and CBRN (chemical, biological, radiological, nuclear) misuse before wider release, and is using red teams and real-time monitoring during this limited rollout to cybersecurity-focused testers.
Should I wait for Gemini 4 Argon instead of using Claude or ChatGPT now?
Not if you need a model today β Argon has no public access path yet. If your work genuinely depends on comparing it against current options, that comparison isn’t possible until Google ships a broader release with real pricing and independently checked benchmarks.
Shurah is the founder of AI Tools Daily, tracking pricing, licensing and policy changes across AI tools so readers can make decisions without wading through marketing claims themselves.