[AINews] Gemini 4 Argon: GDM’s answer to Astra/Fable, with 1M output
- ID
- 30679
- Status
- summarized
- Published
- 01 Oct 2026, 2:45 PM
- Fetched
- 01 Oct 2026, 2:55 PM
- Provider
- Latent Space
- Category
- developer-ai
- Original URL
- https://www.latent.space/p/ainews-gemini-4-argon-gdms-answer
- Source URL
- https://www.latent.space/feed
Summary
- Score
- 6.0
- Created
- 01 Oct 2026, 2:56 PM
- Tags
- Audience
- developersvibe_codersai_ml_learnersai_agent_usersstartup_founders
What happened
Google DeepMind introduced Gemini 4 Argon, claiming first place on 13 of 19 benchmarks against GPT-6 Astra and Claude Opus 5.5, with a 1M-token output limit via the new Long Decode Continuation API feature. Standard pricing is $4/$20 per 1M input/output tokens, with a 50% introductory discount to $2/$10 and 95% off cached input. Access is limited to government users and trusted cyber defenders in the Fairwind Program, with broader developer, enterprise, and consumer access promised later.
Why it matters
The actionable details are gated: Argon is not generally available, and the 1M output is delivered via Long Decode Continuation, which pauses and resumes responses across calls, while Vals lists 262K max output. Don't re-architect around 1M single-call output yet; if you evaluate it later, compare the $4/$20 standard or $2/$10 intro pricing against your current model, and note cached input is 95% off.
Discussion angle
Is a 1M-token output limit useful if it requires Long Decode Continuation pause/resume across calls, and what would you actually build with it versus chunking outputs?