AI Weekly Malaysia

Back to items Summaries

[AINews] Gemini 4 Argon: GDM’s answer to Astra/Fable, with 1M output

ID
30679
Status
summarized
Published
01 Oct 2026, 2:45 PM
Fetched
01 Oct 2026, 2:55 PM
Provider
Latent Space
Category
developer-ai
Original URL
https://www.latent.space/p/ainews-gemini-4-argon-gdms-answer
Source URL
https://www.latent.space/feed

Summary

Score
6.0
Created
01 Oct 2026, 2:56 PM
Tags
Audience
developersvibe_codersai_ml_learnersai_agent_usersstartup_founders

What happened

Google DeepMind introduced Gemini 4 Argon, claiming first place on 13 of 19 benchmarks against GPT-6 Astra and Claude Opus 5.5, with a 1M-token output limit via the new Long Decode Continuation API feature. Standard pricing is $4/$20 per 1M input/output tokens, with a 50% introductory discount to $2/$10 and 95% off cached input. Access is limited to government users and trusted cyber defenders in the Fairwind Program, with broader developer, enterprise, and consumer access promised later.

Why it matters

The actionable details are gated: Argon is not generally available, and the 1M output is delivered via Long Decode Continuation, which pauses and resumes responses across calls, while Vals lists 262K max output. Don't re-architect around 1M single-call output yet; if you evaluate it later, compare the $4/$20 standard or $2/$10 intro pricing against your current model, and note cached input is 95% off.

Discussion angle

Is a 1M-token output limit useful if it requires Long Decode Continuation pause/resume across calls, and what would you actually build with it versus chunking outputs?

Top