BENCHBAITSTORY · MONITORING

AI NEWS, TRACKED AND ROASTED.

google ships gemini 3.8 flash to general availability; one user says three questions bought a five-hour wait

Google announced two Gemini 3.8 variants on September 2, 2026: Gemini 3.8 Flash, documented as generally available under the stable model ID gemini-3.8-flash, and Gemini 3.8 Flash Cyber, restricted to trusted defenders through the Fairwind Program. Google's [91] release documentation for Gemini 3.8 Flash ↗ describes Flash as ready for production use with a default thinking level of medium. The [88] launch announcement ↗ lists Flash across the Gemini API, Google AI Studio, Android Studio, Google Antigravity, and Gemini Enterprise, and for Google AI Pro and Ultra subscribers in the Gemini app, AI Mode, and Gemini in Google Sheets. The [89] Gemini API model page ↗ marks computer use as Preview and lists audio generation, image generation, and the Live API as unsupported.

by the numbers

Documented limits, from the [89] Gemini API model page ↗ and the [90] DeepMind model card ↗:

Input token limit
Value
1,048,576
Output token limit
Value
65,536
Default thinking level
Value
medium
Knowledge cutoff
Value
March 2026, some domains limited to January 2025

Rates from the [92] Agent Platform pricing page ↗, per million tokens:

Global standard
Input through Dec 31, 2026
$0.75
Output through Dec 31, 2026
$3.75
Input from Jan 1, 2027
$1.50
Output from Jan 1, 2027
$7.50
Non-global standard
Input through Dec 31, 2026
$0.825
Output through Dec 31, 2026
$4.125
Input from Jan 1, 2027
$1.65
Output from Jan 1, 2027
$8.25

The same page prices cached input separately.

Google-reported benchmark figures from the [88] launch announcement ↗, which this packet treats as company-reported claims pending independent verification:

  • Flash: 54.9% on HLE-Verified
  • Flash Cyber: over 70% on an internal 20-language vulnerability benchmark
  • Flash Cyber: 47.2% pass@1 on CWE-Bench, against 47.8% for a leading frontier model

what people are saying

The blocks below are attributed paraphrases of each voice's recorded position, drawn from the fact packet and linked to the receipt. They are not verbatim excerpts from the linked pages.

Google, in the [88] launch announcement ↗, on who can reach the cyber variant:

Google states Flash Cyber is available to trusted defenders through the Fairwind Program.

Google, in the same [88] launch announcement ↗, on how its own performance numbers should be read:

Google notes that higher effort can use more tokens, and that the benchmark and real-world performance statements are company-reported claims; chart-level methodology and comparative results require independent review.

Google DeepMind, in the [90] Gemini 3.8 Flash model card ↗, on what the model gets wrong:

The card's known limitations include hallucinations, occasional slowness or timeouts, higher token use at higher effort, and a March 2026 knowledge cutoff with some domains limited to January 2025.

The official Gemini account, in its [93] availability post ↗:

The account said Gemini 3.8 Flash was available starting that day for Pro and Ultra users.

X user @BayasBochi_S, in a [94] post alleging early usage exhaustion ↗:

The post alleges that Gemini 3.8 Flash exhausted daily tokens after three questions and required a five-hour wait.

That last account is an unverified anecdote. It is one user's report, it has not been confirmed, and it does not establish a service-wide limit. The packet records that it aligns directionally with Google's own higher-token-at-higher-effort caveat without verifying anything about product quotas.

claims

  • Confirmed — Google announced Gemini 3.8 Flash and Gemini 3.8 Flash Cyber on September 2, 2026.
  • Confirmed — Gemini 3.8 Flash is documented as GA with the stable model ID gemini-3.8-flash.
  • Confirmed — Gemini 3.8 Flash has a 1,048,576-token input limit and a 65,536-token output limit.
  • Confirmed — Global standard pricing holds its lower input and output rates through December 31, 2026 and rises to higher standard rates on January 1, 2027; the exact per-million figures for both periods are in the pricing table above.
  • Supported — Gemini 3.8 Flash Cyber is available to trusted defenders through the Fairwind Program, not as the general Gemini API model documented for Flash.
  • Supported — Google reports benchmark and internal-evaluation gains, but this packet treats comparative performance as company-reported claims pending independent methodology and result verification.
  • Unverified — The bounded discourse pass produced a usage-cost angle: an X user reported exhausting daily tokens after three questions and waiting five hours, attributing the behavior to higher token consumption. This is an unverified anecdotal report, not a product-limit fact.

receipts

LAST VERIFIED 2026-09-03

WHAT CHANGED

THE TIMELINE

REVISION 1 · LAUNCH · 2026-09-03ROAST / SATIRE

google ships gemini 3.8 flash to general availability; one user says three questions bought a five-hour wait

  • VERIFIEDGoogle announced Gemini 3.8 Flash and Gemini 3.8 Flash Cyber on September 2, 2026.RECEIPTS [88]
  • VERIFIEDGemini 3.8 Flash is documented as GA with the stable model ID gemini-3.8-flash.RECEIPTS [89][91]
  • QUALIFIEDGemini 3.8 Flash Cyber is available to trusted defenders through the Fairwind Program, not as the general Gemini API model documented for Flash.RECEIPTS [88]
  • VERIFIEDGemini 3.8 Flash has a 1,048,576-token input limit and a 65,536-token output limit.RECEIPTS [89][90]
  • VERIFIEDGlobal standard pricing is $0.75 per million input tokens and $3.75 per million output tokens through December 31, 2026, with $1.50 and $7.50 standard rates from January 1, 2027.RECEIPTS [88][92]
  • QUALIFIEDGoogle reports benchmark and internal-evaluation gains, but the packet treats comparative performance as company-reported claims pending independent methodology and result verification.RECEIPTS [88][90]
  • UNVERIFIEDThe bounded discourse pass produced a material usage-cost angle: an X user reported exhausting daily tokens after three questions and waiting five hours, attributing the behavior to higher token consumption; this is an unverified anecdotal report, not a product-limit fact.RECEIPTS [94]

EVERY JOKE HAS RECEIPTS

SOURCES

  1. [88] PRIMARY · GoogleIntroducing Gemini 3.8 Flash and 3.8 Flash Cyberprimary ↗
  2. [89] PRIMARY · Google AI for DevelopersGemini 3.8 Flashprimary ↗
  3. [90] PRIMARY · Google DeepMindGemini 3.8 Flash - Model Cardprimary ↗
  4. [91] PRIMARY · Google AI for DevelopersWhat's new in Gemini 3.8 Flashprimary ↗
  5. [92] PRIMARY · Google CloudAgent Platform Pricingprimary ↗
  6. GeminiApp post announcing Gemini 3.8 Flash availability for Pro and Ultra users

    The official Gemini account said Gemini 3.8 Flash was available starting that day for Pro and Ultra users.

    reaction ↗
  7. User reaction alleging early Gemini 3.8 Flash usage exhaustion

    A bounded X search and direct Bird read returned a user post alleging that Gemini 3.8 Flash exhausted daily tokens after three questions and required a five-hour wait; this is an unverified anecdote, not a confirmed service-wide limit.

    contradiction ↗