One prompt can now drive more of your workMore support for defending infrastructure and open sourceClaude becomes easier to use safely under updated rulesOne prompt window can handle more of the workGPT-6 Intelligent UI makes conversations visual and interactiveClaude Haiku 5.5 delivers low-cost, high-performance AICreate visual, interactive answers from a simple chatRun high-volume tasks with a cheaper fast modelGPT-6 Sol and Luna let you match speed to budgetEasier access to advanced Claude models for security workRun image, audio, and video search on-device with one modelAtlassian integration makes company knowledge easier to useAnthropic expands safer access to advanced cyber featuresEnable text watermarking via API for EU complianceText provenance becomes easier to pilot in the EUClaude training becomes easier for enterprise teamsAnthropic invests in workforce training for enterprise adoptionEnterprise adoption and training get easierGemini 4 Argon is built for long professional tasksUse Astra-level performance affordably in daily workOne prompt can now drive more of your workMore support for defending infrastructure and open sourceClaude becomes easier to use safely under updated rulesOne prompt window can handle more of the workGPT-6 Intelligent UI makes conversations visual and interactiveClaude Haiku 5.5 delivers low-cost, high-performance AICreate visual, interactive answers from a simple chatRun high-volume tasks with a cheaper fast modelGPT-6 Sol and Luna let you match speed to budgetEasier access to advanced Claude models for security workRun image, audio, and video search on-device with one modelAtlassian integration makes company knowledge easier to useAnthropic expands safer access to advanced cyber featuresEnable text watermarking via API for EU complianceText provenance becomes easier to pilot in the EUClaude training becomes easier for enterprise teamsAnthropic invests in workforce training for enterprise adoptionEnterprise adoption and training get easierGemini 4 Argon is built for long professional tasksUse Astra-level performance affordably in daily work
Official sources only. Rumors, leaks, and get-rich schemes are excluded.
← Back to top
AI BriefingGooglePricing & Plans00:00

AI summarized from verified sources

Flex & Priority Inference Tiers for Gemini API

Run background jobs at 50% cost, saving budgets significantly.

SOURCE CHECK

3 sources

VERIFIED

Sources

Key Points

  • 1Flex: Cost-opt, lower priority
  • 2Priority: Low-latency, high priority
  • 3For Gemini 2.5/3.1 models
  • 4Available now

Google added Flex (cost-optimized) and Priority (latency-optimized) tiers to Gemini API. Flex offers up to 50% savings for tolerant workloads; Priority prioritizes traffic. Balances cost, speed, reliability for devs.

Key point

Google updated the Gemini Developer API pricing page with clearer input, output, and caching rates by model. It also makes free and paid tiers easier to compare for prototype and production planning.

Impact

Separate prototype and production costs more clearly. Key checks: Pricing clarified by model / Input, output, and cache listed / Free vs paid tiers are clearer.

h
hayami

Stay on top of OpenAI, Google & Anthropic updates. An AI digest for business professionals.

Source Policy

We use only official sources. Each article links to the original announcement so you can verify it yourself.

© 2026 hayami. All rights reserved.