Gemini 3.8 Flash
google/gemini-3.8-flash
Coming soon
Listed, not sold yet: the provider's resale terms are being confirmed.
Google's newest Flash model, aimed at long-horizon software engineering, agents and enterprise workflows. Accepts text, image, video, audio and PDF input.
Price
- Input
- $0.75 per 1M tokens
- Output
- $3.75 per 1M tokens
- Notes
- Paid tier. $0.75 in / $3.75 out (output includes thinking tokens) through 2026-12-31, rises to $1.50 / $7.50 from 2027-01-01.
- Source
- Google pricing , last verified October 8, 2026
No markup on model usage: you pay the provider's list price.
Model
- Context
- 1.05M (1,048,576 tokens)
- Max output
- 65,536 tokens
- Modality
- Text and image
- Input
- text, image, video, audio, file
- Output
- text
- Weights
- Closed
- Released
- September 2, 2026
- Tags
- coding, agents, multimodal, long-context
Data policy
- Prompts go to
- Google (Gemini Developer API)
- Retention
- Paid tier: prompts and responses not used to improve Google products, but logged for a limited period for abuse monitoring, free tier content may be used to improve products
- Zero data retention
- Not offered
- Region
- No region guarantee, data may be stored transiently or cached in any country where Google operates. Google points ZDR and enterprise processing needs to Vertex AI (Gemini Enterprise Agent Platform)
- Policy
- Google data policy
Summarised from the provider's published terms. The provider's own policy is what applies.
Call it with any OpenAI SDK
Available when the model is live
from openai import OpenAI
client = OpenAI(
base_url="https://api.safeguard.sh/v1",
api_key="SAFEGUARD_API_KEY",
)
response = client.chat.completions.create(
model="google/gemini-3.8-flash",
messages=[{"role": "user", "content": "Hello"}],
)
print(response.choices[0].message.content)Base URL https://api.safeguard.sh/v1, model google/gemini-3.8-flash.
Self-healing security runs on Safeguard.
Your first fix PR is minutes away.
No sales call required, even your agent can complete the purchase over MCP.