Safeguard

Google

Gemini 3.8 Flash

google/gemini-3.8-flash

Coming soon

Listed, not sold yet: the provider's resale terms are being confirmed.

Google's newest Flash model, aimed at long-horizon software engineering, agents and enterprise workflows. Accepts text, image, video, audio and PDF input.

Price

Input
$0.75 per 1M tokens
Output
$3.75 per 1M tokens
Notes
Paid tier. $0.75 in / $3.75 out (output includes thinking tokens) through 2026-12-31, rises to $1.50 / $7.50 from 2027-01-01.
Source
Google pricing , last verified October 8, 2026

No markup on model usage: you pay the provider's list price.

Model

Context
1.05M (1,048,576 tokens)
Max output
65,536 tokens
Modality
Text and image
Input
text, image, video, audio, file
Output
text
Weights
Closed
Released
September 2, 2026
Tags
coding, agents, multimodal, long-context

Data policy

Prompts go to
Google (Gemini Developer API)
Retention
Paid tier: prompts and responses not used to improve Google products, but logged for a limited period for abuse monitoring, free tier content may be used to improve products
Zero data retention
Not offered
Region
No region guarantee, data may be stored transiently or cached in any country where Google operates. Google points ZDR and enterprise processing needs to Vertex AI (Gemini Enterprise Agent Platform)

Summarised from the provider's published terms. The provider's own policy is what applies.

Call it with any OpenAI SDK

Available when the model is live

from openai import OpenAI

client = OpenAI(
    base_url="https://api.safeguard.sh/v1",
    api_key="SAFEGUARD_API_KEY",
)

response = client.chat.completions.create(
    model="google/gemini-3.8-flash",
    messages=[{"role": "user", "content": "Hello"}],
)
print(response.choices[0].message.content)

Base URL https://api.safeguard.sh/v1, model google/gemini-3.8-flash.

Self-healing security runs on Safeguard.

Your first fix PR is minutes away.

No sales call required, even your agent can complete the purchase over MCP.