Safeguard

Google: Gemini 3.8 Flash

google/gemini-3.8-flashAvailable via Safeguard
Public cloudPrivate cloudEU🇺🇸

Use with Safeguard

Gemini 3.8 Flash is available via Safeguard. Leave your work email and we will enable it for your workspace.

We will also add you to our newsletter; unsubscribe any time.

Released
September 2, 2026
Context
1,048,576 tokens
Input
$0.75 / 1M tokens
Output
$3.75 / 1M tokens

Overview

Google's newest Flash model, aimed at long-horizon software engineering, agents and enterprise workflows. Accepts text, image, video, audio and PDF input.

Provider
Google
Modality
Text and image
Input
text, image, video, audio, file
Output
text
Max output
65,536 tokens
Weights
Closed
Supported parameters
tools, structured_outputs, reasoning, web_searchProvider's model page, checked .
Distillable
NoLicence or terms, checked .
Tags
coding, agents, multimodal, long-context

Data policy

Prompts go to
Google (Gemini Developer API)
Retention
Paid tier: prompts and responses not used to improve Google products, but logged for a limited period for abuse monitoring, free tier content may be used to improve products
Zero data retention
Not offered
Region
No region guarantee, data may be stored transiently or cached in any country where Google operates. Google points ZDR and enterprise processing needs to Vertex AI (Gemini Enterprise Agent Platform)

Summarised from the provider's published terms. The provider's own policy is what applies.

Where you can run it

How Gemini 3.8 Flash can be deployed, bought and reached through Safeguard.

Deployment
Public cloudPrivate cloud
Procurement
Not published yet
Inference regions
EUEuropean Union🇺🇸United StatesCloud provider's availability page, checked .

Closed weights stay with the provider, so this model runs in the provider's cloud, never on-prem or air-gapped.

Pricing

ProviderInput / 1M tokensOutput / 1M tokensContext
Google$0.75$3.751.05M
Notes
Paid tier. $0.75 in / $3.75 out (output includes thinking tokens) through 2026-12-31, rises to $1.50 / $7.50 from 2027-01-01.
Source
Google pricing , last verified October 8, 2026

No markup on model usage: you pay the provider's list price.

API

OpenAI-compatible. Your API key and access to this model are set up for your workspace when you choose Use with Safeguard.

from openai import OpenAI

client = OpenAI(
    base_url="https://api.safeguard.sh/gpt/v1",
    api_key="SAFEGUARD_API_KEY",
)

response = client.chat.completions.create(
    model="google/gemini-3.8-flash",
    messages=[{"role": "user", "content": "Hello"}],
)
print(response.choices[0].message.content)

Base URL https://api.safeguard.sh/gpt/v1, model google/gemini-3.8-flash.

Product names and logos are trademarks of their respective owners and identify each model's publisher. Their use does not imply endorsement, sponsorship or affiliation.

Self-healing security runs on Safeguard.

Your first fix PR is minutes away.

No sales call required, even your agent can complete the purchase over MCP.