Google: Gemini 3.5 Flash
Overview
Earlier GA Flash model for routine, high-throughput workloads with multimodal input.
- Provider
- Modality
- Text and image
- Input
- text, image, video, audio, file
- Output
- text
- Max output
- 65,536 tokens
- Weights
- Closed
- Supported parameters
- tools, structured_outputs, reasoning, web_searchProvider's model page, checked .
- Distillable
- NoLicence or terms, checked .
- Tags
- multimodal, long-context
Data policy
- Prompts go to
- Google (Gemini Developer API)
- Retention
- Paid tier: prompts and responses not used to improve Google products, but logged for a limited period for abuse monitoring, free tier content may be used to improve products
- Zero data retention
- Not offered
- Region
- No region guarantee, data may be stored transiently or cached in any country where Google operates. Google points ZDR and enterprise processing needs to Vertex AI (Gemini Enterprise Agent Platform)
- Policy
- Google data policy
Summarised from the provider's published terms. The provider's own policy is what applies.
Where you can run it
How Gemini 3.5 Flash can be deployed, bought and reached through Safeguard.
- Deployment
- Public cloudPrivate cloud
- Procurement
- Not published yet
- Inference regions
- 🇦🇺Australia🇨🇦Canada🇩🇪GermanyEUEuropean Union🇬🇧United Kingdom🇮🇳India🇯🇵Japan🇸🇬Singapore🇺🇸United StatesCloud provider's availability page, checked .
Closed weights stay with the provider, so this model runs in the provider's cloud, never on-prem or air-gapped.
Pricing
| Provider | Input / 1M tokens | Output / 1M tokens | Context |
|---|---|---|---|
| $1.50 | $9 | 1.05M |
- Notes
- Paid tier, output includes thinking tokens.
- Source
- Google pricing , last verified October 8, 2026
No markup on model usage: you pay the provider's list price.
API
OpenAI-compatible. Your API key and access to this model are set up for your workspace when you choose Use with Safeguard.
from openai import OpenAI
client = OpenAI(
base_url="https://api.safeguard.sh/gpt/v1",
api_key="SAFEGUARD_API_KEY",
)
response = client.chat.completions.create(
model="google/gemini-3.5-flash",
messages=[{"role": "user", "content": "Hello"}],
)
print(response.choices[0].message.content)Base URL https://api.safeguard.sh/gpt/v1, model google/gemini-3.5-flash.
Product names and logos are trademarks of their respective owners and identify each model's publisher. Their use does not imply endorsement, sponsorship or affiliation.
Self-healing security runs on Safeguard.
Your first fix PR is minutes away.
No sales call required, even your agent can complete the purchase over MCP.