Gemini 3.5 Flash

gemini-3.5-flash · Google

Gemini 3.5 Flash provides sustained frontier-level intelligence optimized for real-world tasks at a higher speed and lower cost. Designed for the agentic era, it excels at sub-agent deployment, multi-step workflows, and long-horizon tasks at scale. This model is particularly effective for rapid agentic loops involving complex coding cycles and iterations.

API Pricing

Input$1.5 / 1M tokens
Output$9 / 1M tokens
Cache read$0.15 / 1M tokens

Specifications

Context1.05M tokens
Max output65.5K tokens
Modalitiestext, image, video, audio, PDF
CapabilitiesThinking, Tool calling, Web search, URL context, Code interpreter, Computer use, File search, Structured outputs, Prompt caching
Endpointschat_completions, gemini_api, claude_api

Frequently asked questions

What is gemini-3.5-flash?

Gemini 3.5 Flash provides sustained frontier-level intelligence optimized for real-world tasks at a higher speed and lower cost. Designed for the agentic era, it excels at sub-agent deployment, multi-step workflows, and long-horizon tasks at scale. This model is particularly effective for rapid agentic loops involving complex coding cycles and iterations.

What is the context length of gemini-3.5-flash?

gemini-3.5-flash has a 1,048,576 token context window. It supports up to 65,536 output tokens.

How much does gemini-3.5-flash cost?

On AIHubMix, gemini-3.5-flash costs $1.5 per million input tokens and $9 per million output tokens. Cached input reads are billed at $0.15 per million tokens.

What modalities does gemini-3.5-flash support?

gemini-3.5-flash accepts text, image, video, audio and PDF input.

What capabilities does gemini-3.5-flash support?

gemini-3.5-flash supports thinking, tool calling, function calling, structured outputs, web search, deep search and long context. Per-protocol parameter support is listed in the capability table on this page.

How do I call gemini-3.5-flash via API?

gemini-3.5-flash is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to gemini-3.5-flash — no other code changes needed.

Who created gemini-3.5-flash?

gemini-3.5-flash is developed by Google. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

More models from Google

gemini-3.6-flash

by Google

Gemini 3.6 Flash provides sustained frontier-level intelligence optimized for real-world…

$1.5/1M in · $7.5/1M out
1,048,576 tokens context

gemini-3.1-flash-lite-image

by Google

Google's newest, most compact, and most cost-effective image generation and editing…

$0.25/1M in · $1.5/1M out

gemini-3.5-flash-lite

by Google

Gemini 3.5 Flash-Lite is a low-latency, cost-effective multimodal model optimized for…

$0.3/1M in · $2.5/1M out
1,048,576 tokens context

gemini-3.5-flash-lite-free

by Google

Gemini 3.5 Flash-Lite free version: Free model resources are limited and provided only…

1,048,576 tokens context

gemini-3.6-flash-free

by Google

Gemini 3.6 Flash free version: fFree model resources are limited and provided only for…

1,000,000 tokens context

gemini-3.1-flash-image

by Google

gemini-3.1-flash-image (Nano Banana 2) features professional-grade visual intelligence…

$0.5/1M in · $3/1M out

Use gemini-3.5-flash via the AIHubMix unified API — one interface for every major LLM.