Gemini 3.5 Flash

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution loops, supporting text, image, video, audio, and PDF inputs. Defaults to medium thinking effort for faster and more cost-efficient responses, with full support for thinking levels (minimal, low, medium, high) for fine-grained cost/performance trade-offs.

Speech-to-textGoogle1M tokens$1.6351 / min

google/gemini-3.5-flash

Context
1M tokens
Completion cap
65,536
Tools
Yes
JSON
Yes
Released
2026-05-19

Call it from Geek Hub

Same OpenAI SDK. Change the base URL and the model id.

curl https://api.geekhub.mx/v1/chat/completions \
  -H "Authorization: Bearer $GEEKHUB_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "google/gemini-3.5-flash",
    "messages": [{"role": "user", "content": "Hola"}]
  }'
Get an API key

Frequently asked questions

What is Gemini 3.5 Flash?
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution loops, supporting text, image, video, audio, and PDF inputs. Defaults to medium thinking effort for faster and more cost-efficient responses, with full support for thinking levels (minimal, low, medium, high) for fine-grained cost/performance trade-offs. Gemini 3.5 Flash runs on the Geek Hub API (OpenAI-compatible). Model id: google/gemini-3.5-flash.
Is Gemini 3.5 Flash free?
No. $1.6351 per minute on Geek Hub.
What is the context length of Gemini 3.5 Flash?
Gemini 3.5 Flash has a 1M tokens context window. It supports up to 65,536 completion tokens.
When was Gemini 3.5 Flash released?
Gemini 3.5 Flash was released on 2026-05-19.

More models from Google