# Gemini 3.6 Flash Model id on AIHubMix: `gemini-3.6-flash` Create an API key: https://console.aihubmix.com/ > Gemini 3.6 Flash provides sustained frontier-level intelligence optimized for real-world tasks at a higher speed and lower cost. Designed for the agentic era, it excels at code generation, agentic execution, and spatial reasoning. This model is particularly effective for rapid agentic loops involving complex coding cycles and iterations. - Developer: Google - Context window: 1,048,576 tokens - Input modalities: text, image, audio, video - Capabilities: Thinking, Tool calling, Web search, URL context, Code interpreter, Computer use, File search, Structured outputs, Prompt caching - Release date: 2026-07-21 - Pricing: $1.5/M input tokens, $7.5/M output tokens, $0.15/M cached input ## Endpoints (base URL: https://aihubmix.com) - `POST /v1/chat/completions` — OpenAI Chat Completions (`Authorization: Bearer $AIHUBMIX_API_KEY`) - `POST /gemini/v1beta/models/gemini-3.6-flash:generateContent` — Google Gemini (`x-goog-api-key: $AIHUBMIX_API_KEY`; streaming: `:streamGenerateContent?alt=sse`) - `POST /v1/messages` — Anthropic Messages (`x-api-key: $AIHUBMIX_API_KEY` + `anthropic-version: 2023-06-01`) ## Example ```bash curl -s https://aihubmix.com/v1/chat/completions \ -H "Authorization: Bearer $AIHUBMIX_API_KEY" \ -H "Content-Type: application/json" \ -d '{"model":"gemini-3.6-flash","messages":[{"role":"user","content":"Hello"}]}' ``` ## Response Without `stream`, `/v1/chat/completions` returns a standard Chat Completions object: ```json {"id":"...","object":"chat.completion","model":"gemini-3.6-flash","choices":[{"message":{"role":"assistant","content":"..."}}],"usage":{"prompt_tokens":12,"completion_tokens":24,"total_tokens":36}} ``` With `"stream": true` the response is `text/event-stream`: read each `data:` JSON chunk until `data: [DONE]`. The Messages and Gemini endpoints return their protocols' native response shapes (Anthropic / Google). ## Errors Error responses carry a `tid` (trace id) — include it when contacting support. Reference: https://docs.aihubmix.com/en/FAQs/HTTP-Codes.md - 400 — parameter error; most are passed through from the upstream provider (media: `prompt_missing`, `size_not_supported`, `n_not_within_range`, …) - 401 — missing `Authorization` header, or the key is invalid/expired - 403 — `insufficient_user_quota` (top up at https://console.aihubmix.com/), account suspended, or this key is not allowed to use this model - 429 — rate limited; back off and retry - 503 — no channel can serve the request (check the model id and your access), or the upstream provider is throttling; retry later ## More - Model page: https://aihubmix.com/model/gemini-3.6-flash - Try in browser: https://playground.aihubmix.com/?model=gemini-3.6-flash - Compare with another model (human-facing, side-by-side specs and pricing): https://aihubmix.com/compare/gemini-3.6-flash/{other_model_id} - Full parameter schema (machine-readable, authoritative): https://aihubmix.com/model-data/models/gemini-3.6-flash.80c49b35.json — per-protocol parameters with types, ranges, enums and defaults. Refreshed together with this page; if it ever 404s, re-resolve via `https://aihubmix.com/model-data/index.json` (find this id, fetch its `path`) - Generate runnable code programmatically: npm `@aihubmix/codegen` — the generator behind the Playground's "Get Code" (4 protocols × 7 languages, media endpoints included); the body it builds is the exact wire body the Playground sends, so generated snippets and real requests cannot diverge. `@aihubmix/model-schema` (npm) translates the parameter schema above into codegen input - Site index for agents: https://aihubmix.com/llms.txt · Onboarding: https://aihubmix.com/agents.md --- Canonical version of this document: https://aihubmix.com/model/gemini-3.6-flash/llms.txt