Gemini 3.6 FlashAPI
Google's newest Flash — thinking-by-default at Flash latency, 1M context, native audio input.
Input-Preis
$1.05
per 1M tokens
Output-Preis
$5.25
per 1M tokens
Prompt-Caching
Gib cache_control für stabile Präfixe — Cache-Hits werden zu einem Bruchteil des Input-Preises abgerechnet. Cache-Writes laufen zum normalen Input-Tarif.
Input × 0.2 (automatisch berechnet)
Spezifikationen
- Modell-ID
gemini-3-6-flash- Endpoint
POST /v1/chat/completions- Kategorie
- Text
- Anbieter
- Capabilities
- visionfunctionstreamingthinkinglong-context
Drop-in OpenAI-Kompatibilität
Richte dein bestehendes OpenAI SDK auf api.kunavo.com/v1 und ersetze die Modell-ID. Keine Streaming-Änderungen, kein neues SDK.
Doku ansehenGemini API pricing guideSelbst ausprobieren
Setze KUNAVO_API_KEY aus /app/keys und führe das Snippet aus.
curl https://api.kunavo.com/v1/chat/completions \
-H "Authorization: Bearer $KUNAVO_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "gemini-3-6-flash",
"messages": [
{"role": "user", "content": "Hello, Gemini 3.6 Flash"}
],
"stream": false
}'Kunavo vs. Direktzugriff auf Google
Dasselbe Gemini 3.6 Flash, dieselben Antworten. Ein Gateway ändert Konto, SDK und Rechnung — hier der ehrliche Vergleich.
| Kunavo | Google direkt | |
|---|---|---|
| Preis (per 1M tokens) | $1.05 / $5.25 (−30%) | $1.50 / $7.50 |
| Konto | Ein Kunavo-Konto, Key in 2 Minuten, Mindestaufladung $5 | Ein Google-Konto samt eigener Abrechnungseinrichtung |
| SDK | OpenAI SDK behalten — base_url ändern, model auf "gemini-3-6-flash" setzen | Google-eigenes SDK oder deren OpenAI-kompatible Schicht, sofern vorhanden |
| Derselbe Key erreicht außerdem | Claude, Gemini, GPT sowie Bild-, Video- und Audiomodelle | Nur der Katalog von Google |
| Abrechnung | Ein Stripe-Guthaben, Pay-as-you-go, verfällt nie; fehlgeschlagene Anfragen sind kostenlos | Eine eigene Rechnung pro Anbieter |
| Ratenlimits | Geteilte Gateway-Kapazität, kein vertragliches Limit pro Konto | Tarifstufen-Limits von Google, anhebbar über Nutzungshistorie oder Vertrag |
Gehen Sie direkt zu Google, wenn Sie ein vertraglich zugesichertes Ratenlimit, ein Enterprise-SLA oder ein anbieterexklusives Feature ab Tag eins brauchen. Kunavo passt, wenn Sie einen Key, eine Rechnung und über alle Anbieter hinweg einen niedrigeren Tokenpreis wollen.
FAQ
How much does Gemini 3.6 Flash cost?
On Kunavo, Gemini 3.6 Flash is $1.05 per 1M input tokens and $5.25 per 1M output tokens — about 30% under Google's official price. Failed requests are never billed, and it's pay-as-you-go — you only pay for successful calls.
Can I call Gemini 3.6 Flash with the OpenAI SDK?
Yes. Set base_url to https://api.kunavo.com/v1, pass your Kunavo key, and set model to "gemini-3-6-flash". Requests and responses are OpenAI-compatible.
What endpoint does Gemini 3.6 Flash use?
Gemini 3.6 Flash is called via POST /v1/chat/completions on Kunavo's OpenAI-compatible API.
What is Gemini 3.6 Flash good for?
Gemini 3.6 Flash is a text model from Google, with support for vision, function, streaming, thinking, long-context.
What is different from calling Google directly?
The model is the same, at about 30% under Google's list price. What changes is around it: one key and one Stripe balance instead of a Google account with its own billing, the OpenAI SDK with base_url swapped instead of a provider-specific SDK, and the same key also reaching Claude, Gemini, GPT, image, video and audio models. The tradeoff is that gateway capacity is shared — if you need a contractual rate limit or an enterprise SLA on Gemini 3.6 Flash specifically, go direct to Google.
Is Gemini 3.6 Flash cheaper on Kunavo?
Yes — Gemini 3.6 Flash is about 30% under Google's official list price on Kunavo, with no subscription and no minimum monthly spend. You pay per token from a $5 top-up, failed requests are never billed, and the balance never expires.