Zurück zum Katalog
Google·Textnew30 %+ unter offiziell

Gemini 3.6 FlashAPI

Google's newest Flash — thinking-by-default at Flash latency, 1M context, native audio input.

Input-Preis

$1.05

per 1M tokens

Output-Preis

$5.25

per 1M tokens

Prompt-Caching

Gib cache_control für stabile Präfixe — Cache-Hits werden zu einem Bruchteil des Input-Preises abgerechnet. Cache-Writes laufen zum normalen Input-Tarif.

Cache-Read
$0.21

Input × 0.2 (automatisch berechnet)

Spezifikationen

Modell-ID
gemini-3-6-flash
Endpoint
POST /v1/chat/completions
Kategorie
Text
Anbieter
Google
Capabilities
visionfunctionstreamingthinkinglong-context

Drop-in OpenAI-Kompatibilität

Richte dein bestehendes OpenAI SDK auf api.kunavo.com/v1 und ersetze die Modell-ID. Keine Streaming-Änderungen, kein neues SDK.

Doku ansehenGemini API pricing guide

Selbst ausprobieren

Setze KUNAVO_API_KEY aus /app/keys und führe das Snippet aus.

curl
curl https://api.kunavo.com/v1/chat/completions \
  -H "Authorization: Bearer $KUNAVO_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "gemini-3-6-flash",
    "messages": [
      {"role": "user", "content": "Hello, Gemini 3.6 Flash"}
    ],
    "stream": false
  }'

Kunavo vs. Direktzugriff auf Google

Dasselbe Gemini 3.6 Flash, dieselben Antworten. Ein Gateway ändert Konto, SDK und Rechnung — hier der ehrliche Vergleich.

KunavoGoogle direkt
Preis (per 1M tokens)$1.05 / $5.25 (−30%)$1.50 / $7.50
KontoEin Kunavo-Konto, Key in 2 Minuten, Mindestaufladung $5Ein Google-Konto samt eigener Abrechnungseinrichtung
SDKOpenAI SDK behalten — base_url ändern, model auf "gemini-3-6-flash" setzenGoogle-eigenes SDK oder deren OpenAI-kompatible Schicht, sofern vorhanden
Derselbe Key erreicht außerdemClaude, Gemini, GPT sowie Bild-, Video- und AudiomodelleNur der Katalog von Google
AbrechnungEin Stripe-Guthaben, Pay-as-you-go, verfällt nie; fehlgeschlagene Anfragen sind kostenlosEine eigene Rechnung pro Anbieter
RatenlimitsGeteilte Gateway-Kapazität, kein vertragliches Limit pro KontoTarifstufen-Limits von Google, anhebbar über Nutzungshistorie oder Vertrag

Gehen Sie direkt zu Google, wenn Sie ein vertraglich zugesichertes Ratenlimit, ein Enterprise-SLA oder ein anbieterexklusives Feature ab Tag eins brauchen. Kunavo passt, wenn Sie einen Key, eine Rechnung und über alle Anbieter hinweg einen niedrigeren Tokenpreis wollen.

FAQ

How much does Gemini 3.6 Flash cost?

On Kunavo, Gemini 3.6 Flash is $1.05 per 1M input tokens and $5.25 per 1M output tokens — about 30% under Google's official price. Failed requests are never billed, and it's pay-as-you-go — you only pay for successful calls.

Can I call Gemini 3.6 Flash with the OpenAI SDK?

Yes. Set base_url to https://api.kunavo.com/v1, pass your Kunavo key, and set model to "gemini-3-6-flash". Requests and responses are OpenAI-compatible.

What endpoint does Gemini 3.6 Flash use?

Gemini 3.6 Flash is called via POST /v1/chat/completions on Kunavo's OpenAI-compatible API.

What is Gemini 3.6 Flash good for?

Gemini 3.6 Flash is a text model from Google, with support for vision, function, streaming, thinking, long-context.

What is different from calling Google directly?

The model is the same, at about 30% under Google's list price. What changes is around it: one key and one Stripe balance instead of a Google account with its own billing, the OpenAI SDK with base_url swapped instead of a provider-specific SDK, and the same key also reaching Claude, Gemini, GPT, image, video and audio models. The tradeoff is that gateway capacity is shared — if you need a contractual rate limit or an enterprise SLA on Gemini 3.6 Flash specifically, go direct to Google.

Is Gemini 3.6 Flash cheaper on Kunavo?

Yes — Gemini 3.6 Flash is about 30% under Google's official list price on Kunavo, with no subscription and no minimum monthly spend. You pay per token from a $5 top-up, failed requests are never billed, and the balance never expires.