Gemini 3.6 FlashAPI
Google's newest Flash — thinking-by-default at Flash latency, 1M context, native audio input.
Harga input
$1.05
per 1M tokens
Harga output
$5.25
per 1M tokens
Prompt caching
Kirim cache_control pada prefiks stabil — hit cache ditagih sepersekian dari harga input. Tulis cache selalu dengan harga input normal.
Dihitung otomatis sebagai input × 0.2
Spesifikasi
- ID model
gemini-3-6-flash- Endpoint
POST /v1/chat/completions- Kategori
- Text
- Penyedia
- Kemampuan
- visionfunctionstreamingthinkinglong-context
Kompatibilitas drop-in OpenAI
Arahkan SDK OpenAI kamu ke api.kunavo.com/v1 dan ganti ID model. Tanpa perubahan streaming, tanpa SDK baru.
Lihat dokumenGemini API pricing guideCoba sendiri
Atur KUNAVO_API_KEY dari /app/keys dan jalankan snippet di bawah.
curl https://api.kunavo.com/v1/chat/completions \
-H "Authorization: Bearer $KUNAVO_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "gemini-3-6-flash",
"messages": [
{"role": "user", "content": "Hello, Gemini 3.6 Flash"}
],
"stream": false
}'Kunavo vs memanggil Google langsung
Model Gemini 3.6 Flash dan responsnya sama persis. Yang diubah gateway adalah akun, SDK, dan tagihan — ini perbandingan jujurnya.
| Kunavo | Google langsung | |
|---|---|---|
| Harga (per 1M tokens) | $1.05 / $5.25 (−30%) | $1.50 / $7.50 |
| Akun | Satu akun Kunavo, kunci dalam 2 menit, isi ulang minimum $5 | Akun Google beserta pengaturan penagihannya sendiri |
| SDK | Tetap pakai OpenAI SDK — ganti base_url, isi model dengan "gemini-3-6-flash" | SDK milik Google, atau lapisan kompatibel-OpenAI bila tersedia |
| Kunci yang sama juga menjangkau | Claude, Gemini, GPT plus model gambar, video, dan audio | Hanya katalog Google sendiri |
| Penagihan | Satu saldo Stripe, bayar sesuai pemakaian, tidak pernah hangus; permintaan gagal tidak ditagih | Faktur terpisah untuk tiap penyedia |
| Batas laju | Kapasitas gateway bersama, tanpa batas kontraktual per akun | Batas tier Google sendiri, dinaikkan lewat riwayat pemakaian atau kontrak |
Pilih Google langsung bila Anda butuh batas laju kontraktual, SLA enterprise, atau fitur khusus penyedia sejak hari pertama. Kunavo cocok bila Anda ingin satu kunci, satu tagihan, dan harga per token lebih rendah di semua penyedia.
FAQ
How much does Gemini 3.6 Flash cost?
On Kunavo, Gemini 3.6 Flash is $1.05 per 1M input tokens and $5.25 per 1M output tokens — about 30% under Google's official price. Failed requests are never billed, and it's pay-as-you-go — you only pay for successful calls.
Can I call Gemini 3.6 Flash with the OpenAI SDK?
Yes. Set base_url to https://api.kunavo.com/v1, pass your Kunavo key, and set model to "gemini-3-6-flash". Requests and responses are OpenAI-compatible.
What endpoint does Gemini 3.6 Flash use?
Gemini 3.6 Flash is called via POST /v1/chat/completions on Kunavo's OpenAI-compatible API.
What is Gemini 3.6 Flash good for?
Gemini 3.6 Flash is a text model from Google, with support for vision, function, streaming, thinking, long-context.
What is different from calling Google directly?
The model is the same, at about 30% under Google's list price. What changes is around it: one key and one Stripe balance instead of a Google account with its own billing, the OpenAI SDK with base_url swapped instead of a provider-specific SDK, and the same key also reaching Claude, Gemini, GPT, image, video and audio models. The tradeoff is that gateway capacity is shared — if you need a contractual rate limit or an enterprise SLA on Gemini 3.6 Flash specifically, go direct to Google.
Is Gemini 3.6 Flash cheaper on Kunavo?
Yes — Gemini 3.6 Flash is about 30% under Google's official list price on Kunavo, with no subscription and no minimum monthly spend. You pay per token from a $5 top-up, failed requests are never billed, and the balance never expires.