カタログに戻る
Google·Textnew公式比 30%+ オフ

Gemini 3.6 FlashAPI

Google's newest Flash — thinking-by-default at Flash latency, 1M context, native audio input.

入力料金

$1.05

per 1M tokens

出力料金

$5.25

per 1M tokens

プロンプトキャッシング

安定したプレフィックスに cache_control を渡せば、キャッシュヒットは入力料金のごく一部で課金されます。キャッシュ書き込みは通常の入力レート。

キャッシュ読み出し
$0.21

入力 × 0.2 から自動算出

仕様

モデル ID
gemini-3-6-flash
エンドポイント
POST /v1/chat/completions
カテゴリー
Text
プロバイダー
Google
機能
visionfunctionstreamingthinkinglong-context

ドロップインの OpenAI 互換性

既存の OpenAI SDK を api.kunavo.com/v1 に向け、モデル ID を入れ替えるだけ。ストリーミング変更も SDK 変更も不要です。

ドキュメントを見るGemini API pricing guide

試してみる

/app/keys から KUNAVO_API_KEY を設定し、以下のスニペットを実行してください。

curl
curl https://api.kunavo.com/v1/chat/completions \
  -H "Authorization: Bearer $KUNAVO_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "gemini-3-6-flash",
    "messages": [
      {"role": "user", "content": "Hello, Gemini 3.6 Flash"}
    ],
    "stream": false
  }'

Kunavo と Google 直接利用の違い

Gemini 3.6 Flash のモデル自体も出力も同じです。ゲートウェイが変えるのはアカウント・SDK・請求——以下が率直な比較です。

KunavoGoogle 直接
料金(per 1M tokens)$1.05 / $5.25 (−30%)$1.50 / $7.50
アカウントKunavo アカウント 1 つ、キー発行は 2 分、最低 $5 からチャージGoogle のアカウントと、その請求設定が別途必要
SDKOpenAI SDK のまま——base_url を変え、model に "gemini-3-6-flash" を指定Google 独自の SDK、または提供されていれば OpenAI 互換レイヤー
同じキーで呼べる範囲Claude・Gemini・GPT に加え、画像・動画・音声モデルGoogle 自社モデルのみ
請求Stripe 残高 1 つで従量課金、残高は失効しません。失敗したリクエストは課金されませんプロバイダーごとに別々の請求書
レート制限共有のゲートウェイ枠。アカウント単位の契約上限はありませんGoogle のティア上限(利用実績や契約で引き上げ)

契約上のレート制限、エンタープライズ SLA、または提供元固有の新機能を初日から使いたい場合は Google への直接接続が適しています。キー 1 つ・請求 1 つで、各社のモデルをより安い単価で使いたい場合は Kunavo が向いています。

FAQ

How much does Gemini 3.6 Flash cost?

On Kunavo, Gemini 3.6 Flash is $1.05 per 1M input tokens and $5.25 per 1M output tokens — about 30% under Google's official price. Failed requests are never billed, and it's pay-as-you-go — you only pay for successful calls.

Can I call Gemini 3.6 Flash with the OpenAI SDK?

Yes. Set base_url to https://api.kunavo.com/v1, pass your Kunavo key, and set model to "gemini-3-6-flash". Requests and responses are OpenAI-compatible.

What endpoint does Gemini 3.6 Flash use?

Gemini 3.6 Flash is called via POST /v1/chat/completions on Kunavo's OpenAI-compatible API.

What is Gemini 3.6 Flash good for?

Gemini 3.6 Flash is a text model from Google, with support for vision, function, streaming, thinking, long-context.

What is different from calling Google directly?

The model is the same, at about 30% under Google's list price. What changes is around it: one key and one Stripe balance instead of a Google account with its own billing, the OpenAI SDK with base_url swapped instead of a provider-specific SDK, and the same key also reaching Claude, Gemini, GPT, image, video and audio models. The tradeoff is that gateway capacity is shared — if you need a contractual rate limit or an enterprise SLA on Gemini 3.6 Flash specifically, go direct to Google.

Is Gemini 3.6 Flash cheaper on Kunavo?

Yes — Gemini 3.6 Flash is about 30% under Google's official list price on Kunavo, with no subscription and no minimum monthly spend. You pay per token from a $5 top-up, failed requests are never billed, and the balance never expires.