Gemini 2.5 FlashAPI
Previous-gen Gemini Flash — extreme value.
入力料金
$0.09
per 1M tokens
出力料金
$0.75
per 1M tokens
プロンプトキャッシング
安定したプレフィックスに cache_control を渡せば、キャッシュヒットは入力料金のごく一部で課金されます。キャッシュ書き込みは通常の入力レート。
入力 × 0.2 から自動算出
仕様
- モデル ID
gemini-2-5-flash- エンドポイント
POST /v1/chat/completions- カテゴリー
- Text
- プロバイダー
- 機能
- visionfunctionstreaming
ドロップインの OpenAI 互換性
既存の OpenAI SDK を api.kunavo.com/v1 に向け、モデル ID を入れ替えるだけ。ストリーミング変更も SDK 変更も不要です。
ドキュメントを見るGemini API pricing guide試してみる
/app/keys から KUNAVO_API_KEY を設定し、以下のスニペットを実行してください。
curl https://api.kunavo.com/v1/chat/completions \
-H "Authorization: Bearer $KUNAVO_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "gemini-2-5-flash",
"messages": [
{"role": "user", "content": "Hello, Gemini 2.5 Flash"}
],
"stream": false
}'Kunavo と Google 直接利用の違い
Gemini 2.5 Flash のモデル自体も出力も同じです。ゲートウェイが変えるのはアカウント・SDK・請求——以下が率直な比較です。
| Kunavo | Google 直接 | |
|---|---|---|
| 料金(per 1M tokens) | $0.09 / $0.75 (−70%) | $0.30 / $2.50 |
| アカウント | Kunavo アカウント 1 つ、キー発行は 2 分、最低 $10 からチャージ | Google のアカウントと、その請求設定が別途必要 |
| SDK | OpenAI SDK のまま——base_url を変え、model に "gemini-2-5-flash" を指定 | Google 独自の SDK、または提供されていれば OpenAI 互換レイヤー |
| 同じキーで呼べる範囲 | Claude・Gemini・GPT に加え、画像・動画・音声モデル | Google 自社モデルのみ |
| 請求 | Stripe 残高 1 つで従量課金、残高は失効しません。失敗したリクエストは課金されません | プロバイダーごとに別々の請求書 |
| レート制限 | 共有のゲートウェイ枠。アカウント単位の契約上限はありません | Google のティア上限(利用実績や契約で引き上げ) |
契約上のレート制限、エンタープライズ SLA、または提供元固有の新機能を初日から使いたい場合は Google への直接接続が適しています。キー 1 つ・請求 1 つで、各社のモデルをより安い単価で使いたい場合は Kunavo が向いています。
FAQ
How much does Gemini 2.5 Flash cost?
On Kunavo, Gemini 2.5 Flash is $0.09 per 1M input tokens and $0.75 per 1M output tokens — about 70% under Google's official price. Failed requests are never billed, and it's pay-as-you-go — you only pay for successful calls.
Can I call Gemini 2.5 Flash with the OpenAI SDK?
Yes. Set base_url to https://api.kunavo.com/v1, pass your Kunavo key, and set model to "gemini-2-5-flash". Requests and responses are OpenAI-compatible.
What endpoint does Gemini 2.5 Flash use?
Gemini 2.5 Flash is called via POST /v1/chat/completions on Kunavo's OpenAI-compatible API.
What is Gemini 2.5 Flash good for?
Gemini 2.5 Flash is a text model from Google, with support for vision, function, streaming.
What is different from calling Google directly?
The model is the same, at about 70% under Google's list price. What changes is around it: one key and one Stripe balance instead of a Google account with its own billing, the OpenAI SDK with base_url swapped instead of a provider-specific SDK, and the same key also reaching Claude, Gemini, GPT, image, video and audio models. The tradeoff is that gateway capacity is shared — if you need a contractual rate limit or an enterprise SLA on Gemini 2.5 Flash specifically, go direct to Google.
Is Gemini 2.5 Flash cheaper on Kunavo?
Yes — Gemini 2.5 Flash is about 70% under Google's official list price on Kunavo, with no subscription and no minimum monthly spend. You pay per token from a $10 top-up, failed requests are never billed, and the balance never expires.