返回模型市场
Google·Textnew比官方便宜 30%+

Gemini 3.6 FlashAPI

Google's newest Flash — thinking-by-default at Flash latency, 1M context, native audio input.

输入价格

$1.05

per 1M tokens

输出价格

$5.25

per 1M tokens

Prompt 缓存

给稳定 prefix 加 cache_control,命中按 input 价的一小部分计费。cache_write 按 input 原价计费。

Cache read
$0.21

按 input × 0.2 自动派生

参数

模型 ID
gemini-3-6-flash
端点
POST /v1/chat/completions
类别
Text
供应商
Google
能力
visionfunctionstreamingthinkinglong-context

无缝兼容 OpenAI SDK

把 OpenAI SDK 的 base_url 指向 api.kunavo.com/v1,换 model id 即可。流式、工具调用、SDK 行为完全一致。

查看文档Gemini API 定价指南

立即试用

在 /app/keys 创建 API key,复制为 KUNAVO_API_KEY,然后任选下面一段代码运行。

curl
curl https://api.kunavo.com/v1/chat/completions \
  -H "Authorization: Bearer $KUNAVO_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "gemini-3-6-flash",
    "messages": [
      {"role": "user", "content": "Hello, Gemini 3.6 Flash"}
    ],
    "stream": false
  }'

Kunavo 与直连 Google 的区别

同一个 Gemini 3.6 Flash,同样的输出。网关改变的是账号、SDK 和账单——下面是如实对照。

Kunavo直连 Google
价格(per 1M tokens)$1.05 / $5.25 (−30%)$1.50 / $7.50
账号一个 Kunavo 账号,2 分钟拿到 Key,最低充值 $5需要 Google 账号,并单独配置其计费
SDK继续用 OpenAI SDK——改 base_url,model 填 "gemini-3-6-flash"Google 官方 SDK,或其提供的 OpenAI 兼容层
同一个 Key 还能调用Claude、Gemini、GPT 以及图像、视频、音频模型仅 Google 自家模型
计费一个 Stripe 余额,按量付费,余额不过期;失败请求不计费每家厂商一张独立账单
速率限制共享网关容量,无按账号的合约级限额Google 自家档位限额,可凭用量或合约提升

如果你需要合约级速率限制、企业 SLA,或首日就要用上厂商独有功能,直连 Google 更合适。如果你想要一个 Key、一张账单,并在所有厂商上拿到更低的单价,就用 Kunavo。

FAQ

How much does Gemini 3.6 Flash cost?

On Kunavo, Gemini 3.6 Flash is $1.05 per 1M input tokens and $5.25 per 1M output tokens — about 30% under Google's official price. Failed requests are never billed, and it's pay-as-you-go — you only pay for successful calls.

Can I call Gemini 3.6 Flash with the OpenAI SDK?

Yes. Set base_url to https://api.kunavo.com/v1, pass your Kunavo key, and set model to "gemini-3-6-flash". Requests and responses are OpenAI-compatible.

What endpoint does Gemini 3.6 Flash use?

Gemini 3.6 Flash is called via POST /v1/chat/completions on Kunavo's OpenAI-compatible API.

What is Gemini 3.6 Flash good for?

Gemini 3.6 Flash is a text model from Google, with support for vision, function, streaming, thinking, long-context.

What is different from calling Google directly?

The model is the same, at about 30% under Google's list price. What changes is around it: one key and one Stripe balance instead of a Google account with its own billing, the OpenAI SDK with base_url swapped instead of a provider-specific SDK, and the same key also reaching Claude, Gemini, GPT, image, video and audio models. The tradeoff is that gateway capacity is shared — if you need a contractual rate limit or an enterprise SLA on Gemini 3.6 Flash specifically, go direct to Google.

Is Gemini 3.6 Flash cheaper on Kunavo?

Yes — Gemini 3.6 Flash is about 30% under Google's official list price on Kunavo, with no subscription and no minimum monthly spend. You pay per token from a $5 top-up, failed requests are never billed, and the balance never expires.