ガイド一覧へ戻る
Pricing·2026年6月8日·最終更新 2026年9月4日·読了9分

Anthropic Claude API pricing September 2026 (official rates)

Claude is the default for reasoning, coding and agents. Here are the current Claude API prices per model, verified September 2026 — Anthropic's official list alongside Kunavo's rates at 30–60% less — with worked examples and the levers that cut a Claude bill the most.

Last reviewed on .

This is Anthropic Claude API pricing as of September 2026: Anthropic's official per-token list price for every current Claude model, side by side with what the same model costs on Kunavo, plus worked cost examples and the levers that cut a Claude bill the most. Claude is the default choice for reasoning, coding and agentic work, and Kunavo undercuts Anthropic's list on every Claude model but one — about 60% off the Claude 4.x generation, 60% off Opus 5 and 30% off Fable 5 — behind one OpenAI-compatible API. The exception is claude-sonnet-5, which Kunavo prices at Anthropic's list rate rather than under it; the note below says why and what to use instead.

Rates last verified September 4, 2026. Kunavo's per-token prices below are read live from the model catalog, and the “Anthropic list” column tracks Anthropic's published rate — the official source is anthropic.com/pricing (developer detail at docs.anthropic.com).

Official Anthropic API pricing — Claude Opus, Sonnet and Haiku (September 2026)

Rates are per 1M tokens, in USD, as billed on Kunavo. The “Anthropic list” column is Anthropic's official published rate for the same model, taken from anthropic.com/pricing on the verification date above — so this table is the official price list and the discounted one in a single view.

ModelInput / 1MOutput / 1MAnthropic list (in / out)You save
claude-haiku-4-5$0.40$2.00$1.00 / $5.00~60%
claude-sonnet-4-6$1.20$6.00$3.00 / $15.00~60%
claude-sonnet-5$2.00$10.00$2.00 / $10.00none — same as list
claude-opus-4-7$2.00$10.00$5.00 / $25.00~60%
claude-opus-5$2.00$10.00$5.00 / $25.00~60%
claude-opus-5-fast$7.00$35.00$10.00 / $50.00~30%
claude-fable-5$7.00$35.00$10.00 / $50.00~30%

Live rates always show on the pricing page and each model page. Haiku is the cheap workhorse, Sonnet the balanced default, Opus the heavy reasoner — and the Claude 5 family adds Sonnet 5 (near-Opus coding at Sonnet cost), Opus 5 and Fable 5.1, Anthropic's most capable model for frontier reasoning and long-horizon agents (its predecessor Fable 5 is still served, at the same price).

Claude Opus 5 is the value pick of the Claude 5 family. Anthropic lists it at $5.00 / $25.00 per 1M — half of Fable 5's $10.00 / $50.00 — for near-flagship reasoning. On Kunavo it is $2.00 / $10.00, i.e. ~60% under Anthropic's list and the same rate as the older claude-opus-4-7, so moving a workload from Opus 4.7 to Opus 5 is a one-word model change at no extra cost. A frontier-grade run that costs $0.42 on Fable 5 (50K in / 2K out) costs $0.12 on Opus 5.

What about Claude Opus 5 Fast? Anthropic also ships a latency-tuned claude-opus-5-fast — the same reasoning as Opus 5, with faster output, listed at $10.00 / $50.00 per 1M, exactly 2× standard Opus 5. On Kunavo it is $7.00 / $35.00, about 30% under Anthropic's list. Reach for it when time-to-answer is the thing you are optimizing — an interactive coding session, a user waiting on a response. For batch, agentic or long-context work, standard claude-opus-5 returns the same answers for a fraction of the price.

The one exception, stated plainly: Claude Sonnet 5

Anthropic originally launched claude-sonnet-5 at $2.00 / $10.00 per 1M as an introductory rate, with a rise to $3 / $15 scheduled for September 1, 2026. On August 31, 2026 Anthropic cancelled that increase and made $2.00 / $10.00 the permanent standard price. Kunavo matches it exactly — $2.00 / $10.00, the same as going direct. For this one model there is no price saving on Kunavo, and any page of ours still implying a ~30% discount on Sonnet 5 is out of date. What Kunavo still gives you on Sonnet 5 is the part that has nothing to do with per-token price: one key and one balance across Claude, GPT, Gemini and the image/video models, no Anthropic account, no monthly minimum, and no charge for failed calls.

If per-token price is what decides it, the better answer is to move up rather than across. claude-opus-5 costs $2.00 / $10.00 on Kunavo — the same $10.00-per-1M-output price Anthropic charges for Sonnet 5, but with Opus-tier reasoning, and 60% under Opus 5's own $5.00 / $25.00 list price. Claude Fable 5 is a straight ~30% saving from the first call: $7.00 / $35.00 against Anthropic's $10.00 / $50.00.

Anthropic API pricing by model — Haiku, Sonnet, Opus and Fable

The same figures broken out per model family, since Anthropic prices each tier separately and most cost questions are really about which tier a workload belongs in. Every rate is per 1M tokens, input / output.

Claude Haiku 4.5 API pricing

Anthropic lists claude-haiku-4-5 at $1.00 / $5.00 — the cheapest model in the Claude family and the one to reach for on high-volume, low-difficulty work: classification, extraction, routing, support replies. On Kunavo it is $0.40 / $2.00, about 60% under list. At that rate an 800-in / 200-out support ticket costs about $0.00072, so 10,000 tickets a day is roughly $7.20/day. Haiku 4.5 has a 200K context window rather than the 1M the rest of the current family carries.

Claude Sonnet API pricing — Sonnet 4.6 and Sonnet 5

The two Sonnet generations no longer carry the same Anthropic list price. claude-sonnet-4-6 — the everyday production workhorse — lists at $3.00 / $15.00 and is $1.20 / $6.00 on Kunavo, about 60% under list. claude-sonnet-5 brings near-Opus coding and agentic quality at Sonnet cost and lists lower, at $2.00 / $10.00 — the introductory rate Anthropic made permanent on August 31, 2026. Kunavo prices it at $2.00 / $10.00, matching Anthropic exactly rather than undercutting it — see the note below. A 6K-in / 500-out RAG answer is about $0.0102 on Sonnet 4.6.

Claude Opus API pricing — Opus 4.7, Opus 5 and Opus 5 Fast

The Opus tier is the heavy reasoner. Anthropic lists both claude-opus-4-7 and claude-opus-5 at $5.00 / $25.00; Kunavo serves both at $2.00 / $10.00, about 60% under list — the deepest discount in the family, and the reason moving an Opus 4.7 workload to Opus 5 costs nothing extra. The latency-tuned claude-opus-5-fast lists at $10.00 / $50.00, exactly 2× standard Opus 5, and is $7.00 / $35.00 here. A 50K-in / 2K-out long-context analysis on Opus 5 costs about $0.12.

Claude Fable 5 API pricing

claude-fable-5-1 is Anthropic's most capable model — frontier reasoning, long-horizon agents, 1M context — and its most expensive. It succeeded claude-fable-5 on September 1, 2026 in the same tier at the same rate, so the pricing below applies to both. Listed at $10.00 / $50.00 per 1M tokens, double the Opus tier. On Kunavo it is $7.00 / $35.00, about 30% under Anthropic's list, so that saving applies from the first call. The cost math that matters: a 50K-in / 2K-out frontier agent run is about $0.42 on Fable 5 against $0.12 for the same run on Opus 5 — so reserve Fable 5 for work where the Opus tier has actually been tried and fallen short. Thinking is always on for this model, and reasoning tokens bill at the output rate, which makes the output line the one to watch.

How Anthropic API pricing works per token

Anthropic prices the Claude API per token, quoted per 1M tokens and billed on the exact count — there is no per-request fee, no minimum and no subscription. You pay for input tokens (system prompt, context, messages) and output tokens (the completion). Output is billed at roughly 5× the input rate across the family, so the largest lever on cost is how much the model writes. The second lever is prompt caching: stable, repeated context can be cached and billed at a fraction of the normal input rate.

A token is about 4 characters of English, so 1M tokens ≈ 750,000 words. At Kunavo's $6.00 per 1M output on Sonnet 4.6, a 500-token answer costs about $0.00300 in output. Reasoning models add a wrinkle: hidden reasoning tokens are billed at the output rate even though you never see them, which is why an output cap saves more than it looks like it should.

Worked cost examples

Real numbers at Kunavo's rates:

WorkloadTokens (in / out)ModelCost
Support ticket800 / 200Haiku 4.5$0.00072
RAG answer6,000 / 500Sonnet 4.6$0.0102
Coding assistant call10,000 / 1,500Sonnet 4.6$0.021
Coding agent turn10,000 / 1,500Sonnet 5$0.037
Long-context analysis50,000 / 2,000Opus 4.7$0.12
Frontier agent run50,000 / 2,000Fable 5$0.42

So 10,000 support tickets a day on Haiku is about $7.20/day. The math, runnable:

claude_cost.py
# Kunavo Claude rates (USD per 1M tokens): (input, output)
RATES = {
    "claude-haiku-4-5":  (0.40, 2.00),
    "claude-sonnet-4-6": (1.20, 6.00),
    "claude-sonnet-5":   (2.00, 10.00),
    "claude-opus-4-7":   (2.00, 10.00),
    "claude-opus-5":     (2.00, 10.00),
    "claude-fable-5":    (7.00, 35.00),
}

def cost(model: str, in_tokens: int, out_tokens: int) -> float:
    i, o = RATES[model]
    return in_tokens / 1_000_000 * i + out_tokens / 1_000_000 * o

print(cost("claude-haiku-4-5", 800, 200))      # support ticket -> $0.00072
print(cost("claude-sonnet-4-6", 6_000, 500))   # RAG answer     -> $0.0102
print(cost("claude-sonnet-5", 10_000, 1_500))  # coding agent   -> $0.037
print(cost("claude-opus-4-7", 50_000, 2_000))  # long analysis  -> $0.12
print(cost("claude-opus-5", 50_000, 2_000))    # Opus 5 analysis-> $0.12
print(cost("claude-fable-5", 50_000, 2_000))   # frontier run   -> $0.42

Prompt caching — the biggest single saving

If your system prompt or retrieved context is large and stable, prompt caching bills cached input at 10% of the input rate. For a long, reused system prompt that is a 60–90% cut on the input line. It is one extra field on the request — see the Anthropic prompt caching deep dive for the exact pattern and the Messages API reference. What that looks like on a bill, including the three ways a gateway can quietly break your hit rate, is in Claude prompt caching; the Claude token cost calculator has a cached-tokens field so you can price it against your own traffic.

Kunavo pricing and Stripe billing

No subscription, no Anthropic Console billing setup. Top up a balance (Stripe or local payment methods) and calls draw down at the rates above. Pay-as-you-go from a $10 minimum top-up, the balance never expires, and larger top-ups carry bonus credit. The same balance covers Claude, Gemini, GPT, image, video and audio models — one wallet, one invoice.

Claude vs Gemini: the cost decision

Match the model to the job rather than defaulting to the most capable one:

NeedPickKunavo in / out per 1M
Cheapest viablegemini-2-5-flash$0.09 / $0.75
Cheap + Anthropic qualityclaude-haiku-4-5$0.40 / $2.00
Balanced reasoningclaude-sonnet-4-6$1.20 / $6.00
Near-Opus coding at Sonnet costclaude-sonnet-5$2.00 / $10.00
Hardest reasoning on a budgetclaude-opus-4-7$2.00 / $10.00
Frontier reasoning & agentsclaude-fable-5$7.00 / $35.00

See the Gemini and GPT pricing guides for the other providers, and the cost optimization guide for the difficulty-routing pattern in code.

Claude Code API pricing

Claude Code — Anthropic's agentic coding CLI — runs on these same Claude models, so its cost is ordinary Claude API token cost: the files and tool output it reads count as input tokens, and the edits and explanations it writes count as output tokens. Sonnet 4.6 ($1.20 / $6.00 per 1M) is the value driver, Sonnet 5 ($2.00 / $10.00) the newest near-Opus coder, and Opus 4.7 ($2.00 / $10.00) or the frontier Fable 5 ($7.00 / $35.00) handle the hardest refactors. Agentic sessions re-read context each turn, so they are token-heavy — which makes prompt caching the biggest lever, billing cached context at 10% of the input rate. Point Claude Code (or another agentic coding tool) at Kunavo's Anthropic-compatible Messages endpoint to pay the rates above — 30–60% under Anthropic's list depending on the model — instead of list price.

Claude Code reads ANTHROPIC_BASE_URL natively, so that redirect is three environment variables and no extra software — the Claude Code setup guide has the exact variables and the credential mistake behind most 401s. If you are still deciding between paying per token and paying for a plan, Claude Code pricing puts the Pro/Max plan fees next to the per-token rates and works through the break-even, is Claude Code free covers what is and isn't included, and getting an API key for Claude Code covers where the key goes. Not installed yet? Install Claude Code has the per-OS command.

FAQ

What is Anthropic's official API pricing in September 2026?

As of September 4, 2026, Anthropic's official list price per 1M tokens is: Claude Haiku 4.5 $1.00 input / $5.00 output, Claude Sonnet 4.6 $3.00 / $15.00, Claude Sonnet 5 $2.00 / $10.00, Claude Opus 4.7 $5.00 / $25.00, Claude Opus 5 $5.00 / $25.00, and Claude Fable 5 $10.00 / $50.00. Anthropic publishes these at platform.claude.com/docs/en/about-claude/pricing. The same models on Kunavo are $0.40/$2.00, $1.20/$6.00, $2.00/$10.00, $2.00/$10.00, $2.00/$10.00 and $7.00/$35.00 respectively — 30–60% under list depending on the model. One exception worth stating plainly: Claude Sonnet 5. Anthropic made its $2.00 / $10.00 rate the permanent standard price on August 31, 2026 (the increase to $3 / $15 that had been scheduled for September 1 was cancelled), and Kunavo matches that price exactly at $2.00 / $10.00 rather than beating it. For that one model there is no saving against Anthropic — the reason to route it through Kunavo is the shared key and balance, not price. If price is the deciding factor, claude-opus-5 costs the same $2.00 / $10.00 on Kunavo and gives Opus-tier reasoning for it.

Is this the official Anthropic price list?

No — Anthropic's own page at anthropic.com/pricing is the official source, and it is linked at the top of this guide. This page reproduces those official list prices, states the date they were verified, and puts Kunavo's rate for the same model next to each one. Kunavo is an independent gateway that resells these models, not Anthropic.

What is the official Anthropic API pricing for Claude Opus, Sonnet and Haiku?

Anthropic's official list price per 1M tokens, by tier: Claude Haiku 4.5 is $1.00 input / $5.00 output; Claude Sonnet 4.6 and Claude Sonnet 5 are both $3.00 / $15.00; Claude Opus 4.7 and Claude Opus 5 are both $5.00 / $25.00; Claude Opus 5 Fast and Claude Fable 5 are $10.00 / $50.00. Kunavo serves the same models at $0.40/$2.00 (Haiku 4.5), $1.20/$6.00 (Sonnet 4.6), $2.00/$10.00 (Sonnet 5), $2.00/$10.00 (Opus 4.7 and Opus 5), $7.00/$35.00 (Opus 5 Fast) and $7.00/$35.00 (Fable 5) — between 30% and 60% under list depending on the model. Verified September 4, 2026 against anthropic.com/pricing.

How much does the Claude Haiku 4.5 API cost?

Anthropic lists Claude Haiku 4.5 at $1.00 input / $5.00 output per 1M tokens — the cheapest model in the Claude family. On Kunavo it is $0.40 / $2.00 per 1M, about 60% under list. In practical terms an 800-input / 200-output support ticket costs roughly $0.00072, so 10,000 of them a day is about $7.20. Haiku 4.5 carries a 200K context window rather than the 1M window on the current Sonnet, Opus and Fable models. Model-by-model totals for your own volumes are in the token cost calculator.

Where are the official Anthropic API pricing docs?

Anthropic publishes its official API price list at anthropic.com/pricing, and the per-model developer documentation — context windows, output limits and model IDs — at docs.anthropic.com under the model overview page. Both are linked at the top of this guide. This page is not an Anthropic property: it reproduces those official list prices with the date they were verified (September 4, 2026) and puts Kunavo's rate for the same model beside each one.

How much does the Anthropic API cost per token in 2026?

Anthropic quotes per 1M tokens and bills the exact token count — no per-request fee, no minimum, no subscription. A token is roughly 4 characters of English, so 1M tokens is about 750,000 words. In practical terms, at Kunavo's rates: a support ticket on Haiku 4.5 costs about $0.0007, a RAG answer on Sonnet 4.6 about $0.01, and a long-context analysis on Opus 4.7 about $0.12. Input and output are priced separately, with output about 5× input.

Is the Claude API free?

The Claude API is not free on an ongoing basis. Anthropic's pricing FAQ states that new users receive a small amount of free credits to test the API; Anthropic does not publish the size of that grant, and there is no free tier beyond it, so usage is billed per token once the initial credits are spent. Kunavo does not give new accounts free credit either: it is pay-as-you-go from a $10 minimum top-up, with per-token rates 30–60% below Anthropic's list price depending on the model, a balance that never expires, and no billing for failed calls. The one model where Kunavo is not below Anthropic is Claude Sonnet 5, which is priced at Anthropic's list rate rather than under it.

How much does Claude Sonnet cost?

On Kunavo, Claude Sonnet 5 is $2.00 per 1M input tokens and $10.00 per 1M output — exactly Anthropic's $2.00 / $10.00 list price, not under it. Sonnet 5 is the one Claude model Kunavo does not discount. The previous-gen Claude Sonnet 4.6 is $1.20 / $6.00, about 60% under its $3.00 / $15.00 list price. Claude Haiku 4.5 is $0.40 / $2.00 and Opus 4.7 is $2.00 / $10.00, both about 60% under list.

How much does the Claude Opus 5 API cost?

Anthropic lists Claude Opus 5 at $5.00 input / $25.00 output per 1M tokens — half the price of Claude Fable 5 ($10.00 / $50.00) for near-flagship reasoning. On Kunavo it is $2.00 / $10.00 per 1M, about 60% under Anthropic's list, pay-as-you-go with no subscription. That is the same rate Kunavo charges for the older Claude Opus 4.7, so upgrading is a one-word model change at no extra cost. The latency-tuned claude-opus-5-fast, listed at $10.00 / $50.00 (2× standard Opus 5), is $7.00 / $35.00 on Kunavo.

What is the difference between Claude Opus 5 and Claude Opus 5 Fast?

They are the same model with different latency and price. Standard claude-opus-5 lists at $5.00 / $25.00 per 1M tokens. claude-opus-5-fast delivers the same reasoning quality with faster output and lists at $10.00 / $50.00 per 1M — exactly 2×. Pick Fast only when time-to-answer matters more than cost, for example an interactive coding session; for batch, agentic or long-context work, standard Opus 5 gives identical answers for far less money. Kunavo serves both: claude-opus-5 at $2.00 / $10.00 and claude-opus-5-fast at $7.00 / $35.00, about 60% and 30% under Anthropic's list respectively.

How much does the Claude Fable 5 API cost?

Claude Fable 5 lists at $10.00 input / $50.00 output per 1M tokens. On Kunavo it is $7.00 / $35.00 per 1M, about 30% under Anthropic's list, pay-as-you-go with no subscription. A 50K-input / 2K-output frontier agent run costs about $0.42 at those rates. Anthropic released Claude Fable 5.1 on September 1, 2026 as its successor and Anthropic's most capable model; it lists at the same $10.00 / $50.00 and costs the same $7.00 / $35.00 on Kunavo, so switching to claude-fable-5-1 changes the model string and nothing on the bill.

How do I reduce Claude API cost?

Prompt caching first (cached input at 10% of rate), then tier down to Haiku for simple tasks and cap output. Details in the cost optimization guide.

How much does the Claude Code API cost?

Claude Code runs on the standard Claude models, so it's billed as ordinary Claude API token cost — files and tool output count as input tokens, edits and explanations as output tokens. On Kunavo that's Sonnet 4.6 at $1.20 / $6.00 per 1M, Sonnet 5 at $2.00 / $10.00, and Opus 4.7 at $2.00 / $10.00 — 30–60% under Anthropic depending on the model, except Sonnet 5, which is priced at Anthropic's list rate rather than under it. Agentic sessions are token-heavy, so prompt caching (cached input at 10% of the input rate) is the biggest saving.

Does Kunavo support the Anthropic Messages API?

Yes — Claude is available through both the native Messages API (/v1/messages) and the OpenAI-compatible /v1/chat/completions endpoint. To start, see how to get a Claude API key.