Skip to main content
Claude API Pricing: $0.30–$5.50 per 1M Input Tokens (2026)
Guides

Claude API Pricing: $0.30–$5.50 per 1M Input Tokens

Claude API pricing in 2026: Opus 5 lists at $5/$25 and Sonnet 5 at $2/$10 per 1M tokens. Unifically serves the same models from $0.30 per 1M input tokens.

Unifically Model Research Team
8 min read

Claude API pricing is billed per token: you pay one rate for the tokens you send and a higher rate for the tokens the model writes back. As of August 18, 2026, Anthropic lists Claude Opus 5 at $5 input / $25 output per million tokens and Claude Sonnet 5 at $2 / $10. Unifically serves the same ten Claude models through OpenAI-compatible and Anthropic-compatible endpoints, with Opus 5 at $1.50 / $7.50 and Haiku 4.5 from $0.30 per million input tokens. This guide covers every model's rate, how token billing works, and what a real request costs.

TL;DR: Anthropic's list prices per 1M tokens (input / output): Claude Fable 5 $10 / $50, Claude Opus 5 $5 / $25, Claude Sonnet 5 $2 / $10, Claude Haiku 4.5 $1 / $5. On Unifically the same models cost $5.50 / $27.50 (Fable 5), $1.50 / $7.50 (Opus 5), $0.90 / $4.50 (Sonnet 5), and $0.30 / $1.50 (Haiku 4.5), on the same /v1/chat/completions and /v1/messages request shapes. New accounts get $0.20 of free balance on signup, enough for several hundred thousand Claude tokens. See live rates on the pricing page.

Claude API pricing per model (2026)

Anthropic API pricing is published on Anthropic's pricing page. These are Anthropic's own list prices per million tokens as of August 18, 2026:

ModelInputOutputCache readCache write (5m)
Claude Fable 5$10.00$50.00$1.00$12.50
Claude Opus 5 / 4.8 / 4.7 / 4.6 / 4.5$5.00$25.00$0.50$6.25
Claude Sonnet 5$2.00$10.00$0.20$2.50
Claude Sonnet 4.6 / 4.5$3.00$15.00$0.30$3.75
Claude Haiku 4.5$1.00$5.00$0.10$1.25

One recent change worth knowing: Sonnet 5 launched at an introductory $2 / $10, with a planned increase to $3 / $15 on September 1, 2026. Anthropic has since made $2 / $10 the standard price, so the increase will not happen. That leaves Sonnet 5 cheaper than the older Sonnet 4.6.

Unifically serves all ten Claude models. Claude model pricing on Unifically, per million tokens:

ModelInputOutputCache readCache writevs Anthropic list
Claude Fable 5$5.50$27.50$0.55$6.87545% lower
Claude Opus 5$1.50$7.50$0.15$1.87670% lower
Claude Opus 4.8$1.50$7.50$0.15$1.87670% lower
Claude Opus 4.7$1.50$7.50$0.15$1.87670% lower
Claude Opus 4.6$1.50$7.50$0.15$1.87670% lower
Claude Opus 4.5$1.50$7.50$0.15$1.87670% lower
Claude Sonnet 5$0.90$4.50$0.09$1.12655% lower
Claude Sonnet 4.6$0.90$4.50$0.09$1.12670% lower
Claude Sonnet 4.5$0.90$4.50$0.09$1.12670% lower
Claude Haiku 4.5$0.30$1.50$0.03$0.37670% lower

Which model to pick:

  • Claude Sonnet 5 is the default. Near-Opus quality on coding and agentic work, a 1M-token context window, and the best price-to-capability ratio in the family.
  • Claude Opus 5 is good for hard agentic coding, multi-file refactors, and long autonomous runs.
  • Claude Haiku 4.5 is the cheapest Claude API option. Use it for classification, extraction, and high-volume pipelines where latency and cost matter more than depth.
  • Claude Fable 5 is Anthropic's most capable model, priced double Opus 5. Reserve it for the hardest reasoning and long-horizon agent work.

How Claude token pricing works (input vs output vs cache)

Every Claude request bills three things:

  1. Input tokens. Everything you send: system prompt, conversation history, documents, tool definitions. Billed at the input rate.
  2. Output tokens. Everything the model writes back. On every current Claude model, output costs 5x the input rate. Reasoning ("thinking") tokens also bill at the output rate, even when the response hides them.
  3. Cache tokens. Prompt caching stores a stable prefix (system prompt, tool list, long documents) so repeat requests reread it at a fraction of the cost. Cache reads bill at 10% of the input rate; writing the cache costs 25% more than plain input for the 5-minute tier.

Two details change the real Claude token cost more than the headline rates:

  • The tokenizer changed. Claude Opus 4.7 and later models (including Sonnet 5 and Fable 5) use a newer tokenizer that produces roughly 30% more tokens for the same text than Sonnet 4.6 and earlier. Compare models by cost per task, not cost per token.
  • Batch requests are half price on Anthropic. Anthropic's Batch API processes async jobs at 50% of standard rates, with most batches finishing within an hour.

Because history is resent on every turn, a long chat's cost is dominated by input tokens. Prompt caching and a frozen system prompt cut that by up to 90% on the cached portion.

Claude API cost examples (what a real request costs)

The Claude API cost of a task follows directly from the token counts. Four worked examples, priced at Anthropic list rates and at Unifically rates:

TaskTokens (in / out)Anthropic listUnifically
One chat turn, Sonnet 51,000 / 500$0.007$0.0032
Summarize a 50-page PDF, Opus 530,000 / 2,000$0.20$0.06
Classify 10,000 support tickets, Haiku 4.53M / 200K$4.00$1.20
Chatbot month, Sonnet 520M / 4M$80.00$36.00

The pattern: single requests cost fractions of a cent on Sonnet and Haiku, and even Opus-tier work stays under a dollar for most one-off jobs. Costs only get interesting at volume, which is where the per-token rate and prompt caching decide your bill.

How to access the Claude API

Two routes:

  1. Anthropic directly. Create a key in the Anthropic Console and call api.anthropic.com at the list prices above.
  2. Unifically. One API key covers Claude alongside GPT, image, video, and audio models, billed pay-per-use from a single balance. No subscription, and credits do not expire. Claude models work on three endpoint formats: OpenAI-compatible /v1/chat/completions and /v1/responses, plus the Anthropic-compatible /v1/messages.

A working request against Claude Sonnet 5:

curl -X POST https://api.unifically.com/v1/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -d '{
    "model": "anthropic/claude-sonnet-5",
    "messages": [
      {"role": "user", "content": "Explain prompt caching in two sentences."}
    ]
  }'

Swap the model id for any row in the table above (anthropic/claude-opus-5, anthropic/claude-haiku-4-5, and so on). Existing OpenAI SDK code works by pointing base_url at https://api.unifically.com/v1, and existing Anthropic SDK code works by swapping the base URL and using /v1/messages.

FAQ

How much does the Claude API cost?

Anthropic lists Claude Opus 5 at $5 input / $25 output per million tokens, Sonnet 5 at $2 / $10, Haiku 4.5 at $1 / $5, and Fable 5 at $10 / $50. Unifically serves the same models at $1.50 / $7.50 (Opus 5), $0.90 / $4.50 (Sonnet 5), $0.30 / $1.50 (Haiku 4.5), and $5.50 / $27.50 (Fable 5). A typical single request costs well under a cent on Sonnet or Haiku.

Is the Claude API free?

No. Both Anthropic and Unifically bill per token. Claude is free to try on Unifically though: new accounts get $0.20 of free balance on signup, which at Haiku 4.5 rates covers over 600,000 input tokens or about 130,000 output tokens. There is no subscription, and unused balance does not expire.

Which Claude model is the cheapest?

Claude Haiku 4.5: $1 / $5 per million tokens at Anthropic list, $0.30 / $1.50 on Unifically. It is built for high-volume, low-latency work like classification and extraction.

What does Claude Opus 5 cost per million tokens?

$5 input / $25 output at Anthropic's list price. On Unifically, Claude Opus 5 costs $1.50 input / $7.50 output per million tokens, with cache reads at $0.15.

Why is Anthropic Claude API pricing different from what I pay on Unifically?

Anthropic Claude API pricing is the first-party list rate. Unifically resells the same models at lower per-token rates: 70% below list on the Opus, older Sonnet, and Haiku models, 55% below on Sonnet 5, and 45% below on Fable 5. The request and response formats are the same, so switching is a base-URL change.

Do cached tokens cost less?

Yes. Cache reads bill at 10% of the input rate on both platforms. On Unifically, Sonnet 5 cache reads are $0.09 per million tokens against $0.90 for fresh input, so a chatbot with a large stable system prompt saves most of its input spend.


Prices verified against Anthropic's pricing page and Unifically's live pricing feed on August 18, 2026. We will update this post when Anthropic changes list prices, when new Claude models land on the platform, or if Unifically rates move. For current numbers at any time, check the pricing page.

Last updated: August 18, 2026

Continue reading

More Blogs