Rates verified 27 июля 2026
Калькулятор стоимости Claude Fable 5
Большинство token calculators умножают count на list price и stop. Три вещи двигают real number — здесь смоделированы все.
Что добавляет поверх plain token calculator
First — effort parameter двигает output volume в несколько раз между low и max — expensive model at low effort часто cheap option. Second — share Fable 5 requests hits safety classifier, ответ Opus 4.8, billed differently, invisible error handling. Third — token counts с Opus 4.6 или раньше low ~30%: tokenizer changed at Opus 4.7, every current model inherits. Skip любой из трёх — estimate wrong в direction you did not choose.
Оценка месячных расходов
Token counts из модели до Opus 4.7
Tokenizer с Opus 4.7 даёт ~30% больше токенов на тот же текст. Fable 5, Opus 4.8 и Opus 5 делят его — включайте только если мерили на Opus 4.6 или раньше.
Маршрут через Batch API
50% off в обе стороны, но ответы не real-time.
Claude Fable 5
$238.19
в месяц · $5.81/MTok effective
Claude Opus 5
$122.15
в месяц · $2.98/MTok effective
Claude Opus 5 дешевле на $116.04 в месяц на этой нагрузке.
- Fresh input
- $111.15
- Cache reads
- $25.93
- Output
- $95.00
- Safeguard fallback
- $6.11
Оценки, не quote. Два input смоделированы, не published (1.3x tokenizer, effort estimates) · Rates verified 2026-08-04
Rates DeepSeek, Qwen, GLM и Kimi — vendor-published август 2026, не independently verified. Effort-tier multipliers ниже — Anthropic-only estimates, к competitors как rough proxy.
Три слоя, которые другие не моделируют
Documented behaviour, не speculation — но none in list price.
Tokenizer changed at Opus 4.7 — но не между today's models
Token-counting documentation: Fable 5 и Mythos 5 — tokenizer с Opus 4.7, ~30% больше tokens чем pre-Opus 4.7 models на same text. Migration guide 1.0x–1.35x по content, code/JSON/XML/YAML high end. Opus 4.8 или Opus 5 → Fable 5 — no tokenizer change. Toggle matters только baseline Opus 4.6 или earlier — budget short ~third.
Effort двигает output volume в разы
Effort — API parameter (low, medium, high, xhigh, max — default high), не prompt trick; меняет сколько tokens model spends. На одном published measurement identical prompt — ~1,900 output tokens at low и ~14,400 at max. Anthropic не published cost curve — multipliers interpolated от single anchor, labelled estimates.
Safeguard fallbacks billed — и не free
Cybersecurity или biology classifiers trip — ответ Opus 4.8, HTTP 200 с stop_reason: refusal. Pay for model that served: block before output — common — declined Fable 5 attempt not billed, turn bills Opus 4.8 rates. Mid-stream block — expensive: prefix streamed bills Fable 5 rates. Anthropic reports fallbacks under 5% conversations average; Artificial Analysis 9% on one hard benchmark; prompts narrating reasoning raise rate sharply.
Где calculator честен про not knowing
Token counts estimated from inputs, не measured — exact figure: count_tokens endpoint с real prompt. Effort multipliers interpolated от single public measurement. Cache-write costs excluded — prefix change frequency. Planning range, не invoice.
Снизьте number до оплаты
Caching stable system prompt, effort one notch down на routine work, убрать reasoning-narration instructions — each moves bill больше чем model switch.