COST MODELLING / VERIFIED SEP 30 2026

    Model the cost before it scales.

    Compare current models from OpenAI, Anthropic, Google, and xAI. Model your workload and cache reads using standard text API rates in USD per 1M tokens.

    Workload inputs

    Estimated / month

    $748.00

    GPT-6.1 Sol

    Cache savings

    $152.00

    vs. fully uncached

    Lowest modeled costMonthly

    Standard text-token estimate. Include billed reasoning tokens in output. Cache reads are modeled; cache writes, storage, tools, search, long-context uplifts, regional premiums, taxes, and discounts are excluded.

    PRICE BOARD

    Public model rates.

    A quick comparison—not a benchmark. Choose models on quality, latency, reliability, and task fit before optimizing token cost.

    LIVE CATALOG / OPENROUTER

    Pull the latest listed rates.

    The endpoint is public, CORS-enabled, and returns per-token prompt, completion, and cache fields. Treat it as an aggregator feed; keep official provider pages as the source of record.

    API docs

    GET https://openrouter.ai/api/v1/models?output_modalities=text&model_authors=openai,anthropic,google,x-ai&sort=pricing-low-to-high

    Ready to query the live catalog.

    Provider / modelInputCachedOutputNoteSource
    OpenAIGPT-6.1 Sol$2.00$0.10$10.00Standard; prompts ≤272K tokensOfficial
    OpenAIGPT-6 Astra$10.00$1.00$50.00Standard; prompts ≤272K tokensOfficial
    OpenAIGPT-6 Luna$0.10$0.01$0.50Standard; prompts ≤272K tokensOfficial
    OpenAIGPT-5.6 Sol$4.00$0.40$20.00Promo through at least Nov 21, 2026; ≤272KOfficial
    AnthropicClaude Fable 5.1$10.00$0.25$50.00Standard; cache writes billed separatelyOfficial
    AnthropicClaude Opus 5.5$4.00$0.20$20.00Standard; cache writes billed separatelyOfficial
    AnthropicClaude Sonnet 5.5$2.00$0.20$10.00Standard; cache writes billed separatelyOfficial
    AnthropicClaude Opus 5$5.00$0.50$25.00Previous generation; standard rateOfficial
    AnthropicClaude Haiku 4.5$1.00$0.10$5.00Standard rateOfficial
    GoogleGemini 3.8 Flash$0.75$0.075$3.75Promo through Dec 31, 2026Official
    GoogleGemini 3.1 Pro Preview$2.00$0.20$12.00Prompts ≤200K tokens; cache storage extraOfficial
    GoogleGemini 3.5 Flash-Lite$0.30$0.03$2.50Standard rateOfficial
    xAIGrok 4.7$2.00$0.50$6.00Standard; prompts ≤200K tokensOfficial

    OPERATING NOTES

    The cheapest token is the one you do not send.

    Cache stable system context and reusable documents.

    Route routine work to smaller models; escalate on uncertainty.

    Cap outputs and stop agent loops when marginal value drops.

    Track cost per successful task—not simply cost per token.