Doubleword
    Pricing

    Same intelligence. Up to 99% cheaper.

    One API. Open-weight models. Pick your delivery window — Async or Batch — and pay only for what you use.

    Per-token pricing

    Same Intelligence. Fraction of the price.

    Cost to process 1 billion tokens in + 1 billion tokens out at comparable intelligence.

    Model
    Anthropic
    $60K
    OpenAI (Astra)
    $60K
    Anthropic (Fable 5.1)
    $60K
    OpenAI
    $24K
    Industry Average
    $18K
    Doubleword (Async)
    $13.4K
    $0$15K$30K$45K$60K

    Intelligence via Artificial Analysis Index v4.3 · Hover any bar for full pricing details · Want access to a model you don't see here — just ask us!

    No credit card required · No minimum spend · Pay only for tokens used

    Try Kimi-K3

    We can offer custom pricing for bulk discounts, large workloads, and dedicated deployments — reach out to [email protected].

    Three speeds. One API.

    Pick the delivery window that fits your workflow. All tiers use the same OpenAI-compatible API.

    Real Time

    Iterate on prompts with real-time responses. Full price, zero wait.

    Async Inference

    up to 50% off RT

    Background agents that need results fast. High throughput inference.

    Batch (~24 hours)

    up to 80% off RT

    Big batch jobs where cost matters most. Deepest discounts.

    Full pricing

    Model-by-model breakdown

    Model SLA Input $/MTok Output $/MTok Cost / 1B in+out vs Big Token
    DeepSeek-V4.1-FlashNew Async $0.12 $0.48 $600 — Try DeepSeek-V4.1-Flash
    ↳ Batch $0.08 $0.30 $380 —
    DeepSeek-V4-Pro-0813New Async $0.99 $2.97 $4.0K 25% cheaper Try DeepSeek-V4-Pro-0813
    ↳ Batch $0.66 $1.98 $2.6K 50% cheaper
    MiMo-V2.5-ProNew Async $0.33 $0.65 $980 81% cheaper Try MiMo-V2.5-Pro
    ↳ Batch $0.22 $0.44 $660 87% cheaper
    InklingNew Async $0.90 $3.00 $3.9K 78% cheaper Try Inkling
    ↳ Batch $0.60 $2.00 $2.6K 86% cheaper
    Qwen3.8-27B Async $0.35 $2.25 $2.6K 91% cheaper Try Qwen3.8-27B
    ↳ Batch $0.25 $1.50 $1.8K 94% cheaper
    DeepSeek-V4-Pro Async $0.98 $1.95 $2.9K 90% cheaper Try DeepSeek-V4-Pro
    ↳ Batch $0.65 $1.30 $2.0K 94% cheaper
    DeepSeek-V4-Flash-0731New Async $0.07 $0.14 $210 96% cheaper Try DeepSeek-V4-Flash-0731
    ↳ Batch $0.05 $0.09 $140 97% cheaper
    DeepSeek-V4-Flash Async $0.07 $0.14 $210 96% cheaper Try DeepSeek-V4-Flash
    ↳ Batch $0.05 $0.09 $140 97% cheaper
    Kimi-K3New Async $2.15 $11.25 $13.4K 78% cheaper Try Kimi-K3
    ↳ Batch $1.50 $7.50 $9K 85% cheaper
    Kimi-K2.6 Async $0.50 $2.56 $3.1K 83% cheaper Try Kimi-K2.6
    ↳ Batch $0.33 $1.71 $2.0K 89% cheaper
    GLM-5.3New Async $1.05 $3.30 $4.3K 93% cheaper Try GLM-5.3
    ↳ Batch $0.70 $2.20 $2.9K 95% cheaper
    GLM-5.3-FlashNew Async $0.11 $0.38 $490 99% cheaper Try GLM-5.3-Flash
    ↳ Batch $0.08 $0.25 $330 99% cheaper
    GLM-5.2New Async $0.70 $2.25 $3.0K 84% cheaper Try GLM-5.2
    ↳ Batch $0.47 $1.50 $2.0K 89% cheaper
    GLM-5.1 Async $0.79 $2.63 $3.4K 81% cheaper Try GLM-5.1
    ↳ Batch $0.53 $1.75 $2.3K 87% cheaper
    Hy3 Async $0.11 $0.44 $550 97% cheaper Try Hy3
    ↳ Batch $0.07 $0.29 $360 98% cheaper
    Qwen3.5-397B-A17B Async $0.29 $1.84 $2.1K 93% cheaper Try Qwen3.5-397B-A17B
    ↳ Batch $0.19 $1.23 $1.4K 95% cheaper
    Qwen3.6-35B-A3B Async $0.11 $0.75 $860 95% cheaper Try Qwen3.6-35B-A3B
    ↳ Batch $0.07 $0.50 $570 97% cheaper
    Qwen3.5-35B-A3B Async $0.07 $0.30 $370 94% cheaper Try Qwen3.5-35B-A3B
    ↳ Batch $0.05 $0.20 $250 96% cheaper
    Qwen3.5-4B Async $0.05 $0.08 $130 99% cheaper Try Qwen3.5-4B
    ↳ Batch $0.04 $0.06 $100 99% cheaper
    Qwen3.5-9B Async $0.08 $0.11 $190 97% cheaper Try Qwen3.5-9B
    ↳ Batch $0.05 $0.08 $130 98% cheaper
    Muse-Glimmer-30BNew Async $0.07 $0.24 $310 95% cheaper Try Muse-Glimmer-30B
    ↳ Batch $0.05 $0.15 $200 97% cheaper
    Gemma-4-31B Async $0.09 $0.26 $350 94% cheaper Try Gemma-4-31B
    ↳ Batch $0.06 $0.18 $240 96% cheaper
    Nemotron-3-Ultra-550B-A55B Async $0.38 $1.65 $2.0K 93% cheaper Try Nemotron-3-Ultra-550B-A55B
    ↳ Batch $0.25 $1.10 $1.4K 96% cheaper
    Nemotron-3-Super-120B-A12B Async $0.06 $0.34 $400 93% cheaper Try Nemotron-3-Super-120B-A12B
    ↳ Batch $0.04 $0.23 $270 96% cheaper
    GPT-OSS-20B Async $0.02 $0.10 $120 99% cheaper Try GPT-OSS-20B
    ↳ Batch $0.02 $0.07 $90 99% cheaper
    Qwen3-VL-235B-A22B Async $0.16 $1.43 $1.6K 67% cheaper Try Qwen3-VL-235B-A22B
    ↳ Batch $0.11 $0.95 $1.1K 78% cheaper
    Qwen3-VL-30B-A3B Async $0.11 $0.45 $560 88% cheaper Try Qwen3-VL-30B-A3B
    ↳ Batch $0.08 $0.30 $380 92% cheaper
    Qwen3-14B Async $0.03 $0.30 $330 93% cheaper Try Qwen3-14B
    ↳ Batch $0.02 $0.20 $220 95% cheaper
    DeepSeek-OCR-2OCR Async $0.08 $0.08 $160 — Try DeepSeek-OCR-2
    ↳ Batch $0.05 $0.05 $100 —
    olmOCR-2-7BOCR Async $0.15 $0.15 $300 — Try olmOCR-2-7B
    ↳ Batch $0.10 $0.10 $200 —
    LightOnOCR-2-1BOCR Async $0.08 $0.08 $160 — Try LightOnOCR-2-1B
    ↳ Batch $0.05 $0.05 $100 —
    Qwen3-Embedding-8B Async $0.03 — $30 — Try Qwen3-Embedding-8B
    ↳ Batch $0.02 — $20 —

    No surprises. No lock-in.

    No credit card required
    No minimum spend
    Pay only for tokens used
    Results stream as they're ready
    OpenAI-compatible API
    Cancel or retry any batch, any time

    We can offer custom pricing for bulk discounts, large workloads, and dedicated deployments - reach out to [email protected].

    Stop overpaying for inference.

    Run your background agents and workloads at a fraction of the price and double the scale.

    If you can wait an hour, you can save a lot.