Doubleword
    Pricing

    Same intelligence. Up to 99% cheaper.

    One API. Open-weight models. Pick your delivery window — Async or Batch — and pay only for what you use.

    Per-token pricing

    Same Intelligence. Fraction of the price.

    Cost to process 1 billion tokens in + 1 billion tokens out at comparable intelligence.

    Model
    Anthropic
    $60K
    OpenAI
    $35K
    Industry Average
    $18K
    Doubleword (Async)
    $13.4K
    $0$15K$30K$45K$60K

    Intelligence via Artificial Analysis Index v4.0 · Hover any bar for full pricing details · Want access to a model you don't see here — just ask us!

    No credit card required · No minimum spend · Pay only for tokens used

    Try Kimi-K3

    We can offer custom pricing for bulk discounts, large workloads, and dedicated deployments — reach out to hello@doubleword.ai.

    Three speeds. One API.

    Pick the delivery window that fits your workflow. All tiers use the same OpenAI-compatible API.

    Real Time

    Iterate on prompts with real-time responses. Full price, zero wait.

    Async Inference

    up to 50% off RT

    Background agents that need results fast. High throughput inference.

    Batch (~24 hours)

    up to 80% off RT

    Big batch jobs where cost matters most. Deepest discounts.

    Full pricing

    Model-by-model breakdown

    ModelSLAInput $/MTokOutput $/MTokCost / 1B in+outvs Big Token
    Inkling-NVFP4NewAsync$0.90$3.00$3.9K78% cheaperTry Inkling-NVFP4
    Batch$0.60$2.00$2.6K86% cheaper
    Qwen3.8-27B-FP8Async$0.35$2.25$2.6K91% cheaperTry Qwen3.8-27B-FP8
    Batch$0.25$1.50$1.8K94% cheaper
    DeepSeek-V4-ProAsync$0.98$1.95$2.9K90% cheaperTry DeepSeek-V4-Pro
    Batch$0.65$1.30$2.0K94% cheaper
    DeepSeek-V4-Flash-0731NewAsync$0.07$0.14$21093% cheaperTry DeepSeek-V4-Flash-0731
    Batch$0.05$0.09$14095% cheaper
    DeepSeek-V4-FlashAsync$0.07$0.14$21093% cheaperTry DeepSeek-V4-Flash
    Batch$0.05$0.09$14095% cheaper
    Kimi-K3NewAsync$2.15$11.25$13.4K78% cheaperTry Kimi-K3
    Batch$1.50$7.50$9K85% cheaper
    Kimi-K2.6Async$0.50$2.56$3.1K83% cheaperTry Kimi-K2.6
    Batch$0.33$1.71$2.0K89% cheaper
    GLM-5.2-FP8NewAsync$0.70$2.25$3.0K84% cheaperTry GLM-5.2-FP8
    Batch$0.47$1.50$2.0K89% cheaper
    GLM-5.1-FP8Async$0.79$2.63$3.4K81% cheaperTry GLM-5.1-FP8
    Batch$0.53$1.75$2.3K87% cheaper
    Hy3-FP8Async$0.11$0.44$55097% cheaperTry Hy3-FP8
    Batch$0.07$0.29$36098% cheaper
    Qwen3.5-397B-A17BAsync$0.29$1.84$2.1K93% cheaperTry Qwen3.5-397B-A17B
    Batch$0.19$1.23$1.4K95% cheaper
    Qwen3.6-35B-A3B-FP8Async$0.11$0.75$86095% cheaperTry Qwen3.6-35B-A3B-FP8
    Batch$0.07$0.50$57097% cheaper
    Qwen3.5-35B-A3B-FP8Async$0.07$0.30$37094% cheaperTry Qwen3.5-35B-A3B-FP8
    Batch$0.05$0.20$25096% cheaper
    Qwen3.5-4BAsync$0.05$0.08$13099% cheaperTry Qwen3.5-4B
    Batch$0.04$0.06$10099% cheaper
    Qwen3.5-9BAsync$0.08$0.11$19097% cheaperTry Qwen3.5-9B
    Batch$0.05$0.08$13098% cheaper
    Muse-Glimmer-30BNewAsync$0.07$0.24$31095% cheaperTry Muse-Glimmer-30B
    Batch$0.05$0.15$20097% cheaper
    Gemma-4-31BAsync$0.09$0.26$35094% cheaperTry Gemma-4-31B
    Batch$0.06$0.18$24096% cheaper
    Nemotron-3-Ultra-550B-A55BAsync$0.38$1.65$2.0K93% cheaperTry Nemotron-3-Ultra-550B-A55B
    Batch$0.25$1.10$1.4K96% cheaper
    Nemotron-3-Super-120B-A12BAsync$0.06$0.34$40093% cheaperTry Nemotron-3-Super-120B-A12B
    Batch$0.04$0.23$27096% cheaper
    GPT-OSS-20BAsync$0.02$0.10$12099% cheaperTry GPT-OSS-20B
    Batch$0.02$0.07$9099% cheaper
    Qwen3-VL-235B-A22BAsync$0.16$1.43$1.6K29% cheaperTry Qwen3-VL-235B-A22B
    Batch$0.11$0.95$1.1K53% cheaper
    Qwen3-VL-30B-A3BAsync$0.11$0.45$56063% cheaperTry Qwen3-VL-30B-A3B
    Batch$0.08$0.30$38075% cheaper
    Qwen3-14B-FP8Async$0.03$0.30$33078% cheaperTry Qwen3-14B-FP8
    Batch$0.02$0.20$22085% cheaper
    DeepSeek-OCR-2OCRAsync$0.08$0.08$160Try DeepSeek-OCR-2
    Batch$0.05$0.05$100
    olmOCR-2-7BOCRAsync$0.15$0.15$300Try olmOCR-2-7B
    Batch$0.10$0.10$200
    LightOnOCR-2-1BOCRAsync$0.08$0.08$160Try LightOnOCR-2-1B
    Batch$0.05$0.05$100
    Qwen3-Embedding-8BAsync$0.03$30Try Qwen3-Embedding-8B
    Batch$0.02$20

    No surprises. No lock-in.

    No credit card required
    No minimum spend
    Pay only for tokens used
    Results stream as they're ready
    OpenAI-compatible API
    Cancel or retry any batch, any time

    We can offer custom pricing for bulk discounts, large workloads, and dedicated deployments - reach out to hello@doubleword.ai.

    Stop overpaying for inference.

    Run your background agents and workloads at a fraction of the price and double the scale.

    If you can wait an hour, you can save a lot.