OpenRouter vs Fireworks AI. Compare a fixed OpenRouter DeepInfra backend with Fireworks Standard hosting, not OpenAI direct API pricing.
USD token-only estimates using each offering’s published default tier. 1M tokens are an aggregate workload, not a single request. Cache savings, batch discounts, long-context surcharges, tools, taxes and negotiated rates are excluded. Tokenizers and capabilities may differ; lower price does not establish equivalent quality.
OpenRouter · OpenAI: gpt-oss-120b · OpenRouter / deepinfra/bf16 is 72.4% cheaper for this workload. Absolute difference: $0.543.
OpenRouter route quote for API model openai/gpt-oss-120b, fixed backend deepinfra/bf16 (bf16). Not the upstream developer's direct API price or an OpenRouter-wide minimum. Text-token estimate only; pin this backend and disable fallbacks to match this quote. Taxes, credit-purchase fees and optional tools are excluded.
Fireworks AI Standard global serverless token rates, not the model developer's direct API price. Priority, Fast, US-only serving, dedicated hosting and optional tools are excluded. Model identifier is the official Fireworks catalog slug.