{"uid":"cap_YYW1chjOruNiJxPYaLgk2","slug":"modelprices-xyz-qwen-pricing-table-a272febe","name":"modelprices.xyz Qwen Pricing Table","description":"Qwen pricing table: what every Qwen model costs per token right now — Qwen 3, Qwen 2.5, Qwen Coder — from Alibaba and every provider that resells or hosts them (Bedrock, Azure, Vertex, OpenRouter, Fireworks). Input, output, cache and batch USD per 1M tokens, sorted cheapest first, so you can compare Qwen inference cost across hosts and pick the cheapest place to run one. Normalized, cross-checked, refreshed hourly.","url":"https://modelprices.xyz/qwen-pricing","method":"GET","headers":{},"bodySchema":{"type":"object","$schema":"https://json-schema.org/draft/2020-12/schema","required":["input"],"properties":{"input":{"type":"object","required":["type","method"],"properties":{"type":{"type":"string","const":"http"},"method":{"enum":["GET"],"type":"string"},"queryParams":{"type":"object","properties":{}}},"additionalProperties":false},"output":{"type":"object","required":["type"],"properties":{"type":{"type":"string"},"example":{"type":"object"}}}}},"responseSchema":null,"example":null,"exampleRequest":null,"tags":["x402"],"displayCostAmount":"0.01","displayCostAsset":"USDC","priceDynamic":false,"priceHint":null,"priceStatus":"priced","priceSource":"probe","requiresHandshake":false,"reviewCount":0,"rating":{"score":"0.00","successRate":"0.00","reviews":0,"stars":null,"state":"unrated"},"availabilityStatus":"unknown","priceObserved":null,"sessionDeposit":null,"pricing":{"kind":"static","summary":"$0.01/call","primary":{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.01","per":"call","confidence":"exact"},"accepted":[{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.01","per":"call","confidence":"exact"}]},"paymentMethods":[{"uid":"pm_hLnXKNUWtIZ12Unjj7R2S","protocol":"x402","methodType":"crypto","chain":"base","mode":"charge","costAmount":"0.01","costPer":"request","priority":0,"asset":"0x833589fCD6eDb6E08f4c7C32D4f71b54bdA02913","unit":"request","depositMicros":null,"planRef":null}],"brandName":null,"brandSlug":null,"brandBaseUrl":null,"brandDocsUrl":null,"whatItDoes":"Returns per-token pricing for every Qwen model across all major hosting providers, sorted cheapest first, refreshed hourly.","exampleAgentPrompt":"What's the cheapest place to run Qwen 3 right now — can you pull the latest per-token prices across all providers like Bedrock, Azure, Vertex, OpenRouter, and Fireworks so I can compare input and output costs?","exampleUseCases":[{"title":"Cost-optimize Qwen Coder deployment","prompt":"I'm building a coding assistant using Qwen Coder — can you check which provider is cheapest for input and output tokens right now so I know where to host it?"},{"title":"Cross-provider budget comparison for Qwen 2.5","prompt":"We're deciding between Azure, Vertex, and OpenRouter for Qwen 2.5 — pull the latest token pricing for all of them so I can see the cost difference before we commit."},{"title":"Hourly cost estimate for Qwen inference","prompt":"I'm processing about 50 million tokens a month with Qwen — what are the current per-token rates across all hosts so I can estimate my monthly bill and pick the cheapest option?"}],"resultDescription":"A normalized table of Qwen models and their per-token costs (input, output, cache, batch in USD per 1M tokens) across all providers that host them — including Alibaba, AWS Bedrock, Azure, Vertex AI, OpenRouter, and Fireworks — sorted from cheapest to most expensive, refreshed hourly.","failureModes":["Pricing data temporarily stale if upstream provider APIs are unavailable during refresh cycle","New Qwen models may not appear immediately after release","Provider-specific promotional or enterprise pricing not reflected","Cache/batch pricing missing for providers that do not publish those rates","HTTP 402 if payment is not included or fails"],"whenToPreferThis":"Use this endpoint when you need to compare Qwen-family model inference costs across multiple providers in a single call. Prefer it over manually querying each provider's pricing page, or when building cost-optimization logic specifically for Qwen models. It is particularly useful when you need cache and batch pricing alongside standard input/output rates, all normalized to the same USD-per-1M-token unit.","instructions":null,"reviewSummary":null,"reviewSummaryHighlights":null,"reviewSummaryConcerns":null,"reviewSummaryGeneratedAt":null,"activationCount":0,"lastUsedAt":null,"lastSuccessfullyRanAt":null,"lastHealthCheckAt":"2026-09-14T12:41:07.481Z","isFirstParty":false}