{"uid":"cap_CZG9Z8F0eoTeh-JsIw77h","slug":"modelprices-xyz-llama-pricing-table-acaea3d0","name":"modelprices.xyz Llama Pricing Table","description":"Llama pricing table: what every Llama model costs per token right now — Llama 4 Scout, Llama 4 Maverick, Llama 3.3 — from Meta and every provider that resells or hosts them (Bedrock, Azure, Vertex, OpenRouter, Fireworks). Input, output, cache and batch USD per 1M tokens, sorted cheapest first, so you can compare Llama inference cost across hosts and pick the cheapest place to run one. Normalized, cross-checked, refreshed hourly.","url":"https://modelprices.xyz/llama-pricing","method":"GET","headers":{},"bodySchema":{"type":"object","$schema":"https://json-schema.org/draft/2020-12/schema","required":["input"],"properties":{"input":{"type":"object","required":["type","method"],"properties":{"type":{"type":"string","const":"http"},"method":{"enum":["GET"],"type":"string"},"queryParams":{"type":"object","properties":{}}},"additionalProperties":false},"output":{"type":"object","required":["type"],"properties":{"type":{"type":"string"},"example":{"type":"object"}}}}},"responseSchema":null,"example":null,"exampleRequest":null,"tags":["x402"],"displayCostAmount":"0.01","displayCostAsset":"USDC","priceDynamic":false,"priceHint":null,"priceStatus":"priced","priceSource":"probe","requiresHandshake":false,"reviewCount":0,"rating":{"score":"0.00","successRate":"0.00","reviews":0,"stars":null,"state":"unrated"},"availabilityStatus":"unknown","priceObserved":null,"sessionDeposit":null,"pricing":{"kind":"static","summary":"$0.01/call","primary":{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.01","per":"call","confidence":"exact"},"accepted":[{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.01","per":"call","confidence":"exact"}]},"paymentMethods":[{"uid":"pm_5pYa4e2mg41PsqdcMv55S","protocol":"x402","methodType":"crypto","chain":"base","mode":"charge","costAmount":"0.01","costPer":"request","priority":0,"asset":"0x833589fCD6eDb6E08f4c7C32D4f71b54bdA02913","unit":"request","depositMicros":null,"planRef":null}],"brandName":null,"brandSlug":null,"brandBaseUrl":null,"brandDocsUrl":null,"whatItDoes":"Returns a normalized, cross-provider pricing table for all Llama models (input, output, cache, batch USD per 1M tokens), sorted cheapest first, refreshed hourly.","exampleAgentPrompt":"Show me the full Llama pricing table across all providers — Bedrock, Azure, Vertex, OpenRouter, Fireworks — sorted cheapest first so I can see input, output, and batch prices per million tokens for Llama 4 Scout, Llama 4 Maverick, and Llama 3.3 right now.","exampleUseCases":[{"title":"Pick cheapest Llama 4 Scout host","prompt":"I need to run Llama 4 Scout at scale — can you pull the latest pricing table and tell me which provider charges the least for input and output tokens right now?"},{"title":"Cost comparison before cloud migration","prompt":"We're deciding between AWS Bedrock and Azure for our Llama 3.3 deployment — fetch the current per-token prices from both so I can see the difference in input, output, and batch rates."},{"title":"Budget LLM selection for startup","prompt":"We have a tight inference budget and want to use a Llama model — show me every Llama option available across all providers sorted by cheapest output token price so I can pick the most affordable one."}],"resultDescription":"A normalized table of all Llama models (Llama 4 Scout, Llama 4 Maverick, Llama 3.3, etc.) with per-provider pricing rows including input, output, cache, and batch costs in USD per 1 million tokens, sorted from cheapest to most expensive, cross-checked and refreshed hourly across Meta, AWS Bedrock, Azure, Google Vertex, OpenRouter, and Fireworks.","failureModes":["Pricing data temporarily stale if provider APIs are unavailable during refresh cycle","Missing rows if a provider has recently changed their pricing page structure","Empty result if no Llama models are currently listed by any provider","HTTP 402 if payment of $0.01 USDC is not included with the request"],"whenToPreferThis":"Use this endpoint when you specifically need Llama model pricing across multiple hosting providers in a single normalized call, especially when comparing Meta's Llama family (Scout, Maverick, 3.3) across AWS Bedrock, Azure, Vertex, OpenRouter, and Fireworks. Prefer this over general model search endpoints when the user's query is specifically about Llama-family cost or provider selection for Llama inference.","instructions":null,"reviewSummary":null,"reviewSummaryHighlights":null,"reviewSummaryConcerns":null,"reviewSummaryGeneratedAt":null,"activationCount":0,"lastUsedAt":null,"lastSuccessfullyRanAt":null,"lastHealthCheckAt":"2026-09-15T00:45:03.677Z","isFirstParty":false}