{"uid":"cap_m-IKkKxhiaLuwAutu4Z8A","slug":"modelprices-xyz-largest-context-window-leaderboard-6f634375","name":"modelprices.xyz Largest Context Window Leaderboard","description":"Largest-context-window leaderboard: the 100 AI models with the biggest context windows, ranked descending — how many tokens each model can actually take, its max output tokens, modality support, and what those tokens cost. Compare GPT-5, Claude 5, Gemini 3 Pro, Llama 4 Scout and 2,000+ more on context capacity and price together. Refreshed hourly.","url":"https://modelprices.xyz/llm/context-windows","method":"GET","headers":{},"bodySchema":{"type":"object","$schema":"https://json-schema.org/draft/2020-12/schema","required":["input"],"properties":{"input":{"type":"object","required":["type","method"],"properties":{"type":{"type":"string","const":"http"},"method":{"enum":["GET"],"type":"string"},"queryParams":{"type":"object","properties":{}}},"additionalProperties":false},"output":{"type":"object","required":["type"],"properties":{"type":{"type":"string"},"example":{"type":"object"}}}}},"responseSchema":null,"example":null,"exampleRequest":null,"tags":["x402"],"displayCostAmount":"0.01","displayCostAsset":"USDC","priceDynamic":false,"priceHint":null,"priceStatus":"priced","priceSource":"probe","requiresHandshake":false,"reviewCount":0,"rating":{"score":"0.00","successRate":"0.00","reviews":0,"stars":null,"state":"unrated"},"availabilityStatus":"unknown","priceObserved":null,"sessionDeposit":null,"pricing":{"kind":"static","summary":"$0.01/call","primary":{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.01","per":"call","confidence":"exact"},"accepted":[{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.01","per":"call","confidence":"exact"}]},"paymentMethods":[{"uid":"pm_bE8_pvWFU8xwO5ez89aEi","protocol":"x402","methodType":"crypto","chain":"base","mode":"charge","costAmount":"0.01","costPer":"request","priority":0,"asset":"0x833589fCD6eDb6E08f4c7C32D4f71b54bdA02913","unit":"request","depositMicros":null,"planRef":null}],"brandName":null,"brandSlug":null,"brandBaseUrl":null,"brandDocsUrl":null,"whatItDoes":"Returns the top 100 AI models ranked by context window size (descending), with token limits, max output tokens, modality support, and per-token pricing.","exampleAgentPrompt":"Show me the 100 AI models with the biggest context windows right now — I need to see their token limits, max output, modality support, and what they cost per token, all ranked from largest to smallest context.","exampleUseCases":[{"title":"Long-document processing model selection","prompt":"I need to process entire legal contracts that can be hundreds of pages long — which AI models have the largest context windows and what do they cost? Show me the top options ranked by how many tokens they can handle."},{"title":"Budget-aware context capacity comparison","prompt":"I'm comparing models for a RAG pipeline and need to balance context window size against cost — can you pull up the leaderboard of the biggest context window models and their per-token prices so I can find the best value?"},{"title":"Multimodal long-context model research","prompt":"Which multimodal AI models — ones that can handle images as well as text — have the largest context windows available right now? I want to see the full ranked list with their token limits and pricing."}],"resultDescription":"A ranked list of up to 100 AI models ordered by context window size (largest first), each entry including the model name, context window token count, maximum output tokens, supported modalities (text, image, etc.), and input/output pricing per token. Data is refreshed hourly.","failureModes":["Payment failure (x402 micropayment not processed — endpoint returns 402 status)","Temporary data unavailability if hourly refresh is in progress","Empty or partial list if fewer than 100 models are currently tracked","Rate limiting if called too frequently"],"whenToPreferThis":"Use this endpoint when you need to compare AI models specifically on context window capacity alongside pricing in a single call. Prefer this over provider-specific pricing tables (e.g., OpenAI or AWS Bedrock endpoints) when your primary concern is raw token capacity across all providers simultaneously. Choose this over the cheapest-long-context leaderboard when you want the absolute maximum token capacity regardless of cost.","instructions":null,"reviewSummary":null,"reviewSummaryHighlights":null,"reviewSummaryConcerns":null,"reviewSummaryGeneratedAt":null,"activationCount":0,"lastUsedAt":null,"lastSuccessfullyRanAt":null,"lastHealthCheckAt":"2026-09-15T00:51:04.719Z","isFirstParty":false}