{"uid":"cap_UM8NOv2IrwhakRO0qciMG","slug":"openai-gpt-6-astra-long-272k-context-via-mm-family-3d608899","name":"OpenAI GPT-6 Astra Long (>272K context) via mm.family","description":"gpt-6-astra-long: OpenAI GPT 6 Astra (>272K input tokens) chat completions, paid per call in USDC. Per 1M tokens in/out: Flex (default) 10.0/37.5, Standard 20.0/75.0, Fast 40.0/150.0; pick with service_tier. Plus $0.0005 (Base) or $0.0005 (Solana) per call; the quote charges input plus 10% of max_tokens. Levels: low, medium, high, xhigh (model gpt-6-astra-long:<level>). Empty answers are free. Rates: https://openai.mm.family/x402/pricing","url":"https://openai.mm.family/x402/v1/models/gpt-6-astra-long/chat/completions?utm_source=zero.xyz","method":"POST","headers":{},"bodySchema":{"type":"object","properties":{"model":{"enum":["gpt-6-astra-long","gpt-6-astra-long:low","gpt-6-astra-long:medium","gpt-6-astra-long:high","gpt-6-astra-long:xhigh"],"type":"string"},"tools":{"type":"array"},"stream":{"type":"boolean"},"messages":{"type":"array","items":{"type":"object","required":["role","content"],"properties":{"role":{"type":"string"},"content":{"type":"string"}}}},"max_tokens":{"type":"integer"},"service_tier":{"enum":["flex","default","standard","auto","fast","priority"],"type":"string"},"response_format":{"type":"object"},"reasoning_effort":{"type":"string"}}},"responseSchema":{"type":"json","example":{"id":"chatcmpl-x","model":"gpt-6-astra-long","usage":{"total_tokens":9,"prompt_tokens":8,"completion_tokens":1},"object":"chat.completion","choices":[{"index":0,"message":{"role":"assistant","content":"Hello","refusal":null},"finish_reason":"stop"}],"service_tier":"flex"}},"example":null,"exampleRequest":null,"tags":["x402"],"displayCostAmount":"0.002477","displayCostAsset":"USDC","priceDynamic":false,"priceHint":null,"priceStatus":"priced","priceSource":"probe","requiresHandshake":false,"reviewCount":0,"rating":{"score":"0.00","successRate":"0.00","reviews":0,"stars":null,"state":"unrated"},"availabilityStatus":"unknown","priceObserved":null,"sessionDeposit":null,"pricing":{"kind":"static","summary":"$0.002477/call","primary":{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.002477","per":"call","confidence":"exact"},"accepted":[{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.002477","per":"call","confidence":"exact"}]},"paymentMethods":[{"uid":"pm_QTwyGnxhNWJXDzS_FY82k","protocol":"x402","methodType":"crypto","chain":"base","mode":"charge","costAmount":"0.002477","costPer":"request","priority":0,"asset":"0x833589fCD6eDb6E08f4c7C32D4f71b54bdA02913","unit":"request","depositMicros":null,"planRef":null}],"brandName":null,"brandSlug":null,"brandBaseUrl":null,"brandDocsUrl":null,"whatItDoes":"Pay-per-call OpenAI GPT-6 Astra chat completions for long-context inputs (>272K tokens), billed in USDC with configurable speed/cost tiers","exampleAgentPrompt":"Send this 400,000-token legal document and my analysis questions as a chat prompt to GPT-6 Astra Long using the flex service tier with a max_tokens of 4096 — bill the call in USDC.","exampleUseCases":[{"title":"Full codebase review and refactor","prompt":"I need you to load our entire monorepo source code — about 350K tokens — into GPT-6 Astra Long and ask it to identify architectural issues and suggest refactors. Use the standard service tier and set max_tokens to 8192."},{"title":"Long legal contract summarization","prompt":"Take this 300-page contract (it's over 300K tokens) and send it to GPT-6 Astra Long with the question 'Summarize all indemnification and liability clauses.' Use the flex tier to keep costs low and cap the response at 2048 tokens."},{"title":"Book-length research synthesis","prompt":"I have a collection of research papers totalling around 400K tokens — feed them all to GPT-6 Astra Long and ask it to synthesize the key findings into a structured report. Use the fast tier because I need this quickly, and allow up to 16384 output tokens."}],"resultDescription":"Returns a standard OpenAI-compatible chat completion JSON object with the assistant's reply message, finish reason, token usage breakdown (prompt, completion, total), model name, and the service tier used for the request.","failureModes":["Insufficient USDC balance — payment rejected before completion","Input token count does not actually exceed 272K threshold — should use gpt-6-astra instead","Invalid service_tier value — must be one of flex, default, standard, auto, fast, priority","max_tokens too large for available budget — quoted charge may be rejected","Model variant string malformed — must be gpt-6-astra-long or gpt-6-astra-long:<level>","Empty response (zero-length answer) — billed at $0 but counts as a call","Network timeout on very large context processing under fast tier"],"whenToPreferThis":"Choose this endpoint when your input context exceeds 272K tokens and you need GPT-6 Astra specifically — it is the long-context variant designed for inputs that overflow the standard gpt-6-astra endpoint. Prefer it over gpt-6-astra when processing full codebases, lengthy legal documents, or large research corpora. Use the flex tier for cost efficiency, standard for balanced performance, or fast for latency-sensitive workloads. Prefer this over gpt-5.x sibling models when you need the highest capability frontier model for complex long-context reasoning.","instructions":null,"reviewSummary":null,"reviewSummaryHighlights":null,"reviewSummaryConcerns":null,"reviewSummaryGeneratedAt":null,"activationCount":0,"lastUsedAt":null,"lastSuccessfullyRanAt":null,"lastHealthCheckAt":"2026-10-02T12:42:36.378Z","isFirstParty":false,"canonicalSlug":"openai-gpt-6-astra-long-272k-context-via-mm-family-3d608899"}