{"uid":"cap_rx2i21OIDfSWSa7NMTRqU","slug":"openai-multi-model-chat-completions-via-mm-family-x402-0c56e7c6","name":"OpenAI Multi-Model Chat Completions via mm.family (x402)","description":"gpt-6.1-sol (OpenAI's newest model) and 25 more OpenAI chat models (gpt-6, gpt-5.x, gpt-4.1, gpt-4o, o-series), paid per call in USDC. OpenAI token rates on Flex (half price, may answer 429; default where offered), Standard (service_tier default) or Fast (fast), plus $0.0005 (Base) or $0.0005 (Solana) per call; the quote charges input plus 10% of max_tokens. Empty answers are free. One endpoint per model: /x402/v1/models/<key>/chat/completions. Rates: https://openai.mm.family/x402/pricing","url":"https://openai.mm.family/x402/v1/chat/completions?utm_source=zero.xyz","method":"POST","headers":{},"bodySchema":{"type":"object","properties":{"model":{"enum":["chat-latest","gpt-4.1","gpt-4.1-mini","gpt-4.1-nano","gpt-4o","gpt-4o-mini","gpt-5","gpt-5-mini","gpt-5-nano","gpt-5.1","gpt-5.2","gpt-5.4","gpt-5.4-long","gpt-5.4-mini","gpt-5.4-nano","gpt-5.5","gpt-5.5-long","gpt-5.6-luna","gpt-5.6-luna-long","gpt-5.6-sol","gpt-5.6-sol-long","gpt-5.6-terra","gpt-5.6-terra-long","gpt-6-astra","gpt-6-astra-long","gpt-6-luna","gpt-6-luna-long","gpt-6-sol","gpt-6-sol-long","gpt-6.1-sol","gpt-6.1-sol-long","o1","o3","o3-mini","o4-mini"],"type":"string"},"tools":{"type":"array"},"stream":{"type":"boolean"},"messages":{"type":"array","items":{"type":"object","required":["role","content"],"properties":{"role":{"type":"string"},"content":{"type":"string"}}}},"max_tokens":{"type":"integer"},"service_tier":{"enum":["flex","default","standard","auto","fast","priority"],"type":"string"},"response_format":{"type":"object"},"reasoning_effort":{"type":"string"}}},"responseSchema":{"type":"json","example":{"id":"chatcmpl-x","model":"gpt-6-sol","usage":{"total_tokens":9,"prompt_tokens":8,"completion_tokens":1},"object":"chat.completion","choices":[{"index":0,"message":{"role":"assistant","content":"Hello","refusal":null},"finish_reason":"stop"}],"service_tier":"flex"}},"example":null,"exampleRequest":null,"tags":["x402"],"displayCostAmount":"0.001","displayCostAsset":"USDC","priceDynamic":false,"priceHint":null,"priceStatus":"priced","priceSource":"probe","requiresHandshake":false,"reviewCount":0,"rating":{"score":"0.00","successRate":"0.00","reviews":0,"stars":null,"state":"unrated"},"availabilityStatus":"unknown","priceObserved":null,"sessionDeposit":null,"pricing":{"kind":"static","summary":"$0.001/call","primary":{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.001","per":"call","confidence":"exact"},"accepted":[{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.001","per":"call","confidence":"exact"}]},"paymentMethods":[{"uid":"pm_mB8vV_LCvtAR69RnSd6bQ","protocol":"x402","methodType":"crypto","chain":"base","mode":"charge","costAmount":"0.001","costPer":"request","priority":0,"asset":"0x833589fCD6eDb6E08f4c7C32D4f71b54bdA02913","unit":"request","depositMicros":null,"planRef":null}],"brandName":null,"brandSlug":null,"brandBaseUrl":null,"brandDocsUrl":null,"whatItDoes":"Provides pay-per-call access to 25+ OpenAI chat models (GPT-4.1 through GPT-6.1-sol, o-series) via the x402 payment protocol, billed in USDC","exampleAgentPrompt":"Send my conversation to gpt-6.1-sol using the standard service tier with a 2000 token limit and tell me what it says: [system: 'You are a helpful assistant.', user: 'Explain quantum entanglement in simple terms.']","exampleUseCases":[{"title":"Agentic reasoning with latest GPT","prompt":"Use gpt-6.1-sol on the standard tier with up to 4000 tokens to reason through this multi-step problem for me: given these three business constraints, what is the optimal pricing strategy?"},{"title":"Cheap bulk inference with flex pricing","prompt":"Run my list of customer support messages through gpt-5-nano on flex tier, max 500 tokens each, and draft a short reply for each one."},{"title":"Code generation with o-series model","prompt":"Ask o3 to write me a Python function that parses nested JSON and flattens it into a CSV — use the fast tier and give it up to 1500 tokens."}],"resultDescription":"Returns an OpenAI-compatible chat.completion JSON object including the assistant's message content, model used, finish reason, token usage (prompt, completion, total), and the service_tier that handled the request. Empty responses are not charged.","failureModes":["429 Too Many Requests on flex tier (by design — flex may be throttled)","402 Payment Required if USDC payment via x402 is not attached or insufficient","Invalid model name returns 400 or unknown model error","max_tokens quota exceeded causes truncated or refused completion","Empty response returned (billed as free) if model produces no output"],"whenToPreferThis":"Choose this endpoint when you need pay-per-call access to OpenAI's latest and most capable models (GPT-6, GPT-5.x, o-series) without managing an OpenAI subscription, and want to pay in USDC via the x402 protocol. It is especially useful for agentic workflows that need on-demand LLM access billed per call, or when you want to experiment across a wide range of model tiers (nano to sol) and pricing tiers (flex/standard/fast) without committing to a fixed plan.","instructions":null,"reviewSummary":null,"reviewSummaryHighlights":null,"reviewSummaryConcerns":null,"reviewSummaryGeneratedAt":null,"activationCount":0,"lastUsedAt":null,"lastSuccessfullyRanAt":null,"lastHealthCheckAt":"2026-10-02T12:41:58.110Z","isFirstParty":false,"canonicalSlug":"openai-multi-model-chat-completions-via-mm-family-x402-0c56e7c6"}