{"uid":"cap_InIbKI3dkk0urARsppQUX","slug":"omnia-odds-openai-compatible-chat-completions-34258174","name":"Omnia Odds OpenAI-Compatible Chat Completions","description":"OpenAI-compatible chat completions: POST messages[] with model fast (Llama 3.1 8B) or smart (Llama 3.3 70B), max_tokens, temperature; standard chat.completion response. Point any OpenAI SDK at this base URL. LLM inference for agents, no API key, no signup, pay per request","url":"https://odds.rjhsignaltech.workers.dev/v1/chat/completions?utm_source=zero.xyz","method":"POST","headers":{},"bodySchema":{"type":"object","properties":{"model":{"type":"string","description":"\"fast\" (Llama 3.1 8B, default) or \"smart\" (Llama 3.3 70B)"},"messages":{"type":"array","items":{"type":"object","required":["role","content"],"properties":{"role":{"enum":["system","user","assistant"],"type":"string"},"content":{"type":"string"}}}},"max_tokens":{"type":"integer","maximum":1024,"minimum":1},"temperature":{"type":"number","maximum":2,"minimum":0}}},"responseSchema":null,"example":null,"exampleRequest":null,"tags":["x402"],"displayCostAmount":"0.002","displayCostAsset":"USDC","priceDynamic":false,"priceHint":null,"priceStatus":"priced","priceSource":"probe","requiresHandshake":false,"reviewCount":0,"rating":{"score":"0.00","successRate":"0.00","reviews":0,"stars":null,"state":"unrated"},"availabilityStatus":"unknown","priceObserved":null,"sessionDeposit":null,"pricing":{"kind":"static","summary":"$0.002/call","primary":{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.002","per":"call","confidence":"exact"},"accepted":[{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.002","per":"call","confidence":"exact"}]},"paymentMethods":[{"uid":"pm_K6Nq3W_m4VjSjROaNjTZp","protocol":"x402","methodType":"crypto","chain":"base","mode":"charge","costAmount":"0.002","costPer":"request","priority":0,"asset":"0x833589fCD6eDb6E08f4c7C32D4f71b54bdA02913","unit":"request","depositMicros":null,"planRef":null}],"brandName":null,"brandSlug":null,"brandBaseUrl":null,"brandDocsUrl":null,"whatItDoes":"Provides OpenAI-compatible LLM chat completions using Llama 3.1 8B (fast) or Llama 3.3 70B (smart), payable per request via x402 with no API key required.","exampleAgentPrompt":"Using the smart Llama 3.3 70B model, draft a professional apology email to a client explaining a project delay — keep it under 200 words and use a warm but formal tone.","exampleUseCases":[{"title":"Agent task delegation with LLM","prompt":"I need you to summarize this long support ticket into three bullet points using the fast Llama model — here's the text: 'Customer reported that the checkout flow breaks on mobile when they try to apply a discount code, this happens on both iOS and Android, they've tried three times...'"},{"title":"Content drafting for marketing copy","prompt":"Use the smart 70B model to write a catchy product description for my new standing desk, called the FlexDesk Pro — keep it under 150 words, energetic and aimed at remote workers."},{"title":"Classification task inside an agent pipeline","prompt":"Classify this user message as either 'complaint', 'question', or 'compliment' and return just the label — use the fast model: 'Why does your app keep logging me out every hour? It's so frustrating!'"}],"resultDescription":"Returns a standard OpenAI-format chat.completion JSON object containing the assistant's reply message, the model used, token usage counts (prompt, completion, total), and finish reason. The response is drop-in compatible with any OpenAI SDK or client expecting a chat/completions response.","failureModes":["Payment not received or x402 handshake fails — request rejected before inference","max_tokens exceeds 1024 — validation error returned","Invalid model name supplied (not 'fast' or 'smart') — error response","Temperature out of range (0–2) — validation failure","Empty or malformed messages array — bad request error","Cloudflare Worker timeout if model inference takes too long","Upstream Llama inference service unavailable — 5xx error"],"whenToPreferThis":"Choose this endpoint when you need OpenAI-compatible LLM inference with no API key or signup, paying only per request in USDC via x402. It's ideal for autonomous agents that need serverless, frictionless LLM access without managing credentials, and for pipelines already using OpenAI SDK that want a drop-in alternative. Prefer 'fast' (Llama 3.1 8B) for speed-sensitive or high-volume tasks, and 'smart' (Llama 3.3 70B) for quality-critical reasoning or generation tasks.","instructions":null,"reviewSummary":null,"reviewSummaryHighlights":null,"reviewSummaryConcerns":null,"reviewSummaryGeneratedAt":null,"activationCount":0,"lastUsedAt":null,"lastSuccessfullyRanAt":null,"lastHealthCheckAt":"2026-10-02T01:16:09.392Z","isFirstParty":false,"canonicalSlug":"omnia-odds-openai-compatible-chat-completions-34258174"}