{"uid":"cap_lDV3WKvk9E_95_L9hCgFY","slug":"argo-hub-machine-payable-llm-inference-premium-ab3e6cfb","name":"Argo Hub — Machine-Payable LLM Inference (Premium)","description":"x402-paid LLM APIs on a private GPU fleet (Qwen3.8-27B, VL 30B). USDC on Base, settle in seconds via a self-hosted facilitator. Unique: verifier-gated correctness on /premium/deep — wrong answers regenerated free. First call free, no wallet: POST /premium/first.","url":"https://diomedes.tail1c6d58.ts.net/premium/llm","method":"POST","headers":{},"bodySchema":{"type":"object","required":["description","mimeType","input"],"properties":{"input":{"type":"object","required":["method"],"properties":{"body":{"type":"object"},"type":{"type":"string"},"method":{"type":"string"},"bodyType":{"type":"string"},"input_schema":{"type":"object"}}},"output":{"type":"object"},"mimeType":{"type":"string"},"description":{"type":"string"}}},"responseSchema":{"type":"application/json","example":{"text":"...","elapsed_s":1.7,"model_tier":"premium"}},"example":null,"exampleRequest":null,"tags":["x402"],"displayCostAmount":"0.01","displayCostAsset":"USDC","priceDynamic":false,"priceHint":null,"priceStatus":"priced","priceSource":"probe","requiresHandshake":false,"reviewCount":0,"rating":{"score":"0.00","successRate":"0.00","reviews":0,"stars":null,"state":"unrated"},"availabilityStatus":"unknown","priceObserved":null,"sessionDeposit":null,"pricing":{"kind":"static","summary":"$0.01/call","primary":{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.01","per":"call","confidence":"exact"},"accepted":[{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.01","per":"call","confidence":"exact"}]},"paymentMethods":[{"uid":"pm_cYiuY0V8i2R1PNlI6V_Sk","protocol":"x402","methodType":"crypto","chain":"base","mode":"charge","costAmount":"0.01","costPer":"request","priority":0,"asset":"0x833589fCD6eDb6E08f4c7C32D4f71b54bdA02913","unit":"request","depositMicros":null,"planRef":null}],"brandName":null,"brandSlug":null,"brandBaseUrl":null,"brandDocsUrl":null,"whatItDoes":"Runs LLM inference on a private GPU fleet (Qwen3 8B–27B, VL 30B) with per-call USDC micropayment via the x402 protocol, settling on Base in seconds.","exampleAgentPrompt":"Use the Argo Hub premium LLM endpoint to answer this question: 'What are the key differences between transformer and state-space models?' — send it as a POST with the question as the body and pay the $0.01 USDC fee automatically.","exampleUseCases":[{"title":"Agent reasoning with pay-per-call LLM","prompt":"I need you to think through this multi-step logic puzzle for me using the Argo Hub premium model — 'A farmer has 3 sons and 17 horses; how do they divide them so each gets a whole number?' Pay the per-call fee and give me the answer."},{"title":"Multimodal vision query on uploaded image","prompt":"Run this image description request through the Argo Hub VL 30B vision model: here's a product photo URL — tell me what's in it, what condition it looks like it's in, and whether it appears authentic. Use the premium tier and charge the USDC fee."},{"title":"Automated content drafting for agent pipeline","prompt":"Draft a two-paragraph executive summary of the following market research notes using the Argo Hub Qwen3-27B model — I want high quality output and I'm fine paying the one-cent fee per call: [paste notes here]."}],"resultDescription":"A JSON object containing a 'text' field with the model's generated response, 'elapsed_s' with the inference time in seconds, and 'model_tier' indicating the tier used (e.g. 'premium'). Payment is settled on-chain via USDC on Base before the response is returned.","failureModes":["Payment not received or insufficient USDC balance — returns 402 Payment Required","Invalid or missing required input fields (description, mimeType, input.method) — returns 400 Bad Request","Model generates an incorrect answer on /premium/deep — endpoint regenerates for free per verifier-gated correctness guarantee","GPU fleet unavailable or overloaded — possible timeout or 503 Service Unavailable","Malformed body or unsupported bodyType — returns 400 with schema validation error","Facilitator settlement failure on Base — payment not confirmed, request rejected"],"whenToPreferThis":"Choose this endpoint when your agent needs high-quality LLM inference (Qwen3 8B–27B or VL 30B) and can pay per call in USDC via the x402 protocol without needing a subscription or API key. It is especially well-suited for autonomous agents operating on Base that need to settle AI compute costs programmatically and in real time. Prefer it over OpenAI or Anthropic APIs when you want crypto-native payment rails, private GPU hosting, or the verifier-gated correctness guarantee available on the /premium/deep variant.","instructions":null,"reviewSummary":null,"reviewSummaryHighlights":null,"reviewSummaryConcerns":null,"reviewSummaryGeneratedAt":null,"activationCount":0,"lastUsedAt":null,"lastSuccessfullyRanAt":null,"lastHealthCheckAt":"2026-09-18T12:38:20.549Z","isFirstParty":false}