{"uid":"cap_ULMjrvg5-FSSy9cFdPf0-","slug":"argo-hub-deep-llm-inference-x402-paid-09707e1d","name":"Argo Hub — Deep LLM Inference (x402-paid)","description":"x402-paid LLM APIs on a private GPU fleet (Qwen3.8-27B, VL 30B). USDC on Base, settle in seconds via a self-hosted facilitator. Unique: verifier-gated correctness on /premium/deep — wrong answers regenerated free. First call free, no wallet: POST /premium/first.","url":"https://diomedes.tail1c6d58.ts.net/premium/deep","method":"POST","headers":{},"bodySchema":{"type":"object","required":["description","mimeType","input"],"properties":{"input":{"type":"object","required":["method"],"properties":{"body":{"type":"object"},"type":{"type":"string"},"method":{"type":"string"},"bodyType":{"type":"string"},"input_schema":{"type":"object"}}},"output":{"type":"object"},"mimeType":{"type":"string"},"description":{"type":"string"}}},"responseSchema":{"type":"application/json","example":{"text":"...","model_tier":"deep"}},"example":null,"exampleRequest":null,"tags":["x402"],"displayCostAmount":"0.05","displayCostAsset":"USDC","priceDynamic":false,"priceHint":null,"priceStatus":"priced","priceSource":"probe","requiresHandshake":false,"reviewCount":0,"rating":{"score":"0.00","successRate":"0.00","reviews":0,"stars":null,"state":"unrated"},"availabilityStatus":"unknown","priceObserved":null,"sessionDeposit":null,"pricing":{"kind":"static","summary":"$0.05/call","primary":{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.05","per":"call","confidence":"exact"},"accepted":[{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.05","per":"call","confidence":"exact"}]},"paymentMethods":[{"uid":"pm_QMb3kZTD5gUxwjXCjf9gQ","protocol":"x402","methodType":"crypto","chain":"base","mode":"charge","costAmount":"0.05","costPer":"request","priority":0,"asset":"0x833589fCD6eDb6E08f4c7C32D4f71b54bdA02913","unit":"request","depositMicros":null,"planRef":null}],"brandName":null,"brandSlug":null,"brandBaseUrl":null,"brandDocsUrl":null,"whatItDoes":"Runs a deep LLM inference call on a private GPU fleet (Qwen3-27B-class) with verifier-gated correctness guarantees, paid per-call in USDC via x402.","exampleAgentPrompt":"Use the Argo Hub deep inference tier to answer this complex reasoning question: 'What are the key trade-offs between transformer and state-space models for long-context tasks?' — the endpoint should regenerate for free if the answer is wrong.","exampleUseCases":[{"title":"Per-call AI reasoning for agent pipeline","prompt":"I need to run a complex multi-step reasoning query through a pay-per-call LLM without signing up for a subscription — use the Argo Hub deep model and pay the $0.05 USDC fee automatically. The question is: 'Given the following financial data, what is the most likely cause of the Q3 revenue dip?'"},{"title":"Correctness-guaranteed factual Q&A","prompt":"Ask the Argo deep inference endpoint whether the Pythagorean theorem applies in non-Euclidean geometry, and make sure we get a verified correct answer — I want the free regeneration guarantee if it gets it wrong."},{"title":"Vision-language analysis on private GPU","prompt":"Send this product image description to the Argo Hub deep model and ask it to generate a detailed marketing copy paragraph — I want to use the verifier-gated tier so we only pay if the output is actually good."}],"resultDescription":"Returns a JSON object containing a 'text' field with the model's generated response and a 'model_tier' field set to 'deep', indicating which inference tier handled the request. If the verifier deems the answer incorrect, the response is regenerated at no additional cost.","failureModes":["Payment not settled — x402 USDC transaction fails or wallet has insufficient funds, resulting in a 402 response and no inference","Invalid input schema — missing required fields ('description', 'mimeType', 'input.method') returns a 400-class error","Model timeout — long-running deep inference exceeds GPU time limits and returns a 504 or equivalent error","Verifier loop — rare edge case where regenerated answers repeatedly fail correctness checks, potentially stalling the response","Network unreachable — private Tailscale domain (ts.net) may be inaccessible outside the operator's network"],"whenToPreferThis":"Choose this endpoint when you need high-quality, deep reasoning LLM output with a correctness guarantee, are comfortable with per-call USDC micropayments via x402 on Base, and want to avoid subscription-based APIs. Particularly valuable for agent pipelines that require reliable, verifiable answers and want automatic free regeneration on wrong outputs. Prefer over standard hosted APIs when cost-per-call economics matter and you want a self-hosted, private GPU backend rather than shared cloud inference.","instructions":null,"reviewSummary":null,"reviewSummaryHighlights":null,"reviewSummaryConcerns":null,"reviewSummaryGeneratedAt":null,"activationCount":0,"lastUsedAt":null,"lastSuccessfullyRanAt":null,"lastHealthCheckAt":"2026-09-18T06:34:47.630Z","isFirstParty":false}