{"uid":"cap_WwsLW27gjfpdSPeRJ6b6n","slug":"batcave-detective-basic-agent-conclusion-auditor-65fbd513","name":"Batcave Detective Basic — Agent Conclusion Auditor","description":"Batcave Detective Basic: fresh-context supplied-evidence audit of an agent conclusion for unsupported assumptions, contradictions, stale premises, and drift from the original task.","url":"https://control-replacing-spyware-starting.trycloudflare.com/v1/audit-claims","method":"POST","headers":{},"bodySchema":null,"responseSchema":null,"example":null,"exampleRequest":null,"tags":["x402"],"displayCostAmount":"0.02","displayCostAsset":"USDC","priceDynamic":false,"priceHint":null,"priceStatus":"priced","priceSource":"registry","requiresHandshake":false,"reviewCount":0,"rating":{"score":"0.00","successRate":"0.00","reviews":0,"stars":null,"state":"unrated"},"availabilityStatus":"unknown","priceObserved":null,"sessionDeposit":null,"pricing":{"kind":"static","summary":"$0.02/call","primary":{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.02","per":"call","confidence":"exact"},"accepted":[{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.02","per":"call","confidence":"exact"}]},"paymentMethods":[{"uid":"pm_xZZAmAwnJmuJOLMS17JlO","protocol":"x402","methodType":"crypto","chain":"base","mode":"charge","costAmount":"0.02","costPer":"request","priority":0,"asset":"EPjFWdd5AufqSSqeM2qN1xzybapC8G4wEGGkZwyTDt1v","unit":"request","depositMicros":null,"planRef":null}],"brandName":null,"brandSlug":null,"brandBaseUrl":null,"brandDocsUrl":null,"whatItDoes":"Audits an AI agent's conclusion against supplied evidence to detect unsupported assumptions, contradictions, stale premises, and task drift.","exampleAgentPrompt":"Audit my agent's conclusion — it was originally asked to 'identify the best cloud provider for a latency-sensitive fintech startup' and concluded 'AWS us-east-1 is optimal due to lowest latency and mature compliance tools.' Check these three claims against the supplied evidence and flag any unsupported assumptions, contradictions, or stale premises; set importance to high.","exampleUseCases":[{"title":"Research agent conclusion sanity check","prompt":"My research agent just concluded that 'inflation is expected to remain above 4% through 2025' based on these five evidence snippets I'll supply — can you audit that conclusion and its supporting claims for contradictions or unsupported assumptions? The original task was 'summarize current inflation outlook for US markets' and importance is high."},{"title":"Legal brief reasoning drift detection","prompt":"An agent drafting a legal summary concluded 'the defendant has no prior criminal record' after being asked to 'summarize all relevant defendant background information.' Audit this conclusion and the three claims it's based on against the evidence I'm providing — flag any drift from the original task or stale premises, importance medium."},{"title":"Product recommendation freshness audit","prompt":"My shopping agent concluded 'the Sony WH-1000XM5 is the best noise-cancelling headphone under $350' — the original task was 'find the best noise-cancelling headphones available today under $350.' Please audit these four claims against the supplied reviews and check for stale premises or anything that contradicts the evidence; set importance to high and freshness expectation to 'data should be from within the last 6 months.'"}],"resultDescription":"Returns a per-claim audit report identifying which claims are supported, which contain unsupported assumptions, which contradict the supplied evidence, which rely on stale premises, and whether the overall conclusion has drifted from the original task.","failureModes":["Missing required fields (original_task or current_conclusion) returns a validation error","Claims array exceeds 12 items returns a schema rejection","Evidence items exceed 20 or individual text exceeds 1200 chars returns a validation error","Ambiguous or vague original_task may produce low-confidence audit results","Payment failure via x402 returns a 402 status before processing begins"],"whenToPreferThis":"Choose this endpoint when you need a structured, per-claim audit of an agent's conclusion against explicitly supplied evidence — especially when you want to catch task drift, stale premises, or unsupported assumptions in a single fresh-context pass. Prefer it over general LLM self-review when you need a deterministic, auditable record of which specific claims pass or fail against your evidence set.","instructions":null,"reviewSummary":null,"reviewSummaryHighlights":null,"reviewSummaryConcerns":null,"reviewSummaryGeneratedAt":null,"activationCount":0,"lastUsedAt":null,"lastSuccessfullyRanAt":null,"lastHealthCheckAt":"2026-09-15T12:39:39.655Z","isFirstParty":false}