{"uid":"cap_mhdJ81JtsilA5YkliMSS-","slug":"vextorium-web-markdown-extractor-cc4038d0","name":"Vextorium Web Markdown Extractor","description":"Read any public web page or PDF and get its main content as clean, LLM-ready markdown (headings, lists, links, tables, code) plus metadata: title, description, author, date, language, canonical URL. Menus, cookie banners and footers removed. Optional list of links. SSRF-safe fetching.","url":"https://api.vextorium.com/web-markdown-pago?utm_source=zero.xyz","method":"POST","headers":{},"bodySchema":{"type":"object","$schema":"https://json-schema.org/draft/2020-12/schema","required":["input"],"properties":{"input":{"type":"object","required":["type","method","bodyType","body"],"properties":{"body":{"required":["url"],"properties":{"url":{"type":"string","description":"Full http(s) URL of a web page or PDF"},"max_caracteres":{"type":"integer","maximum":100000,"minimum":500,"description":"Maximum markdown length (500-100000, default 20000)"},"incluir_enlaces":{"type":"boolean","description":"Also return the list of links found (default false)"}}},"type":{"type":"string","const":"http"},"method":{"enum":["POST"],"type":"string"},"bodyType":{"enum":["json","form-data","text"],"type":"string"}},"additionalProperties":false},"output":{"type":"object","required":["type"],"properties":{"type":{"type":"string"},"example":{"type":["object","null"],"properties":{"url":{"type":["string","null"]},"tipo":{"type":["string","null"]},"avisos":{"type":["array","null"],"items":{"type":["null","string"]}},"markdown":{"type":["string","null"]},"palabras":{"type":["number","null"]},"metadatos":{"type":["object","null"],"properties":{"sitio":{"type":["string","null"]},"idioma":{"type":["string","null"]},"imagen":{"type":["string","null"]},"titulo":{"type":["string","null"]},"canonica":{"type":["string","null"]},"descripcion":{"type":["string","null"]}}},"recortado":{"type":["boolean","null"]},"url_final":{"type":["string","null"]},"caracteres":{"type":["number","null"]}}}}}}},"responseSchema":{"type":"json","example":{"url":"https://docs.x402.org/introduction","tipo":"html","avisos":["Contenido recortado a 1200 caracteres (de 3590). Sube max_caracteres si necesitas más."],"markdown":"Welcome\n\n# Welcome to x402\n\nThis guide will help you understand x402, the open payment standard, and help you get started building or integrating services with x402.\n\nx402 is the open payment standard that enables services to charge for access to their APIs and content directly over HTTP. It is built around the HTTP `402 Payment Required` status code and allows clients to programmatically pay for resources without accounts, sessions, or credential management.\nWith x402, any web service can require payment before serving a response, using crypto-native payments for speed, privacy, and efficiency.\n**Want to contribute to our docs?** [The documentation source in this repository is open to PRs.](https://github.com/x402-foundation/x402) Our only ask is that you keep these docs as a neutral resource, with no branded content other than linking out to other resources where appropriate.\n**Note about the docs:** These docs are the credibly neutral source of truth for x402, as x402 is a completely open standard under the Apache-2.0 license.\n\n### ​ Why Use x402?\n\nx402 offers:\n\n- **No fees and minimal friction** x402 as a standard has 0 fees built in.\n\n- **Native support for machine-to-machine ","palabras":188,"metadatos":{"sitio":"x402","idioma":"en","imagen":"https://coinbase-5dac824f.mintlify.app/mintlify-assets/_next/image?url=%2F_mintlify%2Fapi%2Fog%3Fdivision%3DWelcome%26title%3DWelcome%2Bto%2Bx402%26description%3DThis%2Bguide%2Bwill%2Bhelp%2Byou%2Bunderstand%2Bx402%252C%2Bthe%2Bopen%2Bpayment%2Bstandard%252C%2Band%2Bhelp%2Byou%2Bget%2Bstarted%2Bbuilding%2Bor%2Bintegrating%2Bservices%2Bwith%2Bx402.%26theme%3D7f62360b0321f9ed397b7eff&w=1200&q=100","titulo":"Welcome to x402 - x402","canonica":"https://docs.x402.org/introduction","descripcion":"This guide will help you understand x402, the open payment standard, and help you get started building or integrating services with x402."},"recortado":true,"url_final":"https://docs.x402.org/introduction","caracteres":1200}},"example":null,"exampleRequest":null,"tags":["x402"],"displayCostAmount":"0.0045","displayCostAsset":"USDC","priceDynamic":false,"priceHint":null,"priceStatus":"priced","priceSource":"registry","requiresHandshake":false,"reviewCount":0,"rating":{"score":"0.00","successRate":"0.00","reviews":0,"stars":null,"state":"unrated"},"availabilityStatus":"unknown","priceObserved":null,"sessionDeposit":null,"pricing":{"kind":"static","summary":"$0.0045/call","primary":{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.0045","per":"call","confidence":"exact"},"accepted":[{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.0045","per":"call","confidence":"exact"}]},"paymentMethods":[{"uid":"pm_GGskQe6_kOeMAxGBsPkap","protocol":"x402","methodType":"crypto","chain":"base","mode":"charge","costAmount":"0.0045","costPer":"request","priority":0,"asset":null,"unit":"request","depositMicros":null,"planRef":null}],"brandName":null,"brandSlug":null,"brandBaseUrl":null,"brandDocsUrl":null,"whatItDoes":"Fetches any public web page or PDF by URL and returns its main content as clean, LLM-ready markdown plus metadata (title, description, author, date, language, canonical URL), with boilerplate removed.","exampleAgentPrompt":"Can you fetch the article at https://example.com/news/article-123 and give me its main content as clean markdown — strip out the menus, ads, and footers — and also return the title, author, and publication date?","exampleUseCases":[{"title":"Research assistant reading news articles","prompt":"Pull the full article from https://techcrunch.com/2024/05/10/openai-latest-news/ and give me the clean markdown text with the title, author, and date — no ads or nav menus."},{"title":"PDF report ingestion for LLM pipeline","prompt":"Fetch this PDF at https://example.com/reports/annual-report-2023.pdf and return its content as markdown so I can feed it into my AI model — limit it to 50000 characters and include any links found in the document."},{"title":"Competitive intelligence scraping","prompt":"Read the page at https://competitor.com/pricing and extract just the main content as clean markdown — I need the headings, tables, and text but none of the cookie banners or footers."}],"resultDescription":"Returns a JSON object with: markdown-formatted main content (headings, lists, tables, code, links), word count, character count, whether the content was truncated, the final resolved URL, content type, metadata object (title, description, canonical URL, language, site name, featured image), optional array of links found on the page, and any warnings encountered during fetch.","failureModes":["URL is not publicly accessible or requires authentication — returns error or empty content","SSRF-blocked URL (private IPs, localhost) — rejected by safety filter","PDF parsing fails for scanned/image-only PDFs with no text layer","Content too long and truncated at max_caracteres limit — recortado flag set to true","Timeout on slow or unresponsive servers","Paywalled content returns only teaser text","Invalid URL format — returns validation error"],"whenToPreferThis":"Choose this endpoint when you need clean, boilerplate-free markdown from any public web page or PDF, especially when feeding content into an LLM context window. It is preferable over raw HTTP fetching because it removes menus, footers, and cookie banners automatically, and over general-purpose scrapers because it outputs structured LLM-ready markdown with rich metadata in a single call. Best for article reading, document ingestion, and research pipelines where content quality and format matter.","instructions":null,"reviewSummary":null,"reviewSummaryHighlights":null,"reviewSummaryConcerns":null,"reviewSummaryGeneratedAt":null,"activationCount":0,"lastUsedAt":null,"lastSuccessfullyRanAt":null,"lastHealthCheckAt":"2026-10-02T18:41:49.397Z","isFirstParty":false,"canonicalSlug":"vextorium-web-markdown-extractor-cc4038d0"}