{"uid":"cap_V-g_ts2DWBdEyrf370LrU","slug":"website-scraper-url-to-clean-markdown-2899b9bc","name":"Website Scraper – URL to Clean Markdown","description":"Scrape a website page: fetches the URL and returns the page content as clean Markdown plus title, description, headings, word count and all links (URL and anchor text). Web scraping for agents without a browser, API key or proxy; handles redirects and strips navigation, ads and scripts. $0.01 per page.","url":"https://intel.rallylive.ca/scrape","method":"GET","headers":{},"bodySchema":{"type":"object","$schema":"https://json-schema.org/draft/2020-12/schema","required":["input"],"properties":{"input":{"type":"object","required":["type","method"],"properties":{"type":{"type":"string","const":"http"},"method":{"enum":["GET"],"type":"string"},"queryParams":{"type":"object","properties":{}}},"additionalProperties":false},"output":{"type":"object","required":["type"],"properties":{"type":{"type":"string"},"example":{"type":"object"}}}}},"responseSchema":null,"example":null,"exampleRequest":null,"tags":["x402"],"displayCostAmount":"0.05","displayCostAsset":"USDC","priceDynamic":false,"priceHint":null,"priceStatus":"priced","priceSource":"probe","requiresHandshake":false,"reviewCount":0,"rating":{"score":"0.00","successRate":"0.00","reviews":0,"stars":null,"state":"unrated"},"availabilityStatus":"unknown","priceObserved":null,"sessionDeposit":null,"pricing":{"kind":"static","summary":"$0.05/call","primary":{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.05","per":"call","confidence":"exact"},"accepted":[{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.05","per":"call","confidence":"exact"}]},"paymentMethods":[{"uid":"pm_EXLmYZKqfFZ7PR04klXom","protocol":"x402","methodType":"crypto","chain":"base","mode":"charge","costAmount":"0.05","costPer":"request","priority":0,"asset":"0x833589fCD6eDb6E08f4c7C32D4f71b54bdA02913","unit":"request","depositMicros":null,"planRef":null}],"brandName":null,"brandSlug":null,"brandBaseUrl":null,"brandDocsUrl":null,"whatItDoes":"Fetches any web page and returns its content as clean Markdown along with title, description, headings, word count, and all extracted links.","exampleAgentPrompt":"Can you scrape the page at https://example.com/blog/post-1 and give me the clean readable text as Markdown, along with its title, any headings, and all the links on the page?","exampleUseCases":[{"title":"Competitor landing page analysis","prompt":"Scrape https://competitor.com/pricing and pull out the clean Markdown text, all the headings, and every link on the page so I can analyze their offer."},{"title":"Article summarization from URL","prompt":"Fetch the article at https://techcrunch.com/2024/05/01/ai-funding-roundup and give me the readable Markdown content and word count — I want to summarize it."},{"title":"Automated link audit for SEO","prompt":"Pull all the links and anchor text from https://mysite.com/resources so I can audit which internal and external pages we're pointing to."}],"resultDescription":"Returns clean Markdown of the page body (navigation, ads, and scripts stripped), plus the page title, meta description, list of headings, total word count, and all hyperlinks found on the page with their URL and anchor text.","failureModes":["URL is unreachable or returns a non-200 HTTP status — endpoint will report the failure","Page is JavaScript-rendered (SPA) and content is not in the initial HTML response — may return minimal or empty content","URL redirects in a loop or to a blocked resource — redirect handling may fail","Page requires login or session cookie — returns login page content instead of target content","Rate limits or bot-blocking on the target site — request may be rejected or return a CAPTCHA page"],"whenToPreferThis":"Choose this endpoint when you need clean, readable text and structural metadata (headings, links, word count) from any publicly accessible web page without running a browser, managing proxies, or holding an API key for a dedicated scraping service. It is especially useful for AI agents that need to read page content, extract links, or convert HTML to Markdown on the fly at $0.01 per page. Prefer alternatives if you need JavaScript rendering of SPAs or need to interact with authenticated sessions.","instructions":null,"reviewSummary":null,"reviewSummaryHighlights":null,"reviewSummaryConcerns":null,"reviewSummaryGeneratedAt":null,"activationCount":0,"lastUsedAt":null,"lastSuccessfullyRanAt":null,"lastHealthCheckAt":"2026-09-14T13:09:36.176Z","isFirstParty":false}