{"uid":"cap_h2CfIQLZ92QvRuq5qtP9X","slug":"bulk-webpage-text-d83bb1f4","name":"bulk-webpage-text","description":"Bulk web page text extraction: up to 5 URLs in one $0.01 call. For each page returns the readable main text (navigation, ads and scripts removed), title, meta description, final URL, status and word count. Batch content extraction for RAG pipelines, research agents, summarization and crawling at a fifth of the per-page price.","url":"https://intel.rallylive.ca/page-text-bulk","method":"GET","headers":{},"bodySchema":{"type":"object","$schema":"https://json-schema.org/draft/2020-12/schema","required":["input"],"properties":{"input":{"type":"object","required":["type","method"],"properties":{"type":{"type":"string","const":"http"},"method":{"enum":["GET"],"type":"string"},"queryParams":{"type":"object","properties":{}}},"additionalProperties":false},"output":{"type":"object","required":["type"],"properties":{"type":{"type":"string"},"example":{"type":"object"}}}}},"responseSchema":null,"example":null,"exampleRequest":null,"tags":["x402"],"displayCostAmount":"0.01","displayCostAsset":"USDC","priceDynamic":false,"priceHint":null,"priceStatus":"priced","priceSource":"probe","requiresHandshake":false,"reviewCount":0,"rating":{"score":"0.00","successRate":"0.00","reviews":0,"stars":null,"state":"unrated"},"availabilityStatus":"unknown","priceObserved":null,"sessionDeposit":null,"pricing":{"kind":"static","summary":"$0.01/call","primary":{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.01","per":"call","confidence":"exact"},"accepted":[{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.01","per":"call","confidence":"exact"}]},"paymentMethods":[{"uid":"pm_JGpNodvZ6yYAcmpSSdKd8","protocol":"x402","methodType":"crypto","chain":"base","mode":"charge","costAmount":"0.01","costPer":"request","priority":0,"asset":"0x833589fCD6eDb6E08f4c7C32D4f71b54bdA02913","unit":"request","depositMicros":null,"planRef":null}],"brandName":null,"brandSlug":null,"brandBaseUrl":null,"brandDocsUrl":null,"whatItDoes":"Extracts readable main text, title, meta description, final URL, status, and word count from up to 5 URLs in a single batch call.","exampleAgentPrompt":"Pull the readable main text from these five URLs for me — strip out the navigation, ads, and scripts — and give me the title, description, word count, and final URL for each: https://example.com/article1, https://example.com/article2, https://example.com/article3, https://example.com/article4, https://example.com/article5.","exampleUseCases":[{"title":"RAG pipeline content ingestion","prompt":"I need to feed these five blog post URLs into my knowledge base — can you extract the clean readable text from each one, no ads or nav menus, and give me the title, word count, and final URL for each so I can index them?"},{"title":"Competitive research content gathering","prompt":"Grab the main article text from these five competitor pages for me: [url1], [url2], [url3], [url4], [url5] — I want clean readable content with titles and descriptions so I can summarize what each one is saying."},{"title":"Batch summarization for newsletter","prompt":"I have five news articles I want summarized — can you fetch the full readable text from each of these URLs and return the title and main body so I can run them through a summarizer? Here are the links: [url1], [url2], [url3], [url4], [url5]."}],"resultDescription":"For each submitted URL, the endpoint returns the cleaned readable main text (navigation, ads, and scripts removed), the page title, meta description, the final resolved URL (after any redirects), the HTTP status code, and the word count of the extracted content.","failureModes":["URL is unreachable or returns a non-200 status — that page's result will reflect the error status","Page uses heavy JavaScript rendering that cannot be executed server-side, resulting in minimal or no text extraction","One or more URLs redirect to a paywall or login page, returning little usable content","Submitting more than 5 URLs may result in an error or truncation","Malformed URLs cause individual page failures while others succeed"],"whenToPreferThis":"Choose this endpoint when you need to extract readable text from multiple web pages in a single API call at reduced cost per page. It is ideal for RAG pipeline ingestion, bulk research, and summarization workflows where you have 2–5 URLs to process simultaneously and want cleaned, ad-free main content rather than raw HTML. Prefer it over single-page extractors when batch efficiency and cost matter.","instructions":null,"reviewSummary":null,"reviewSummaryHighlights":null,"reviewSummaryConcerns":null,"reviewSummaryGeneratedAt":null,"activationCount":0,"lastUsedAt":null,"lastSuccessfullyRanAt":null,"lastHealthCheckAt":"2026-09-14T13:06:05.553Z","isFirstParty":false}