{"uid":"cap_tLy5xwyg98meigt_VkazO","slug":"weblens-pdf-extraction-api-e0eace71","name":"WebLens PDF Extraction API","description":"# WebLens - Web Intelligence API, pay per call\n\nScrape, crawl, map and extract the web. No account, no API key, no monthly\nminimum — you pay for the calls you make and nothing else.\n\n## Pricing\nPage fetching starts at **$0.002**, whole-site crawling at\n**$0.0015/page**, and sitemap discovery at **$0.004**.\nComparable services bill $0.007-0.008 per request, or reach a lower per-page\nrate only on a $99/month commitment. WebLens has no commitment to reach.\n\n## Payment Protocol\nAll paid endpoints use the [x402 protocol](https://x402.org) for HTTP-native\nmicropayments (USDC on Base).\n\n## Cache Discount\nCached responses are **70% cheaper** than fresh fetches.","url":"https://api.weblens.dev/pdf?utm_source=zero.xyz","method":"POST","headers":{},"bodySchema":{"type":"object","properties":{"url":{"type":"string","description":"URL of the PDF document"},"pages":{"type":"array","description":"Specific page numbers to extract (omit for all pages)"}}},"responseSchema":{"type":"json","example":{"url":"https://example.com/document.pdf","pages":[{"content":"Page 1 text content...","pageNumber":1}],"fullText":"Page 1 text content...","metadata":{"title":"Sample Document","author":"John Doe","pageCount":10},"requestId":"req_pdf123","extractedAt":"2026-01-26T12:00:00.000Z"}},"example":null,"exampleRequest":null,"tags":["x402"],"displayCostAmount":"0.004","displayCostAsset":"USDC","priceDynamic":false,"priceHint":null,"priceStatus":"priced","priceSource":"probe","requiresHandshake":false,"reviewCount":0,"rating":{"score":"0.00","successRate":"0.00","reviews":0,"stars":null,"state":"unrated"},"availabilityStatus":"unknown","priceObserved":null,"sessionDeposit":null,"pricing":{"kind":"static","summary":"$0.004/call","primary":{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.004","per":"call","confidence":"exact"},"accepted":[{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.004","per":"call","confidence":"exact"}]},"paymentMethods":[{"uid":"pm_YN2HN32xehDmiEhd6YbdK","protocol":"x402","methodType":"crypto","chain":"base","mode":"charge","costAmount":"0.004","costPer":"request","priority":0,"asset":"0x833589fCD6eDb6E08f4c7C32D4f71b54bdA02913","unit":"request","depositMicros":null,"planRef":null}],"brandName":null,"brandSlug":null,"brandBaseUrl":null,"brandDocsUrl":null,"whatItDoes":"Extracts full text, page-by-page content, and metadata from a PDF document hosted at a given URL","exampleAgentPrompt":"Can you extract all the text from this PDF — https://example.com/report.pdf — and tell me the title, author, and what's on each page?","exampleUseCases":[{"title":"Research paper content extraction","prompt":"Pull the full text and metadata from this academic paper PDF: https://arxiv.org/pdf/2301.00001.pdf — I need the title, author, page count, and the content of each page."},{"title":"Legal document text retrieval","prompt":"Read this contract PDF at https://example.com/contracts/agreement.pdf and extract all the text so I can search through it."},{"title":"Financial report data extraction","prompt":"Extract everything from this annual report PDF at https://company.com/reports/2024-annual.pdf — I need the full text and which page each section appears on."}],"resultDescription":"A JSON object containing the PDF URL, an array of page objects (each with pageNumber and content), the full concatenated text, document metadata (title, author, page count), a unique request ID, and an ISO timestamp of when the extraction occurred.","failureModes":["URL is not reachable or returns a non-200 status — extraction fails with an error","URL does not point to a valid PDF — parsing error returned","PDF is password-protected or DRM-restricted — content cannot be extracted","Very large PDFs may time out or return partial results","Payment failure via x402 protocol blocks the request"],"whenToPreferThis":"Choose this endpoint when you need to programmatically read the contents of a PDF hosted at a public URL, especially when you also need per-page structure and document metadata (title, author, page count). Prefer it over generic web scraping endpoints when the target is specifically a PDF file. The caching discount makes it cost-effective for repeated access to the same document.","instructions":null,"reviewSummary":null,"reviewSummaryHighlights":null,"reviewSummaryConcerns":null,"reviewSummaryGeneratedAt":null,"activationCount":0,"lastUsedAt":null,"lastSuccessfullyRanAt":null,"lastHealthCheckAt":"2026-10-02T01:20:51.706Z","isFirstParty":false,"canonicalSlug":"weblens-pdf-extraction-api-e0eace71"}