{"uid":"cap_Zwn0pJLhT36zSgAhdcLbB","slug":"date-extraction-from-web-page-f77a240e","name":"Date Extraction from Web Page","description":"Date extraction from a page: the published and modified dates from meta tags and schema.org, plus dates found in the text (ISO, US and long forms), the earliest and latest, and a best guess at the publication date. Content freshness and timeline building. $0.01 per page.","url":"https://intel.rallylive.ca/site/dates","method":"GET","headers":{},"bodySchema":{"type":"object","$schema":"https://json-schema.org/draft/2020-12/schema","required":["input"],"properties":{"input":{"type":"object","required":["type","method"],"properties":{"type":{"type":"string","const":"http"},"method":{"enum":["GET"],"type":"string"},"queryParams":{"type":"object","properties":{}}},"additionalProperties":false},"output":{"type":"object","required":["type"],"properties":{"type":{"type":"string"},"example":{"type":"object"}}}}},"responseSchema":null,"example":null,"exampleRequest":null,"tags":["x402"],"displayCostAmount":"0.01","displayCostAsset":"USDC","priceDynamic":false,"priceHint":null,"priceStatus":"priced","priceSource":"probe","requiresHandshake":false,"reviewCount":0,"rating":{"score":"0.00","successRate":"0.00","reviews":0,"stars":null,"state":"unrated"},"availabilityStatus":"unknown","priceObserved":null,"sessionDeposit":null,"pricing":{"kind":"static","summary":"$0.01/call","primary":{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.01","per":"call","confidence":"exact"},"accepted":[{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.01","per":"call","confidence":"exact"}]},"paymentMethods":[{"uid":"pm_HHvHA8nAdG7kUhJBC95Lk","protocol":"x402","methodType":"crypto","chain":"base","mode":"charge","costAmount":"0.01","costPer":"request","priority":0,"asset":"0x833589fCD6eDb6E08f4c7C32D4f71b54bdA02913","unit":"request","depositMicros":null,"planRef":null}],"brandName":null,"brandSlug":null,"brandBaseUrl":null,"brandDocsUrl":null,"whatItDoes":"Extracts published and modified dates from a web page's meta tags, schema.org markup, and body text, returning all found dates plus earliest, latest, and a best-guess publication date.","exampleAgentPrompt":"Can you pull all the dates from this article page — I want to know when it was published, when it was last modified, and the best guess at the actual publication date: https://example.com/article/some-news-story","exampleUseCases":[{"title":"Verifying article publication date","prompt":"I found this news article and I want to confirm when it was actually published — can you extract the publication date from the page, including what's in the meta tags and schema.org: https://www.newssite.com/article/climate-report-2024"},{"title":"Content freshness check for research pipeline","prompt":"I need to know how fresh this page is — pull the published and modified dates from https://docs.example.com/api/overview so I can tell if the documentation is current or stale."},{"title":"Building a timeline from a blog post","prompt":"Can you find all the dates mentioned on this blog post — ISO formats, US dates, long-form dates, earliest and latest — so I can map out the timeline of events described: https://blog.example.com/history-of-our-product"}],"resultDescription":"Returns published and modified dates extracted from meta tags and schema.org markup, a list of all dates found in the page body (ISO, US, and long-form formats), the earliest and latest dates found, and a best-guess publication date inferred from all available signals.","failureModes":["Page has no date signals in meta tags, schema.org, or body text — returns empty or null date fields","URL is unreachable or returns a non-200 HTTP status — extraction fails with error","Page requires JavaScript rendering and dates are only in dynamically loaded content — may miss dates","Ambiguous date formats lead to incorrect parsing or wrong best-guess publication date","Paywalled or bot-blocked pages may not return full content for date scanning"],"whenToPreferThis":"Choose this endpoint when you need to determine when a web page was published or last modified, especially when combining signals from meta tags, schema.org, and body text gives a more reliable answer than any single source. Ideal for content freshness checks, fact-checking pipelines, research timelines, and indexing workflows where publication date accuracy matters. Prefer this over generic scrapers when you specifically need structured temporal metadata rather than full page content.","instructions":null,"reviewSummary":null,"reviewSummaryHighlights":null,"reviewSummaryConcerns":null,"reviewSummaryGeneratedAt":null,"activationCount":0,"lastUsedAt":null,"lastSuccessfullyRanAt":null,"lastHealthCheckAt":"2026-09-14T13:13:22.916Z","isFirstParty":false}