{"uid":"cap_I7g9K2otCJsLZfrqxIPL-","slug":"aayat-ai-web-page-metadata-extractor-6c775c42","name":"Aayat AI Web Page Metadata Extractor","description":"A web page's metadata in one cheap call: title, description, canonical URL, language, site name, preview image, OpenGraph and Twitter card tags, author and dates, icons, RSS/Atom feeds and JSON-LD (schema.org) types. For link previews, SEO checks and crawlers. robots.txt respected. ?url=https://example.com","url":"https://aayatai.com/metadata?utm_source=zero.xyz","method":"GET","headers":{},"bodySchema":{"type":"object","$schema":"https://json-schema.org/draft/2020-12/schema","required":["input"],"properties":{"input":{"type":"object","required":["type","method"],"properties":{"type":{"type":"string","const":"http"},"method":{"enum":["GET"],"type":"string"},"queryParams":{"type":"object","required":["url"],"properties":{"url":{"type":"string","format":"uri","maxLength":2048,"description":"The page address."}}}},"additionalProperties":false},"output":{"type":"object","required":["type"],"properties":{"type":{"type":"string"},"example":{"type":"object","required":["url","title","description","openGraph","feeds"],"properties":{"url":{"type":"string","description":"Final address after redirects."},"type":{"type":["string","null"]},"feeds":{"type":"array","items":{"type":"object"}},"icons":{"type":"array","items":{"type":"string"}},"image":{"type":["string","null"]},"title":{"type":["string","null"]},"trust":{"type":"object","description":"Third-party text, cleaned: read trust.notice; removed = what we stripped."},"author":{"type":["string","null"]},"robots":{"type":["string","null"],"description":"The page's robots meta tag (e.g. noindex)."},"status":{"type":["integer","null"]},"twitter":{"type":"object"},"language":{"type":["string","null"]},"siteName":{"type":["string","null"]},"canonical":{"type":["string","null"]},"generator":{"type":["string","null"]},"openGraph":{"type":"object"},"modifiedAt":{"type":["string","null"]},"description":{"type":["string","null"]},"jsonLdTypes":{"type":"array","items":{"type":"string"}},"publishedAt":{"type":["string","null"]}}}}}}},"responseSchema":{"type":"json","example":{"url":"https://example.org/","type":"website","feeds":[{"url":"https://example.org/feed.xml","type":"application/rss+xml","title":"Blog"}],"icons":["https://example.org/favicon.ico"],"image":"https://example.org/og.png","title":"Example Org","author":null,"robots":null,"status":200,"twitter":{"twitter:card":"summary_large_image"},"language":"en","siteName":"Example","canonical":"https://example.org/","generator":null,"openGraph":{"og:image":"https://example.org/og.png","og:title":"Example Org"},"modifiedAt":null,"description":"We make examples.","jsonLdTypes":["Organization"],"publishedAt":null}},"example":null,"exampleRequest":null,"tags":["x402"],"displayCostAmount":"0.002","displayCostAsset":"USDC","priceDynamic":false,"priceHint":null,"priceStatus":"priced","priceSource":"probe","requiresHandshake":false,"reviewCount":0,"rating":{"score":"0.00","successRate":"0.00","reviews":0,"stars":null,"state":"unrated"},"availabilityStatus":"unknown","priceObserved":null,"sessionDeposit":null,"pricing":{"kind":"static","summary":"$0.002/call","primary":{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.002","per":"call","confidence":"exact"},"accepted":[{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.002","per":"call","confidence":"exact"}]},"paymentMethods":[{"uid":"pm_19nNsfDKpbvqvYnf1a5Tm","protocol":"x402","methodType":"crypto","chain":"base","mode":"charge","costAmount":"0.002","costPer":"request","priority":0,"asset":"0x833589fCD6eDb6E08f4c7C32D4f71b54bdA02913","unit":"request","depositMicros":null,"planRef":null}],"brandName":null,"brandSlug":null,"brandBaseUrl":null,"brandDocsUrl":null,"whatItDoes":"Fetches and returns comprehensive metadata for any public web page, including title, description, OpenGraph tags, Twitter card, JSON-LD types, RSS/Atom feeds, icons, and more in a single call.","exampleAgentPrompt":"Can you pull all the metadata for https://techcrunch.com/2024/05/01/openai-gpt5/ — I need the title, description, OpenGraph image, canonical URL, author, publish date, and any RSS feeds the page links to.","exampleUseCases":[{"title":"Link preview card generation","prompt":"Fetch the metadata for https://github.com/openai/whisper so I can build a rich link preview — I need the title, description, og:image, and site name."},{"title":"SEO audit for a landing page","prompt":"Check the SEO metadata on https://linear.app — give me the title, meta description, canonical URL, robots tag, and whether it has proper OpenGraph and Twitter card tags set up."},{"title":"Discover RSS feed for a blog","prompt":"I want to subscribe to the Stripe blog at https://stripe.com/blog — can you find the RSS or Atom feed URL listed on that page?"}],"resultDescription":"A JSON object containing the final resolved URL, HTTP status, page title, meta description, canonical URL, language, site name, author, published and modified dates, robots directive, preview image, OpenGraph key-value pairs, Twitter card key-value pairs, an array of feed objects (URL, type, title), an array of icon URLs, an array of JSON-LD @type strings, and the detected generator/CMS. All string fields may be null if absent on the page.","failureModes":["URL is unreachable or returns a non-2xx status — status field reflects HTTP error code","Target page is blocked by robots.txt — endpoint respects robots directives and may return limited or no data","Malformed or non-URI input for the url parameter — returns a validation error","Page is JavaScript-rendered only — static metadata may be missing if the page requires JS execution","Rate limiting or network timeouts on the remote server — may result in partial or empty metadata"],"whenToPreferThis":"Choose this endpoint when you need a comprehensive, single-call extraction of all standard web metadata for a URL — including OpenGraph, Twitter cards, JSON-LD, feeds, and icons — without running your own crawler. It is ideal for link preview generation, SEO audits, feed discovery, and content enrichment pipelines. It is more cost-effective and simpler than running a full headless browser scrape when static HTML metadata is sufficient.","instructions":null,"reviewSummary":null,"reviewSummaryHighlights":null,"reviewSummaryConcerns":null,"reviewSummaryGeneratedAt":null,"activationCount":0,"lastUsedAt":null,"lastSuccessfullyRanAt":null,"lastHealthCheckAt":"2026-10-02T00:46:38.407Z","isFirstParty":false,"canonicalSlug":"aayat-ai-web-page-metadata-extractor-6c775c42"}