# Aayat AI Web Page Metadata Extractor

> Aayat AI Web Page Metadata Extractor is a paid API for AI agents from aayatai.com, paid per call via x402, $0.002/call, status unknown (last checked 2026-10-02).

Fetches and returns comprehensive metadata for any public web page, including title, description, OpenGraph tags, Twitter card, JSON-LD types, RSS/Atom feeds, icons, and more in a single call.

## Facts

- Endpoint: GET https://aayatai.com/metadata?utm_source=zero.xyz
- Price: $0.002/call
- Payment: x402
- Status: unknown
- Last checked: 2026-10-02
- Activations on Zero: 0
- Tags: x402
- Canonical page: https://www.zero.xyz/c/aayat-ai-web-page-metadata-extractor-6c775c42
- Structured record (JSON): https://api.zero.xyz/v1/capabilities/cap_I7g9K2otCJsLZfrqxIPL-

Status and success rate cover calls made through Zero and Zero's own probes. Third-party monitors may report differently.

## How to call it through Zero

Zero handles the 402 payment challenge and records the run. With the Zero CLI installed (`npm i -g @zeroxyz/cli`):

```sh
zero fetch --capability aayat-ai-web-page-metadata-extractor-6c775c42
```

Example prompt: Can you pull all the metadata for https://techcrunch.com/2024/05/01/openai-gpt5/ — I need the title, description, OpenGraph image, canonical URL, author, publish date, and any RSS feeds the page links to.

## When to prefer this

Choose this endpoint when you need a comprehensive, single-call extraction of all standard web metadata for a URL — including OpenGraph, Twitter cards, JSON-LD, feeds, and icons — without running your own crawler. It is ideal for link preview generation, SEO audits, feed discovery, and content enrichment pipelines. It is more cost-effective and simpler than running a full headless browser scrape when static HTML metadata is sufficient.

## Known failure modes

- URL is unreachable or returns a non-2xx status — status field reflects HTTP error code
- Target page is blocked by robots.txt — endpoint respects robots directives and may return limited or no data
- Malformed or non-URI input for the url parameter — returns a validation error
- Page is JavaScript-rendered only — static metadata may be missing if the page requires JS execution
- Rate limiting or network timeouts on the remote server — may result in partial or empty metadata

## How this service works

A web page's metadata in one cheap call: title, description, canonical URL, language, site name, preview image, OpenGraph and Twitter card tags, author and dates, icons, RSS/Atom feeds and JSON-LD (schema.org) types. For link previews, SEO checks and crawlers. robots.txt respected. ?url=https://example.com

## Output

A JSON object containing the final resolved URL, HTTP status, page title, meta description, canonical URL, language, site name, author, published and modified dates, robots directive, preview image, OpenGraph key-value pairs, Twitter card key-value pairs, an array of feed objects (URL, type, title), an array of icon URLs, an array of JSON-LD @type strings, and the detected generator/CMS. All string fields may be null if absent on the page.

## Request schema (JSON Schema)

```json
{
 "type": "object",
 "$schema": "https://json-schema.org/draft/2020-12/schema",
 "required": [
  "input"
 ],
 "properties": {
  "input": {
   "type": "object",
   "required": [
    "type",
    "method"
   ],
   "properties": {
    "type": {
     "type": "string",
     "const": "http"
    },
    "method": {
     "enum": [
      "GET"
     ],
     "type": "string"
    },
    "queryParams": {
     "type": "object",
     "required": [
      "url"
     ],
     "properties": {
      "url": {
       "type": "string",
       "format": "uri",
       "maxLength": 2048,
       "description": "The page address."
      }
     }
    }
   },
   "additionalProperties": false
  },
  "output": {
   "type": "object",
   "required": [
    "type"
   ],
   "properties": {
    "type": {
     "type": "string"
    },
    "example": {
     "type": "object",
     "required": [
      "url",
      "title",
      "description",
      "openGraph",
      "feeds"
     ],
     "properties": {
      "url": {
       "type": "string",
       "description": "Final address after redirects."
      },
      "type": {
       "type": [
        "string",
        "null"
       ]
      },
      "feeds": {
       "type": "array",
       "items": {
        "type": "object"
       }
      },
      "icons": {
       "type": "array",
       "items": {
        "type": "string"
       }
      },
      "image": {
       "type": [
        "string",
        "null"
       ]
      },
      "title": {
       "type": [
        "string",
        "null"
       ]
      },
      "trust": {
       "type": "object",
       "description": "Third-party text, cleaned: read trust.notice; removed = what we stripped."
      },
      "author": {
       "type": [
        "string",
        "null"
       ]
      },
      "robots": {
       "type": [
        "string",
        "null"
       ],
       "description": "The page's robots meta tag (e.g. noindex)."
      },
      "status": {
       "type": [
        "integer",
        "null"
       ]
      },
      "twitter": {
       "type": "object"
      },
      "language": {
       "type": [
        "string",
        "null"
       ]
      },
      "siteName": {
       "type": [
        "string",
        "null"
       ]
      },
      "canonical": {
       "type": [
        "string",
        "null"
       ]
      },
      "generator": {
       "type": [
        "string",
        "null"
       ]
      },
      "openGraph": {
       "type": "object"
      },
      "modifiedAt": {
     
… (truncated)
```

## Response schema (JSON Schema)

```json
{
 "type": "json",
 "example": {
  "url": "https://example.org/",
  "type": "website",
  "feeds": [
   {
    "url": "https://example.org/feed.xml",
    "type": "application/rss+xml",
    "title": "Blog"
   }
  ],
  "icons": [
   "https://example.org/favicon.ico"
  ],
  "image": "https://example.org/og.png",
  "title": "Example Org",
  "author": null,
  "robots": null,
  "status": 200,
  "twitter": {
   "twitter:card": "summary_large_image"
  },
  "language": "en",
  "siteName": "Example",
  "canonical": "https://example.org/",
  "generator": null,
  "openGraph": {
   "og:image": "https://example.org/og.png",
   "og:title": "Example Org"
  },
  "modifiedAt": null,
  "description": "We make examples.",
  "jsonLdTypes": [
   "Organization"
  ],
  "publishedAt": null
 }
}
```

## More

- Live health (JSON, refreshed every minute): https://www.zero.xyz/c/aayat-ai-web-page-metadata-extractor-6c775c42/health.json
- [Zero catalog index](https://www.zero.xyz/llms.txt)
- [Other services from aayatai.com](https://www.zero.xyz/host/aayatai.com/llms.txt)
