# Aayat AI Web Page Extractor

> Aayat AI Web Page Extractor is a paid API for AI agents from aayatai.com, paid per call via x402, $0.005/call, status unknown (last checked 2026-10-02).

Fetches any public web page and returns clean Markdown content, the page title, and all links as structured JSON

## Facts

- Endpoint: GET https://aayatai.com/extract?utm_source=zero.xyz
- Price: $0.005/call
- Payment: x402
- Status: unknown
- Last checked: 2026-10-02
- Activations on Zero: 0
- Tags: x402
- Canonical page: https://www.zero.xyz/c/aayat-ai-web-page-extractor-25901259
- Structured record (JSON): https://api.zero.xyz/v1/capabilities/cap_HfHhADaX8jKDX4iLCVhxi

Status and success rate cover calls made through Zero and Zero's own probes. Third-party monitors may report differently.

## How to call it through Zero

Zero handles the 402 payment challenge and records the run. With the Zero CLI installed (`npm i -g @zeroxyz/cli`):

```sh
zero fetch --capability aayat-ai-web-page-extractor-25901259
```

Example prompt: Can you fetch the content of https://openai.com/blog and give me the main text as clean markdown along with all the links on the page?

## When to prefer this

Choose this endpoint when an AI agent needs to read and process the textual content of a specific public web page — especially when clean, clutter-free markdown is needed for downstream LLM consumption. Ideal for single-page reads of articles, documentation, blogs, or product pages. Prefer this over raw HTTP fetching because it automatically strips navigation, scripts, styles, and ads. Not suitable for crawling entire sites, authenticated pages, or private/internal URLs.

## Known failure modes

- Private or internal IP addresses are blocked and return an error
- Pages exceeding 2 MB are rejected and not charged
- Pages taking longer than 10 seconds to load are rejected and not charged
- Invalid or malformed URLs return a validation error
- Non-public pages requiring authentication cannot be accessed
- Paywalled or login-gated content may return partial or empty markdown

## How this service works

Web page extractor for AI agents: fetch any public URL and get clean Markdown, the page title and all links as JSON, with scripts, styles and menus removed. url is the page address (https://...). Private or internal addresses are blocked; pages over 2 MB or slower than 10 seconds are rejected and you are not charged.

## Output

Returns a JSON object containing: the final URL after any redirects, the page title, the main body content converted to clean Markdown (scripts, styles, menus, and clutter stripped), and an array of all links found on the page (each with href and anchor text).

## Request schema (JSON Schema)

```json
{
 "type": "object",
 "$schema": "https://json-schema.org/draft/2020-12/schema",
 "required": [
  "input"
 ],
 "properties": {
  "input": {
   "type": "object",
   "required": [
    "type",
    "method"
   ],
   "properties": {
    "type": {
     "type": "string",
     "const": "http"
    },
    "method": {
     "enum": [
      "GET"
     ],
     "type": "string"
    },
    "queryParams": {
     "type": "object",
     "required": [
      "url"
     ],
     "properties": {
      "url": {
       "type": "string",
       "format": "uri",
       "maxLength": 2048,
       "description": "Full http(s) address of a public web page."
      }
     }
    }
   },
   "additionalProperties": false
  },
  "output": {
   "type": "object",
   "required": [
    "type"
   ],
   "properties": {
    "type": {
     "type": "string"
    },
    "example": {
     "type": "object",
     "required": [
      "url",
      "title",
      "markdown",
      "links"
     ],
     "properties": {
      "url": {
       "type": "string",
       "description": "Final address after redirects."
      },
      "links": {
       "type": "array",
       "items": {
        "type": "object",
        "required": [
         "text",
         "href"
        ],
        "properties": {
         "href": {
          "type": "string"
         },
         "text": {
          "type": "string"
         }
        }
       }
      },
      "title": {
       "type": "string"
      },
      "trust": {
       "type": "object",
       "description": "Third-party text, cleaned: read trust.notice; removed = what we stripped."
      },
      "markdown": {
       "type": "string",
       "description": "Main page content as Markdown."
      }
     }
    }
   }
  }
 }
}
```

## Response schema (JSON Schema)

```json
{
 "type": "json",
 "example": {
  "url": "https://example.com/",
  "links": [
   {
    "href": "https://www.iana.org/domains/example",
    "text": "More information..."
   }
  ],
  "title": "Example Domain",
  "markdown": "# Example Domain\n\nThis domain is for use in illustrative examples in documents..."
 }
}
```

## More

- Live health (JSON, refreshed every minute): https://www.zero.xyz/c/aayat-ai-web-page-extractor-25901259/health.json
- [Zero catalog index](https://www.zero.xyz/llms.txt)
- [Other services from aayatai.com](https://www.zero.xyz/host/aayatai.com/llms.txt)
