# x402 Web Page to Markdown Scraper

> x402 Web Page to Markdown Scraper is a paid API for AI agents from x402-data-api.onrender.com, paid per call via x402, $0.001/call, status unknown (last checked 2026-10-02).

Fetches any public web page URL and returns clean, readable Markdown with ads, navigation, and boilerplate stripped out

## Facts

- Endpoint: GET https://x402-data-api.onrender.com/api/scrape?utm_source=zero.xyz
- Price: $0.001/call
- Payment: x402
- Status: unknown
- Last checked: 2026-10-02
- Activations on Zero: 0
- Tags: x402
- Canonical page: https://www.zero.xyz/c/x402-web-page-to-markdown-scraper-7225a958
- Structured record (JSON): https://api.zero.xyz/v1/capabilities/cap_wdBpy5RqqzszUreMV9XPK

Status and success rate cover calls made through Zero and Zero's own probes. Third-party monitors may report differently.

## How to call it through Zero

Zero handles the 402 payment challenge and records the run. With the Zero CLI installed (`npm i -g @zeroxyz/cli`):

```sh
zero fetch --capability x402-web-page-to-markdown-scraper-7225a958
```

Example prompt: Can you fetch https://en.wikipedia.org/wiki/Large_language_model and give me the main content as clean markdown, with all the ads and navigation stripped out?

## When to prefer this

Use this endpoint when an AI agent needs to read, summarize, or process the text content of a specific public web page and wants it delivered in clean Markdown format without boilerplate, ads, or navigation clutter. Ideal for article ingestion, documentation reading, or any task that requires human-readable page text in a structured format.

## Known failure modes

- Invalid or non-public URL returns an error
- Page behind authentication or paywall cannot be fetched
- Network timeout if the target page is slow or unreachable
- URL that resolves to a non-HTML resource (e.g. PDF, image) may not convert cleanly
- Rate limiting or blocking by the target site may cause fetch failure

## How this service works

Fetch any public web page and return clean, readable Markdown with navigation, ads, and boilerplate stripped. For agents that need article text or documentation as Markdown.

## Output

Returns a JSON object containing the original URL, the page title, the full article/page body as clean Markdown (ads, nav, and boilerplate removed), and a word count integer

## Request schema (JSON Schema)

```json
{
 "type": "object",
 "$schema": "https://json-schema.org/draft/2020-12/schema",
 "required": [
  "input"
 ],
 "properties": {
  "input": {
   "type": "object",
   "required": [
    "type",
    "method"
   ],
   "properties": {
    "type": {
     "type": "string",
     "const": "http"
    },
    "method": {
     "enum": [
      "GET"
     ],
     "type": "string"
    },
    "queryParams": {
     "type": "object",
     "required": [
      "url"
     ],
     "properties": {
      "url": {
       "type": "string",
       "description": "Public http(s) URL of the page to fetch and convert to markdown"
      }
     }
    }
   },
   "additionalProperties": false
  },
  "output": {
   "type": "object",
   "required": [
    "type"
   ],
   "properties": {
    "type": {
     "type": "string"
    },
    "example": {
     "type": "object",
     "properties": {
      "url": {
       "type": "string"
      },
      "title": {
       "type": "string"
      },
      "markdown": {
       "type": "string"
      },
      "word_count": {
       "type": "number"
      }
     }
    }
   }
  }
 }
}
```

## More

- Live health (JSON, refreshed every minute): https://www.zero.xyz/c/x402-web-page-to-markdown-scraper-7225a958/health.json
- [Zero catalog index](https://www.zero.xyz/llms.txt)
- [Other services from x402-data-api.onrender.com](https://www.zero.xyz/host/x402-data-api.onrender.com/llms.txt)
