# dt0ur.online Web Scraper

> dt0ur.online Web Scraper is a paid API for AI agents from dt0ur.online, paid per call via x402, $0.05/call, status unknown (last checked 2026-10-03).

Extracts sanitized, boilerplate-free markdown and clean text from public web URLs, with SSRF protection against private network access.

## Facts

- Endpoint: POST https://dt0ur.online/api/web/scrape?utm_source=zero.xyz
- Price: $0.05/call
- Payment: x402
- Status: unknown
- Last checked: 2026-10-03
- Activations on Zero: 0
- Tags: x402
- Canonical page: https://www.zero.xyz/c/dt0ur-online-web-scraper-76733983
- Structured record (JSON): https://api.zero.xyz/v1/capabilities/cap_tX-YlzMkh3o6qPl1_H6A9

Status and success rate cover calls made through Zero and Zero's own probes. Third-party monitors may report differently.

## How to call it through Zero

Zero handles the 402 payment challenge and records the run. With the Zero CLI installed (`npm i -g @zeroxyz/cli`):

```sh
zero fetch --capability dt0ur-online-web-scraper-76733983 -d '<json body>'
```

Example prompt: Can you scrape this article for me — https://www.theverge.com/2024/5/1/example-article — and give me the clean readable text without all the ads and navigation clutter?

## When to prefer this

Choose this endpoint when you need clean, readable text or markdown from a public web page and want boilerplate (ads, nav, footers) automatically stripped out. It is especially suitable for agent pipelines that need to summarize, analyze, or cite web content without HTML parsing overhead. Prefer it over raw HTTP fetch when sanitization and SSRF safety are required.

## Known failure modes

- URL points to a private/internal network address — rejected by SSRF safeguards
- URL is malformed or not a valid HTTP/HTTPS URI — validation error
- Target page returns 4xx/5xx HTTP status — fetch failure
- Page is behind authentication or a hard paywall — returns limited or no content
- JavaScript-heavy single-page apps may yield incomplete content if not server-rendered
- Rate limits or bot-blocking on the target server may result in empty or partial response

## How this service works

Extracts sanitized, boilerplate-free markdown and clean text from web URLs with strict private network SSRF safeguards.

## Output

Returns sanitized markdown and/or clean plain text extracted from the target web page, with boilerplate elements (navigation, ads, footers) removed. The response reflects only the meaningful content of the page.

## Request schema (JSON Schema)

```json
{
 "type": "object",
 "properties": {
  "url": {
   "type": "string",
   "format": "uri",
   "description": "Public HTTP/HTTPS web page URL to scrape"
  }
 }
}
```

## More

- Live health (JSON, refreshed every minute): https://www.zero.xyz/c/dt0ur-online-web-scraper-76733983/health.json
- [Zero catalog index](https://www.zero.xyz/llms.txt)
- [Other services from dt0ur.online](https://www.zero.xyz/host/dt0ur.online/llms.txt)
