# vaaya.ai CRW Scrape — Single URL Web Scraper

> vaaya.ai CRW Scrape — Single URL Web Scraper is a paid API for AI agents from vaaya.ai, paid per call via x402, $0.01/call, status unknown (last checked 2026-10-01, last successful call 2026-08-13).

Scrapes a single URL and returns clean markdown, HTML, JSON, plain text, links, or screenshot content synchronously

## Facts

- Endpoint: POST https://vaaya.ai/api/run/crw/scrape?utm_source=zero.xyz
- Price: $0.01/call
- Payment: x402
- Status: unknown
- Last checked: 2026-10-01
- Last successful call: 2026-08-13
- Success rate: 100% of calls made through Zero
- Rating: 2.7 / 5 from 1 review
- Activations on Zero: 1
- Tags: x402
- Canonical page: https://www.zero.xyz/c/vaaya-ai-crw-scrape-single-url-web-scraper-d117dc65
- Structured record (JSON): https://api.zero.xyz/v1/capabilities/cap_0-dURQPVXxEcZTZgiprLu

Status and success rate cover calls made through Zero and Zero's own probes. Third-party monitors may report differently.

## How to call it through Zero

Zero handles the 402 payment challenge and records the run. With the Zero CLI installed (`npm i -g @zeroxyz/cli`):

```sh
zero fetch --capability vaaya-ai-crw-scrape-single-url-web-scraper-d117dc65 -d '<json body>'
```

Example prompt: Scrape https://techcrunch.com/2024/01/15/openai-funding and give me the main article content as clean markdown, ignoring nav and footer boilerplate.

## When to prefer this

Choose this endpoint when you need to scrape a single URL synchronously and want the result immediately without managing a crawl job. It is the cheapest full-featured scraping option on vaaya.ai and supports Firecrawl-compatible options including stealth mode, JS rendering, structured JSON extraction, country-based proxying, and multiple output formats. Prefer it over the async crawl endpoint for single-page extraction tasks where you don't need to follow links.

## Known failure modes

- URL unreachable or returns non-200 status — returns error with HTTP status code
- Anti-bot / CAPTCHA blocking without stealth mode — returns empty or blocked content
- JavaScript-heavy page not fully rendered — use renderJs=true or increase waitFor ms
- Invalid or unsupported URL format — returns validation error
- Requested JSON schema does not match page structure — returns empty or partial JSON
- Country code proxy unavailable — returns error or falls back
- Timeout on slow pages — increase waitFor parameter

## How this service works

CRW — Scrape a single URL to clean markdown/HTML/JSON (Firecrawl-compatible). Pass `url`; optional `formats` (markdown|html|rawHtml|plainText|links|json|summary|screenshot, default markdown), `onlyMainContent` (default true), `renderJs` (null = auto-detect), `waitFor` (ms), `stealth` (anti-bot browser), `country` (2-letter proxy egress), `jsonSchema` (structured output). Sync — returns the scraped content directly. Cheapest full-featured scrape rung.

## Output

Returns the scraped content of the requested URL in the specified format(s): clean markdown (default), raw or processed HTML, plain text, a list of links, structured JSON matching a provided schema, a page summary, or a screenshot URL. The response is synchronous and delivered directly.

## Request schema (JSON Schema)

```json
{
 "type": "object",
 "$schema": "https://json-schema.org/draft/2020-12/schema",
 "required": [
  "input"
 ],
 "properties": {
  "input": {
   "type": "object",
   "required": [
    "type",
    "method",
    "bodyType",
    "body"
   ],
   "properties": {
    "body": {
     "type": "object",
     "$schema": "http://json-schema.org/draft-07/schema#",
     "required": [
      "url"
     ],
     "properties": {
      "url": {
       "type": "string",
       "format": "uri"
      },
      "country": {
       "type": "string",
       "maxLength": 2,
       "minLength": 2
      },
      "formats": {
       "type": "array",
       "items": {
        "type": "string"
       }
      },
      "stealth": {
       "type": "boolean"
      },
      "waitFor": {
       "type": "integer",
       "minimum": 0
      },
      "renderJs": {
       "type": [
        "boolean",
        "null"
       ]
      },
      "jsonSchema": {
       "type": "object",
       "additionalProperties": {}
      },
      "onlyMainContent": {
       "type": "boolean"
      }
     },
     "additionalProperties": true
    },
    "type": {
     "type": "string",
     "const": "http"
    },
    "method": {
     "enum": [
      "POST"
     ],
     "type": "string"
    },
    "bodyType": {
     "enum": [
      "json",
      "form-data",
      "text"
     ],
     "type": "string"
    },
    "pathParams": {
     "type": "object"
    }
   },
   "additionalProperties": false
  },
  "output": {
   "type": "object",
   "required": [
    "type"
   ],
   "properties": {
    "type": {
     "type": "string"
    },
    "example": {
     "type": "object"
    }
   }
  }
 }
}
```

## More

- Live health (JSON, refreshed every minute): https://www.zero.xyz/c/vaaya-ai-crw-scrape-single-url-web-scraper-d117dc65/health.json
- [Zero catalog index](https://www.zero.xyz/llms.txt)
- [Other services from vaaya.ai](https://www.zero.xyz/host/vaaya.ai/llms.txt)
