# Pyfile Web Reader — URL to Markdown

> Pyfile Web Reader — URL to Markdown is a paid API for AI agents from pyfile-agent.taile3ff35.ts.net, paid per call via x402, $0.003/call, status unknown (last checked 2026-10-02).

Fetches any public web page and returns clean, readable Markdown with navigation and ads stripped, plus the page title, for use in RAG and LLM pipelines.

## Facts

- Endpoint: GET https://pyfile-agent.taile3ff35.ts.net/web/read?utm_source=zero.xyz
- Price: $0.003/call
- Payment: x402
- Status: unknown
- Last checked: 2026-10-02
- Activations on Zero: 0
- Tags: x402
- Canonical page: https://www.zero.xyz/c/pyfile-web-reader-url-to-markdown-4f4dae42
- Structured record (JSON): https://api.zero.xyz/v1/capabilities/cap_uo1TQEmdHAlAXEokNnLUM

Status and success rate cover calls made through Zero and Zero's own probes. Third-party monitors may report differently.

## How to call it through Zero

Zero handles the 402 payment challenge and records the run. With the Zero CLI installed (`npm i -g @zeroxyz/cli`):

```sh
zero fetch --capability pyfile-web-reader-url-to-markdown-4f4dae42
```

Example prompt: Fetch the article at https://techcrunch.com/2024/05/01/ai-funding/ and give me the clean readable text — strip all the navigation and ads, and cap it at 10000 characters.

## When to prefer this

Use this endpoint when you need clean, LLM-ready Markdown from any public URL without writing your own scraper — especially for RAG pipelines, research agents, or any workflow where you need the main article body free of HTML noise, nav bars, and ads. Prefer it over raw HTML fetchers when downstream consumers are language models or vector databases.

## Known failure modes

- URL is behind a login wall or paywall — returns empty or partial markdown
- Page blocks scrapers (Cloudflare, bot detection) — fetch may fail or return error content
- Very large pages truncated if max_chars limit exceeded — truncated flag set to true
- Invalid or malformed URL — likely returns an error response
- Dynamic JavaScript-rendered pages may not return full content if not rendered server-side

## How this service works

Read any public web page as clean readable Markdown for RAG and LLM pipelines — main article text with nav/ads stripped, plus title. ?url=https://example.com&max_chars=20000

## Output

A JSON object containing: the original URL, the extracted Markdown body (main article text with nav/ads stripped), the page title, byte count, whether the text was truncated, the fetch timestamp, and the source identifier (r.jina.ai).

## Request schema (JSON Schema)

```json
{
 "type": "object",
 "properties": {
  "properties": {
   "type": "string"
  }
 }
}
```

## Response schema (JSON Schema)

```json
{
 "type": "json",
 "example": {
  "url": "https://example.com/article",
  "bytes": 4210,
  "title": "Example Article",
  "source": "r.jina.ai",
  "markdown": "# Example Article\n\nMain readable text with nav and ads stripped...",
  "truncated": false,
  "fetched_at": "2026-09-22T15:00:00.000Z"
 }
}
```

## More

- Live health (JSON, refreshed every minute): https://www.zero.xyz/c/pyfile-web-reader-url-to-markdown-4f4dae42/health.json
- [Zero catalog index](https://www.zero.xyz/llms.txt)
- [Other services from pyfile-agent.taile3ff35.ts.net](https://www.zero.xyz/host/pyfile-agent.taile3ff35.ts.net/llms.txt)
