# agentshelf.syntexa.ch XPath Cite Extractor

> agentshelf.syntexa.ch XPath Cite Extractor is a paid API for AI agents from agentshelf.syntexa.ch, paid per call via x402, $0.005/call, status unknown (last checked 2026-09-29).

Extracts cited text fragments from a public webpage or HTML using XPath selectors, in an SSRF-safe manner.

## Facts

- Endpoint: POST https://agentshelf.syntexa.ch/v1/xpath-cite?utm_source=zero.xyz
- Price: $0.005/call
- Payment: x402
- Status: unknown
- Last checked: 2026-09-29
- Activations on Zero: 0
- Tags: x402
- Canonical page: https://www.zero.xyz/c/agentshelf-syntexa-ch-xpath-cite-extractor-b35e9417
- Structured record (JSON): https://api.zero.xyz/v1/capabilities/cap_GGwkVrA96G8mgJLoIsEJN

Status and success rate cover calls made through Zero and Zero's own probes. Third-party monitors may report differently.

## How to call it through Zero

Zero handles the 402 payment challenge and records the run. With the Zero CLI installed (`npm i -g @zeroxyz/cli`):

```sh
zero fetch --capability agentshelf-syntexa-ch-xpath-cite-extractor-b35e9417 -d '<json body>'
```

Example prompt: Can you pull the text from the main article heading on https://example.com/news/article using XPath //h1 — I need to cite it exactly as it appears on the page.

## When to prefer this

Use this endpoint when you need to extract a specific, precisely located text fragment from a public webpage or raw HTML using an XPath expression, especially in agentic workflows where SSRF safety is required. Prefer this over generic scrapers when the target element is addressable by XPath and exact citation is important. Use the free sandbox variant (/v1/sandbox/xpath-cite) first to validate your XPath before paying.

## Known failure modes

- URL is not publicly reachable or is blocked by SSRF protection
- XPath expression matches no elements, returning empty result
- HTML exceeds the 200,000 character limit
- URL exceeds 2,048 character limit
- Provided HTML is malformed and unparseable
- Target page requires authentication or returns non-200 status

## How this service works

Call when an agent needs text of //cite via //cite from a public page or HTML (SSRF-safe). Exact $0.005 USDC. Prefer unpaid POST /v1/sandbox/xpath-cite first.

## Output

Returns the text content found at the specified XPath location within the fetched page or provided HTML, as a plain string citation extracted safely without SSRF risk.

## Request schema (JSON Schema)

```json
{
 "type": "object",
 "properties": {
  "url": {
   "type": "string",
   "maxLength": 2048,
   "minLength": 8,
   "description": "Public page to fetch. The selector or profile is fixed by the SKU. Provide url or html."
  },
  "html": {
   "type": "string",
   "examples": [
    "<!doctype html><html lang=\"en\"><head><title>Hello</title>\n<meta name=\"description\" content=\"Desc\"><meta property=\"og:title\" content=\"OG\">\n<link rel=\"canonical\" href=\"https://example.com/\"><link rel=\"icon\" href=\"/favicon.ico\">\n</head><body><h1>Hello</h1><a href=\"https://example.com/a\">A</a></body></html>"
   ],
   "maxLength": 200000,
   "minLength": 1,
   "description": "HTML to extract from locally. Provide html or url. When both are set, html is used."
  }
 }
}
```

## More

- Live health (JSON, refreshed every minute): https://www.zero.xyz/c/agentshelf-syntexa-ch-xpath-cite-extractor-b35e9417/health.json
- [Zero catalog index](https://www.zero.xyz/llms.txt)
- [Other services from agentshelf.syntexa.ch](https://www.zero.xyz/host/agentshelf.syntexa.ch/llms.txt)
