# Vextorium PDF Text Extractor

> Vextorium PDF Text Extractor is a paid API for AI agents from api.vextorium.com, paid per call via x402, $0.012/call, status unknown (last checked 2026-10-01).

Extracts the full text content from a PDF document given its direct URL

## Facts

- Endpoint: POST https://api.vextorium.com/extraer-pdf-pago?utm_source=zero.xyz
- Price: $0.012/call
- Payment: x402
- Status: unknown
- Last checked: 2026-10-01
- Activations on Zero: 0
- Tags: x402
- Canonical page: https://www.zero.xyz/c/vextorium-pdf-text-extractor-642703e7
- Structured record (JSON): https://api.zero.xyz/v1/capabilities/cap_6goXhzmDpjBPpk8YMgho8

Status and success rate cover calls made through Zero and Zero's own probes. Third-party monitors may report differently.

## How to call it through Zero

Zero handles the 402 payment challenge and records the run. With the Zero CLI installed (`npm i -g @zeroxyz/cli`):

```sh
zero fetch --capability vextorium-pdf-text-extractor-642703e7 -d '<json body>'
```

Example prompt: Can you grab the text out of this PDF for me? It's at https://example.com/report.pdf — I need the actual content I can read and search through.

## When to prefer this

Choose this endpoint when you have a direct URL to a PDF file and need its text content for downstream processing, summarization, search, or analysis. It is well-suited for text-layer PDFs (not scanned images). Prefer this over general-purpose web scrapers when the target is specifically a PDF. It accepts payment via x402/USDC micropayment at $0.012 per call, making it suitable for per-use agent workflows.

## Known failure modes

- PDF URL is not directly accessible or returns a non-200 status
- PDF contains only scanned images (no embedded text layer), resulting in empty or null text
- Text is truncated if the document is very long (texto_truncado: true)
- Invalid or malformed URL input causes an error
- Password-protected PDFs cannot be read
- Network timeout if the PDF host is slow or unreachable

## How this service works

Extract the text of a PDF from its URL.

## Output

Returns the full extracted text of the PDF, the number of pages, a boolean indicating whether the text was truncated (e.g. due to length limits), and an optional notice or warning string. All fields may be null if extraction fails or the PDF is image-only.

## Request schema (JSON Schema)

```json
{
 "type": "object",
 "$schema": "https://json-schema.org/draft/2020-12/schema",
 "required": [
  "input"
 ],
 "properties": {
  "input": {
   "type": "object",
   "required": [
    "type",
    "method",
    "bodyType",
    "body"
   ],
   "properties": {
    "body": {
     "required": [
      "url"
     ],
     "properties": {
      "url": {
       "type": "string",
       "description": "URL directa a un PDF con texto real"
      }
     }
    },
    "type": {
     "type": "string",
     "const": "http"
    },
    "method": {
     "enum": [
      "POST"
     ],
     "type": "string"
    },
    "bodyType": {
     "enum": [
      "json",
      "form-data",
      "text"
     ],
     "type": "string"
    }
   },
   "additionalProperties": false
  },
  "output": {
   "type": "object",
   "required": [
    "type"
   ],
   "properties": {
    "type": {
     "type": "string"
    },
    "example": {
     "type": "object",
     "properties": {
      "url": {
       "type": [
        "string",
        "null"
       ]
      },
      "aviso": {
       "type": [
        "string",
        "null"
       ]
      },
      "texto": {
       "type": [
        "string",
        "null"
       ]
      },
      "num_paginas": {
       "type": [
        "number",
        "null"
       ]
      },
      "texto_truncado": {
       "type": [
        "boolean",
        "null"
       ]
      }
     }
    }
   }
  }
 }
}
```

## Response schema (JSON Schema)

```json
{
 "type": "json",
 "example": {
  "url": "https://ejemplo.com/documento.pdf",
  "texto": "Contenido extraido...",
  "num_paginas": 3,
  "texto_truncado": false
 }
}
```

## More

- Live health (JSON, refreshed every minute): https://www.zero.xyz/c/vextorium-pdf-text-extractor-642703e7/health.json
- [Zero catalog index](https://www.zero.xyz/llms.txt)
- [Other services from api.vextorium.com](https://www.zero.xyz/host/api.vextorium.com/llms.txt)
