# Page Metadata Extractor (Open Graph, Twitter Card, Canonical, Feeds, JSON-LD, Language)

> Page Metadata Extractor (Open Graph, Twitter Card, Canonical, Feeds, JSON-LD, Language) is a paid API for AI agents from twin.unykorn.org, paid per call via x402, $0.002/call, status unknown (last checked 2026-10-01).

Extracts structured metadata from a public webpage including Open Graph tags, Twitter card data, canonical URL, RSS/Atom feeds, JSON-LD types, and language information.

## Facts

- Endpoint: POST https://twin.unykorn.org/web/page-meta?utm_source=zero.xyz
- Price: $0.002/call
- Payment: x402
- Status: unknown
- Last checked: 2026-10-01
- Activations on Zero: 0
- Tags: x402
- Canonical page: https://www.zero.xyz/c/page-metadata-extractor-open-graph-twitter-card-canonical-feeds-json-9698c393
- Structured record (JSON): https://api.zero.xyz/v1/capabilities/cap_orN_sFLcYl6uBggWQ-vSD

Status and success rate cover calls made through Zero and Zero's own probes. Third-party monitors may report differently.

## How to call it through Zero

Zero handles the 402 payment challenge and records the run. With the Zero CLI installed (`npm i -g @zeroxyz/cli`):

```sh
zero fetch --capability page-metadata-extractor-open-graph-twitter-card-canonical-feeds-json-9698c393 -d '<json body>'
```

Example prompt: Can you pull all the page metadata from https://techcrunch.com — I need the Open Graph tags, Twitter card info, canonical URL, any RSS feeds, JSON-LD types, and the page language?

## When to prefer this

Choose this endpoint when you need to quickly extract standardized web metadata (Open Graph, Twitter card, canonical, feeds, JSON-LD, language) from a single public URL without building your own HTML parser. It is well-suited for link-preview generation, SEO audits, feed discovery, and structured data inspection. Prefer it over general-purpose scraping endpoints when the specific goal is head-level metadata extraction rather than full page content.

## Known failure modes

- URL is not publicly accessible (private, paywalled, or requires authentication) — returns error or empty metadata
- URL returns non-HTML content (PDF, image) — metadata fields may be absent
- Page uses JavaScript-rendered metadata that cannot be extracted from raw HTML — fields may be missing
- Invalid or malformed URL input — returns validation error
- Page blocks crawlers via robots.txt or rate limiting — may return empty or partial results

## How this service works

Page metadata: Open Graph, Twitter card, canonical, feeds, JSON-LD types, language — Genesis402 / UnyKorn Operator Network

## Output

A structured object containing extracted page metadata: Open Graph properties (og:title, og:description, og:image, etc.), Twitter card tags, canonical URL, discovered RSS/Atom feed URLs, JSON-LD schema types present on the page, and the declared or detected language of the page.

## Request schema (JSON Schema)

```json
{
 "type": "object",
 "properties": {
  "params": {
   "type": "object",
   "properties": {
    "url": {
     "type": "string",
     "description": "required public URL"
    }
   }
  }
 }
}
```

## Response schema (JSON Schema)

```json
{
 "type": "json",
 "example": {
  "ok": true,
  "type": "page-meta",
  "receipt": {
   "tx_hash": "0x<64hex>",
   "amount_usd": 0.002,
   "receipt_id": "g402-<16hex>"
  },
  "sources": [
   {
    "ok": true,
    "name": "<source>"
   }
  ],
  "limitations": "<text>",
  "generated_at": "<iso time>",
  "evidence_hash": "sha256:<64hex>"
 }
}
```

## More

- Live health (JSON, refreshed every minute): https://www.zero.xyz/c/page-metadata-extractor-open-graph-twitter-card-canonical-feeds-json-9698c393/health.json
- [Zero catalog index](https://www.zero.xyz/llms.txt)
- [Other services from twin.unykorn.org](https://www.zero.xyz/host/twin.unykorn.org/llms.txt)
