# ZeroReader Llama 3.2 11B Vision

> ZeroReader Llama 3.2 11B Vision is a paid API for AI agents from api.zeroreader.com, paid per call via x402, $0.005/call, status unknown (last checked 2026-10-02).

Runs multimodal chat completions using Meta's Llama 3.2 11B Vision model, capable of understanding both images and text

## Facts

- Endpoint: POST https://api.zeroreader.com/v1/ai/llama-vision?utm_source=zero.xyz
- Price: $0.005/call
- Payment: x402
- Status: unknown
- Last checked: 2026-10-02
- Activations on Zero: 0
- Tags: x402
- Canonical page: https://www.zero.xyz/c/zeroreader-llama-3-2-11b-vision-028971c4
- Structured record (JSON): https://api.zero.xyz/v1/capabilities/cap_GcgZfH-Df_GepKKLphbug

Status and success rate cover calls made through Zero and Zero's own probes. Third-party monitors may report differently.

## How to call it through Zero

Zero handles the 402 payment challenge and records the run. With the Zero CLI installed (`npm i -g @zeroxyz/cli`):

```sh
zero fetch --capability zeroreader-llama-3-2-11b-vision-028971c4 -d '<json body>'
```

Example prompt: Look at this image of a product label and tell me all the ingredients listed — use the vision model with temperature 0.3 and up to 512 tokens in your response.

## When to prefer this

Choose this endpoint when you need a vision-capable model that can process both images and text in a single request. Prefer this over text-only Llama variants (3B, etc.) when the input contains visual content like photos, screenshots, charts, or diagrams. Prefer over larger models when cost and speed matter and the task doesn't require deep reasoning or 100B+ parameter capacity.

## Known failure modes

- Invalid or missing messages array returns a 400 validation error
- Temperature outside 0–2 range causes a 400 error
- max_tokens exceeding 4096 returns a 400 error
- Image content that cannot be parsed or is too large may result in a 422 or model error
- Payment failure (x402) returns a 402 Payment Required before processing
- Network timeout if the model takes too long to generate a long response

## How this service works

Llama 3.2 11B Vision — Vision + text model. Can understand images.

## Output

Returns a chat completion object in OpenAI-compatible format, including the assistant's message content, finish reason (e.g. 'stop'), and a completion ID. The assistant's response will reflect analysis of any image or text provided in the messages array.

## Request schema (JSON Schema)

```json
{
 "type": "object",
 "properties": {
  "stream": {
   "type": "boolean",
   "default": false
  },
  "messages": {
   "type": "array",
   "items": {
    "type": "object",
    "required": [
     "role",
     "content"
    ],
    "properties": {
     "role": {
      "enum": [
       "system",
       "user",
       "assistant"
      ],
      "type": "string"
     },
     "content": {
      "type": "string"
     }
    }
   }
  },
  "max_tokens": {
   "type": "integer",
   "default": 1024,
   "maximum": 4096
  },
  "temperature": {
   "type": "number",
   "default": 0.7,
   "maximum": 2,
   "minimum": 0
  }
 }
}
```

## Response schema (JSON Schema)

```json
{
 "id": "chatcmpl-example",
 "object": "chat.completion",
 "choices": [
  {
   "index": 0,
   "message": {
    "role": "assistant",
    "content": "I'm doing well!"
   },
   "finish_reason": "stop"
  }
 ]
}
```

## More

- Live health (JSON, refreshed every minute): https://www.zero.xyz/c/zeroreader-llama-3-2-11b-vision-028971c4/health.json
- [Zero catalog index](https://www.zero.xyz/llms.txt)
- [Other services from api.zeroreader.com](https://www.zero.xyz/host/api.zeroreader.com/llms.txt)
