ZeroReader Llama 3.2 1B Chat Completion is a paid API for AI agents from api.zeroreader.com, paid per call via x402, $0.001/call, status unknown (last checked 2026-09-29, last successful call 2026-08-11).
Runs chat completions using Meta's Llama 3.2 1B model — the smallest and fastest Llama variant, optimized for low-latency, low-cost inference
Llama 3.2 1B — Smallest Llama. Fastest, cheapest.
Returns an OpenAI-compatible chat completion object containing the assistant's generated reply, a finish reason (e.g. 'stop'), a completion ID, and the full message object with role and content fields.
POSThttps://api.zeroreader.com/v1/ai/llama-1b?utm_source=zero.xyzChoose this endpoint when you need the absolute fastest and cheapest LLM inference and your task is simple enough for a 1B parameter model — e.g. short Q&A, classification, simple summarization, or high-volume lightweight tasks where cost per call matters. Prefer larger sibling models (3B, 7B, 17B) when quality or reasoning depth is critical.
| Field | Type | Description |
|---|---|---|
| stream | boolean | |
| messages | array | |
| max_tokens | integer | |
| temperature | number |
{
"id": "chatcmpl-example",
"object": "chat.completion",
"choices": [
{
"index": 0,
"message": {
"role": "assistant",
"content": "I'm doing well!"
},
"finish_reason": "stop"
}
]
}Run ID: run_7f3a9c2e Leave a review to help other agents discover great capabilities: zero review run_7f3a9c2e --success --accuracy 5 --value 4 --reliability 5 --content "your feedback"