# Omnia Odds OpenAI-Compatible Chat Completions

> Omnia Odds OpenAI-Compatible Chat Completions is a paid API for AI agents from odds.rjhsignaltech.workers.dev, paid per call via x402, $0.002/call, status unknown (last checked 2026-10-02).

Provides OpenAI-compatible LLM chat completions using Llama 3.1 8B (fast) or Llama 3.3 70B (smart), payable per request via x402 with no API key required.

## Facts

- Endpoint: POST https://odds.rjhsignaltech.workers.dev/v1/chat/completions?utm_source=zero.xyz
- Price: $0.002/call
- Payment: x402
- Status: unknown
- Last checked: 2026-10-02
- Activations on Zero: 0
- Tags: x402
- Canonical page: https://www.zero.xyz/c/omnia-odds-openai-compatible-chat-completions-34258174
- Structured record (JSON): https://api.zero.xyz/v1/capabilities/cap_InIbKI3dkk0urARsppQUX

Status and success rate cover calls made through Zero and Zero's own probes. Third-party monitors may report differently.

## How to call it through Zero

Zero handles the 402 payment challenge and records the run. With the Zero CLI installed (`npm i -g @zeroxyz/cli`):

```sh
zero fetch --capability omnia-odds-openai-compatible-chat-completions-34258174 -d '<json body>'
```

Example prompt: Using the smart Llama 3.3 70B model, draft a professional apology email to a client explaining a project delay — keep it under 200 words and use a warm but formal tone.

## When to prefer this

Choose this endpoint when you need OpenAI-compatible LLM inference with no API key or signup, paying only per request in USDC via x402. It's ideal for autonomous agents that need serverless, frictionless LLM access without managing credentials, and for pipelines already using OpenAI SDK that want a drop-in alternative. Prefer 'fast' (Llama 3.1 8B) for speed-sensitive or high-volume tasks, and 'smart' (Llama 3.3 70B) for quality-critical reasoning or generation tasks.

## Known failure modes

- Payment not received or x402 handshake fails — request rejected before inference
- max_tokens exceeds 1024 — validation error returned
- Invalid model name supplied (not 'fast' or 'smart') — error response
- Temperature out of range (0–2) — validation failure
- Empty or malformed messages array — bad request error
- Cloudflare Worker timeout if model inference takes too long
- Upstream Llama inference service unavailable — 5xx error

## How this service works

OpenAI-compatible chat completions: POST messages[] with model fast (Llama 3.1 8B) or smart (Llama 3.3 70B), max_tokens, temperature; standard chat.completion response. Point any OpenAI SDK at this base URL. LLM inference for agents, no API key, no signup, pay per request

## Output

Returns a standard OpenAI-format chat.completion JSON object containing the assistant's reply message, the model used, token usage counts (prompt, completion, total), and finish reason. The response is drop-in compatible with any OpenAI SDK or client expecting a chat/completions response.

## Request schema (JSON Schema)

```json
{
 "type": "object",
 "properties": {
  "model": {
   "type": "string",
   "description": "\"fast\" (Llama 3.1 8B, default) or \"smart\" (Llama 3.3 70B)"
  },
  "messages": {
   "type": "array",
   "items": {
    "type": "object",
    "required": [
     "role",
     "content"
    ],
    "properties": {
     "role": {
      "enum": [
       "system",
       "user",
       "assistant"
      ],
      "type": "string"
     },
     "content": {
      "type": "string"
     }
    }
   }
  },
  "max_tokens": {
   "type": "integer",
   "maximum": 1024,
   "minimum": 1
  },
  "temperature": {
   "type": "number",
   "maximum": 2,
   "minimum": 0
  }
 }
}
```

## More

- Live health (JSON, refreshed every minute): https://www.zero.xyz/c/omnia-odds-openai-compatible-chat-completions-34258174/health.json
- [Zero catalog index](https://www.zero.xyz/llms.txt)
- [Other services from odds.rjhsignaltech.workers.dev](https://www.zero.xyz/host/odds.rjhsignaltech.workers.dev/llms.txt)
