# OpenAI gpt-5.4-long Chat Completions (via mm.family x402)

> OpenAI gpt-5.4-long Chat Completions (via mm.family x402) is a paid API for AI agents from openai.mm.family, paid per call via x402, $0.001091/call, status unknown (last checked 2026-10-02).

Runs GPT-5.4 chat completions for long-context inputs (>272K tokens), billed per call in USDC via x402 micropayment protocol

## Facts

- Endpoint: POST https://openai.mm.family/x402/v1/models/gpt-5.4-long/chat/completions?utm_source=zero.xyz
- Price: $0.001091/call
- Payment: x402
- Status: unknown
- Last checked: 2026-10-02
- Activations on Zero: 0
- Tags: x402
- Canonical page: https://www.zero.xyz/c/openai-gpt-5-4-long-chat-completions-via-mm-family-x402-a4da5ee3
- Structured record (JSON): https://api.zero.xyz/v1/capabilities/cap_ibSTEszZEbiKMnRJ8t4vL

Status and success rate cover calls made through Zero and Zero's own probes. Third-party monitors may report differently.

## How to call it through Zero

Zero handles the 402 payment challenge and records the run. With the Zero CLI installed (`npm i -g @zeroxyz/cli`):

```sh
zero fetch --capability openai-gpt-5-4-long-chat-completions-via-mm-family-x402-a4da5ee3 -d '<json body>'
```

Example prompt: Using GPT-5.4-long with standard service tier and a max of 2000 output tokens, send this entire codebase as context and ask it to find all security vulnerabilities and explain how to fix each one.

## When to prefer this

Choose this endpoint when you need to process inputs exceeding 272K tokens with GPT-5.4 quality, and you want to pay per-call in USDC without a traditional OpenAI subscription. Prefer the flex tier for cost-sensitive batch tasks and standard tier for latency-sensitive or priority workloads. If your context fits under 272K tokens, prefer the non-long gpt-5.4 variant to save on per-call costs.

## Known failure modes

- Insufficient USDC balance causes payment failure and request rejection
- Exceeding the model's absolute context window returns a context length error
- Invalid service_tier enum value returns a 400 validation error
- Empty or malformed messages array causes a 422 unprocessable entity error
- Requesting a model variant not in the enum (e.g. gpt-5.4-long:ultrahigh) returns a 400 error
- Network timeout on very large payloads if the request body is too large

## How this service works

gpt-5.4-long: OpenAI GPT 5.4 (>272K input tokens) chat completions, paid per call in USDC. Per 1M tokens in/out: Flex (default) 2.5/11.25, Standard 5.0/22.5; pick with service_tier. Plus $0.0005 (Base) or $0.0005 (Solana) per call; the quote charges input plus 10% of max_tokens. Levels: none, low, medium, high, xhigh (model gpt-5.4-long:<level>). Empty answers are free. Rates: https://openai.mm.family/x402/pricing

## Output

Returns a JSON chat completion object containing the assistant's reply message, finish reason (e.g. 'stop'), input/output token counts, the model variant used, and the service tier that handled the request (flex or standard).

## Request schema (JSON Schema)

```json
{
 "type": "object",
 "properties": {
  "model": {
   "enum": [
    "gpt-5.4-long",
    "gpt-5.4-long:none",
    "gpt-5.4-long:low",
    "gpt-5.4-long:medium",
    "gpt-5.4-long:high",
    "gpt-5.4-long:xhigh"
   ],
   "type": "string"
  },
  "tools": {
   "type": "array"
  },
  "stream": {
   "type": "boolean"
  },
  "messages": {
   "type": "array",
   "items": {
    "type": "object",
    "required": [
     "role",
     "content"
    ],
    "properties": {
     "role": {
      "type": "string"
     },
     "content": {
      "type": "string"
     }
    }
   }
  },
  "max_tokens": {
   "type": "integer"
  },
  "service_tier": {
   "enum": [
    "flex",
    "default",
    "standard",
    "auto",
    "fast",
    "priority"
   ],
   "type": "string"
  },
  "response_format": {
   "type": "object"
  },
  "reasoning_effort": {
   "type": "string"
  }
 }
}
```

## Response schema (JSON Schema)

```json
{
 "type": "json",
 "example": {
  "id": "chatcmpl-x",
  "model": "gpt-5.4-long",
  "usage": {
   "total_tokens": 9,
   "prompt_tokens": 8,
   "completion_tokens": 1
  },
  "object": "chat.completion",
  "choices": [
   {
    "index": 0,
    "message": {
     "role": "assistant",
     "content": "Hello",
     "refusal": null
    },
    "finish_reason": "stop"
   }
  ],
  "service_tier": "flex"
 }
}
```

## More

- Live health (JSON, refreshed every minute): https://www.zero.xyz/c/openai-gpt-5-4-long-chat-completions-via-mm-family-x402-a4da5ee3/health.json
- [Zero catalog index](https://www.zero.xyz/llms.txt)
- [Other services from openai.mm.family](https://www.zero.xyz/host/openai.mm.family/llms.txt)
