# OpenAI GPT-6 Luna Long (>272K Context) Chat Completions

> OpenAI GPT-6 Luna Long (>272K Context) Chat Completions is a paid API for AI agents from openai.mm.family, paid per call via x402, $0.001/call, status unknown (last checked 2026-10-02).

Provides GPT-6 Luna chat completions with over 272K input token context window, billed per call in USDC via x402 protocol with tiered service levels.

## Facts

- Endpoint: POST https://openai.mm.family/x402/v1/models/gpt-6-luna-long/chat/completions?utm_source=zero.xyz
- Price: $0.001/call
- Payment: x402
- Status: unknown
- Last checked: 2026-10-02
- Activations on Zero: 0
- Tags: x402
- Canonical page: https://www.zero.xyz/c/openai-gpt-6-luna-long-272k-context-chat-completions-0be4906d
- Structured record (JSON): https://api.zero.xyz/v1/capabilities/cap_phQcZukPFojSSZ9x-iqqE

Status and success rate cover calls made through Zero and Zero's own probes. Third-party monitors may report differently.

## How to call it through Zero

Zero handles the 402 payment challenge and records the run. With the Zero CLI installed (`npm i -g @zeroxyz/cli`):

```sh
zero fetch --capability openai-gpt-6-luna-long-272k-context-chat-completions-0be4906d -d '<json body>'
```

Example prompt: Use GPT-6 Luna Long with flex tier and up to 4096 output tokens to analyze this 200,000-word legal document I'm pasting and summarize the key obligations in bullet points.

## When to prefer this

Choose this endpoint when you need more than 272K tokens of input context with GPT-6 Luna specifically, or when cost optimization matters (flex tier is cheapest). Prefer over gpt-6-astra when your prompt exceeds the standard context limit. Use when you need USDC micropayment billing via x402 rather than traditional API key auth. Choose higher service tiers (standard, fast) when latency is critical.

## Known failure modes

- Insufficient USDC balance — payment fails before completion is returned
- max_tokens quota exceeded causing truncated output
- Model overloaded at selected service_tier — try flex or auto tier
- Input exceeds 272K token context window — request rejected
- Invalid model variant string returns 400 error
- Streaming errors if network drops mid-response

## How this service works

gpt-6-luna-long: OpenAI GPT 6 Luna (>272K input tokens) chat completions, paid per call in USDC. Per 1M tokens in/out: Flex (default) 0.1/0.375, Standard 0.2/0.75, Fast 0.4/1.5; pick with service_tier. Plus $0.0005 (Base) or $0.0005 (Solana) per call; the quote charges input plus 10% of max_tokens. Levels: none, low, medium, high, xhigh (model gpt-6-luna-long:<level>). Empty answers are free. Rates: https://openai.mm.family/x402/pricing

## Output

Returns a standard OpenAI-compatible chat completion JSON object including the assistant's reply, token usage breakdown (prompt, completion, total), finish reason, service tier used, and a unique completion ID. Empty responses incur no charge.

## Request schema (JSON Schema)

```json
{
 "type": "object",
 "properties": {
  "model": {
   "enum": [
    "gpt-6-luna-long",
    "gpt-6-luna-long:none",
    "gpt-6-luna-long:low",
    "gpt-6-luna-long:medium",
    "gpt-6-luna-long:high",
    "gpt-6-luna-long:xhigh"
   ],
   "type": "string"
  },
  "tools": {
   "type": "array"
  },
  "stream": {
   "type": "boolean"
  },
  "messages": {
   "type": "array",
   "items": {
    "type": "object",
    "required": [
     "role",
     "content"
    ],
    "properties": {
     "role": {
      "type": "string"
     },
     "content": {
      "type": "string"
     }
    }
   }
  },
  "max_tokens": {
   "type": "integer"
  },
  "service_tier": {
   "enum": [
    "flex",
    "default",
    "standard",
    "auto",
    "fast",
    "priority"
   ],
   "type": "string"
  },
  "response_format": {
   "type": "object"
  },
  "reasoning_effort": {
   "type": "string"
  }
 }
}
```

## Response schema (JSON Schema)

```json
{
 "type": "json",
 "example": {
  "id": "chatcmpl-x",
  "model": "gpt-6-luna-long",
  "usage": {
   "total_tokens": 9,
   "prompt_tokens": 8,
   "completion_tokens": 1
  },
  "object": "chat.completion",
  "choices": [
   {
    "index": 0,
    "message": {
     "role": "assistant",
     "content": "Hello",
     "refusal": null
    },
    "finish_reason": "stop"
   }
  ],
  "service_tier": "flex"
 }
}
```

## More

- Live health (JSON, refreshed every minute): https://www.zero.xyz/c/openai-gpt-6-luna-long-272k-context-chat-completions-0be4906d/health.json
- [Zero catalog index](https://www.zero.xyz/llms.txt)
- [Other services from openai.mm.family](https://www.zero.xyz/host/openai.mm.family/llms.txt)
