# OpenAI GPT-6 Astra Chat Completions (via mm.family x402)

> OpenAI GPT-6 Astra Chat Completions (via mm.family x402) is a paid API for AI agents from openai.mm.family, paid per call via x402, $0.00181/call, status unknown (last checked 2026-10-02).

Provides GPT-6 Astra chat completions with up to 272K input tokens, billed per call in USDC via x402 payment protocol with selectable service tiers

## Facts

- Endpoint: POST https://openai.mm.family/x402/v1/models/gpt-6-astra/chat/completions?utm_source=zero.xyz
- Price: $0.00181/call
- Payment: x402
- Status: unknown
- Last checked: 2026-10-02
- Activations on Zero: 0
- Tags: x402
- Canonical page: https://www.zero.xyz/c/openai-gpt-6-astra-chat-completions-via-mm-family-x402-b53058d4
- Structured record (JSON): https://api.zero.xyz/v1/capabilities/cap_RcTG0H_gMBMmVddHtHUR2

Status and success rate cover calls made through Zero and Zero's own probes. Third-party monitors may report differently.

## How to call it through Zero

Zero handles the 402 payment challenge and records the run. With the Zero CLI installed (`npm i -g @zeroxyz/cli`):

```sh
zero fetch --capability openai-gpt-6-astra-chat-completions-via-mm-family-x402-b53058d4 -d '<json body>'
```

Example prompt: Using GPT-6 Astra on the mm.family x402 gateway, send this conversation to the model with flex pricing and up to 1000 output tokens: system 'You are a helpful assistant', user 'Explain quantum entanglement in simple terms'.

## When to prefer this

Choose this endpoint when you need access to GPT-6 Astra's frontier-level reasoning with up to 272K input context, want per-call USDC micropayment billing via x402 (no subscription), and want to tune cost/speed tradeoffs using flex, standard, or fast service tiers. Prefer it over cheaper sibling models (gpt-5-nano, gpt-5.4-mini) when task complexity demands the most capable model. Use the gpt-6.1-sol-long sibling if your input exceeds 272K tokens.

## Known failure modes

- Insufficient USDC balance — payment rejected before inference runs
- Exceeded 272K token context window — use gpt-6.1-sol-long sibling instead
- Invalid model variant string — must be one of the gpt-6-astra enum values
- max_tokens too large for selected service tier capacity
- Stream mode requested but client not handling chunked responses
- Empty response returned free of charge when model produces no output

## How this service works

gpt-6-astra: OpenAI GPT 6 Astra (<=272K input tokens) chat completions, paid per call in USDC. Per 1M tokens in/out: Flex (default) 5.0/25.0, Standard 10.0/50.0, Fast 20.0/100.0; pick with service_tier. Plus $0.0005 (Base) or $0.0005 (Solana) per call; the quote charges input plus 10% of max_tokens. Levels: low, medium, high, xhigh (model gpt-6-astra:<level>). Empty answers are free. Rates: https://openai.mm.family/x402/pricing

## Output

Returns a JSON chat completion object containing the assistant's reply message, finish reason (e.g. 'stop'), token usage breakdown (prompt, completion, total tokens), model identifier, and service tier used. The response mirrors the OpenAI chat completions API format.

## Request schema (JSON Schema)

```json
{
 "type": "object",
 "properties": {
  "model": {
   "enum": [
    "gpt-6-astra",
    "gpt-6-astra:low",
    "gpt-6-astra:medium",
    "gpt-6-astra:high",
    "gpt-6-astra:xhigh"
   ],
   "type": "string"
  },
  "tools": {
   "type": "array"
  },
  "stream": {
   "type": "boolean"
  },
  "messages": {
   "type": "array",
   "items": {
    "type": "object",
    "required": [
     "role",
     "content"
    ],
    "properties": {
     "role": {
      "type": "string"
     },
     "content": {
      "type": "string"
     }
    }
   }
  },
  "max_tokens": {
   "type": "integer"
  },
  "service_tier": {
   "enum": [
    "flex",
    "default",
    "standard",
    "auto",
    "fast",
    "priority"
   ],
   "type": "string"
  },
  "response_format": {
   "type": "object"
  },
  "reasoning_effort": {
   "type": "string"
  }
 }
}
```

## Response schema (JSON Schema)

```json
{
 "type": "json",
 "example": {
  "id": "chatcmpl-x",
  "model": "gpt-6-astra",
  "usage": {
   "total_tokens": 9,
   "prompt_tokens": 8,
   "completion_tokens": 1
  },
  "object": "chat.completion",
  "choices": [
   {
    "index": 0,
    "message": {
     "role": "assistant",
     "content": "Hello",
     "refusal": null
    },
    "finish_reason": "stop"
   }
  ],
  "service_tier": "flex"
 }
}
```

## More

- Live health (JSON, refreshed every minute): https://www.zero.xyz/c/openai-gpt-6-astra-chat-completions-via-mm-family-x402-b53058d4/health.json
- [Zero catalog index](https://www.zero.xyz/llms.txt)
- [Other services from openai.mm.family](https://www.zero.xyz/host/openai.mm.family/llms.txt)
