# OpenAI GPT-6 Astra Long (>272K context) via mm.family

> OpenAI GPT-6 Astra Long (>272K context) via mm.family is a paid API for AI agents from openai.mm.family, paid per call via x402, $0.002477/call, status unknown (last checked 2026-10-02).

Pay-per-call OpenAI GPT-6 Astra chat completions for long-context inputs (>272K tokens), billed in USDC with configurable speed/cost tiers

## Facts

- Endpoint: POST https://openai.mm.family/x402/v1/models/gpt-6-astra-long/chat/completions?utm_source=zero.xyz
- Price: $0.002477/call
- Payment: x402
- Status: unknown
- Last checked: 2026-10-02
- Activations on Zero: 0
- Tags: x402
- Canonical page: https://www.zero.xyz/c/openai-gpt-6-astra-long-272k-context-via-mm-family-3d608899
- Structured record (JSON): https://api.zero.xyz/v1/capabilities/cap_UM8NOv2IrwhakRO0qciMG

Status and success rate cover calls made through Zero and Zero's own probes. Third-party monitors may report differently.

## How to call it through Zero

Zero handles the 402 payment challenge and records the run. With the Zero CLI installed (`npm i -g @zeroxyz/cli`):

```sh
zero fetch --capability openai-gpt-6-astra-long-272k-context-via-mm-family-3d608899 -d '<json body>'
```

Example prompt: Send this 400,000-token legal document and my analysis questions as a chat prompt to GPT-6 Astra Long using the flex service tier with a max_tokens of 4096 — bill the call in USDC.

## When to prefer this

Choose this endpoint when your input context exceeds 272K tokens and you need GPT-6 Astra specifically — it is the long-context variant designed for inputs that overflow the standard gpt-6-astra endpoint. Prefer it over gpt-6-astra when processing full codebases, lengthy legal documents, or large research corpora. Use the flex tier for cost efficiency, standard for balanced performance, or fast for latency-sensitive workloads. Prefer this over gpt-5.x sibling models when you need the highest capability frontier model for complex long-context reasoning.

## Known failure modes

- Insufficient USDC balance — payment rejected before completion
- Input token count does not actually exceed 272K threshold — should use gpt-6-astra instead
- Invalid service_tier value — must be one of flex, default, standard, auto, fast, priority
- max_tokens too large for available budget — quoted charge may be rejected
- Model variant string malformed — must be gpt-6-astra-long or gpt-6-astra-long:<level>
- Empty response (zero-length answer) — billed at $0 but counts as a call
- Network timeout on very large context processing under fast tier

## How this service works

gpt-6-astra-long: OpenAI GPT 6 Astra (>272K input tokens) chat completions, paid per call in USDC. Per 1M tokens in/out: Flex (default) 10.0/37.5, Standard 20.0/75.0, Fast 40.0/150.0; pick with service_tier. Plus $0.0005 (Base) or $0.0005 (Solana) per call; the quote charges input plus 10% of max_tokens. Levels: low, medium, high, xhigh (model gpt-6-astra-long:<level>). Empty answers are free. Rates: https://openai.mm.family/x402/pricing

## Output

Returns a standard OpenAI-compatible chat completion JSON object with the assistant's reply message, finish reason, token usage breakdown (prompt, completion, total), model name, and the service tier used for the request.

## Request schema (JSON Schema)

```json
{
 "type": "object",
 "properties": {
  "model": {
   "enum": [
    "gpt-6-astra-long",
    "gpt-6-astra-long:low",
    "gpt-6-astra-long:medium",
    "gpt-6-astra-long:high",
    "gpt-6-astra-long:xhigh"
   ],
   "type": "string"
  },
  "tools": {
   "type": "array"
  },
  "stream": {
   "type": "boolean"
  },
  "messages": {
   "type": "array",
   "items": {
    "type": "object",
    "required": [
     "role",
     "content"
    ],
    "properties": {
     "role": {
      "type": "string"
     },
     "content": {
      "type": "string"
     }
    }
   }
  },
  "max_tokens": {
   "type": "integer"
  },
  "service_tier": {
   "enum": [
    "flex",
    "default",
    "standard",
    "auto",
    "fast",
    "priority"
   ],
   "type": "string"
  },
  "response_format": {
   "type": "object"
  },
  "reasoning_effort": {
   "type": "string"
  }
 }
}
```

## Response schema (JSON Schema)

```json
{
 "type": "json",
 "example": {
  "id": "chatcmpl-x",
  "model": "gpt-6-astra-long",
  "usage": {
   "total_tokens": 9,
   "prompt_tokens": 8,
   "completion_tokens": 1
  },
  "object": "chat.completion",
  "choices": [
   {
    "index": 0,
    "message": {
     "role": "assistant",
     "content": "Hello",
     "refusal": null
    },
    "finish_reason": "stop"
   }
  ],
  "service_tier": "flex"
 }
}
```

## More

- Live health (JSON, refreshed every minute): https://www.zero.xyz/c/openai-gpt-6-astra-long-272k-context-via-mm-family-3d608899/health.json
- [Zero catalog index](https://www.zero.xyz/llms.txt)
- [Other services from openai.mm.family](https://www.zero.xyz/host/openai.mm.family/llms.txt)
