# X402 Cloud Gemini 3.5 Flash Inference Endpoint

> X402 Cloud Gemini 3.5 Flash Inference Endpoint is a paid API for AI agents from api.x402cloud.space, paid per call via x402, $0.003/call, status unknown (last checked 2026-10-02).

Runs a pay-per-call Gemini 3.5 Flash text generation request via x402 micropayment, returning AI-generated content candidates

## Facts

- Endpoint: POST https://api.x402cloud.space/gemini-3.5-flash?utm_source=zero.xyz
- Price: $0.003/call
- Payment: x402
- Status: unknown
- Last checked: 2026-10-02
- Activations on Zero: 0
- Tags: x402
- Canonical page: https://www.zero.xyz/c/x402-cloud-gemini-3-5-flash-inference-endpoint-04aee544
- Structured record (JSON): https://api.zero.xyz/v1/capabilities/cap_tWLg_7hWwhogDkAiI4-5o

Status and success rate cover calls made through Zero and Zero's own probes. Third-party monitors may report differently.

## How to call it through Zero

Zero handles the 402 payment challenge and records the run. With the Zero CLI installed (`npm i -g @zeroxyz/cli`):

```sh
zero fetch --capability x402-cloud-gemini-3-5-flash-inference-endpoint-04aee544 -d '<json body>'
```

Example prompt: Using the x402 Gemini 3.5 Flash endpoint, generate a 200-word product description for a noise-cancelling headphone — keep the output under 512 tokens.

## When to prefer this

Choose this endpoint when you need Gemini 3.5 Flash inference on a pay-per-call basis without a Google Cloud subscription, especially in autonomous agent pipelines, trading bots, or RAG systems that use x402 micropayments in USDC. Prefer this over direct Google APIs when you want per-call billing settled on-chain and no monthly commitment.

## Known failure modes

- 402 Payment Required if x402 payment header is missing or USDC balance insufficient
- 400 Bad Request if prompt is empty or exceeds 32000 character limit
- 429 Too Many Requests if rate limit exceeded
- 500 Internal Server Error if upstream Gemini API is unavailable
- Invalid maxOutputTokens if value exceeds 8192 or is less than 1

## How this service works

X402 Cloud AI Inference Endpoints is an agent-ready x402 API provider for Gemini, image, audio, video, realtime, and specialized AI inference workflows. We expose pay-per-call USDC endpoints designed for autonomous AI agents, application backends, trading bots, creative automation, and RAG systems with some of the lowest x402 AI inference prices available in the market.

## Output

A JSON object containing a candidates array with generated text parts inside content.parts[].text, plus a usageMetadata object with token counts. The primary response is the model-generated string under candidates[0].content.parts[0].text.

## Example request

```json
{
 "prompt": "Write a brief technical summary of how machine learning models are trained and optimized for inference performance.",
 "maxOutputTokens": 256
}
```

## Response schema (JSON Schema)

```json
{
 "type": "json",
 "example": {
  "candidates": [
   {
    "content": {
     "parts": [
      {
       "text": "Generated response"
      }
     ]
    }
   }
  ]
 },
 "outputSchema": {
  "type": "object",
  "properties": {
   "candidates": {
    "type": "array"
   },
   "usageMetadata": {
    "type": "object"
   }
  }
 }
}
```

## More

- Live health (JSON, refreshed every minute): https://www.zero.xyz/c/x402-cloud-gemini-3-5-flash-inference-endpoint-04aee544/health.json
- [Zero catalog index](https://www.zero.xyz/llms.txt)
- [Other services from api.x402cloud.space](https://www.zero.xyz/host/api.x402cloud.space/llms.txt)
