# KiteEndpoints.ai - Speech to Text

> KiteEndpoints.ai - Speech to Text is a paid API for AI agents from kite-endpoints.vercel.app, paid per call via x402, $0.09/call, status unknown (last checked 2026-09-14).

Converts spoken audio input into transcribed text via a pay-per-request API using USDC on Base chain

## Facts

- Endpoint: POST https://kite-endpoints.vercel.app/api/speech-to-text
- Price: $0.09/call
- Payment: x402
- Status: unknown
- Last checked: 2026-09-14
- Activations on Zero: 0
- Tags: x402
- Canonical page: https://www.zero.xyz/c/kiteendpoints-ai-speech-to-text-0f6a6303
- Structured record (JSON): https://api.zero.xyz/v1/capabilities/cap_ZS5BZKcnjE3uqsWp5wQ0f

Status and success rate cover calls made through Zero and Zero's own probes. Third-party monitors may report differently.

## How to call it through Zero

Zero handles the 402 payment challenge and records the run. With the Zero CLI installed (`npm i -g @zeroxyz/cli`):

```sh
zero fetch --capability kiteendpoints-ai-speech-to-text-0f6a6303 -d '<json body>'
```

Example prompt: Can you transcribe this audio recording for me — take the speech and convert it into written text so I can read what was said?

## When to prefer this

Choose this endpoint when you need a simple, pay-per-request speech-to-text conversion without a subscription commitment, and when paying in USDC on Base chain is acceptable or preferred. Suitable for agents that process audio infrequently or need to avoid upfront API key setup with major ASR providers.

## Known failure modes

- Unsupported audio format or codec may result in processing failure
- Audio file too large or too long may exceed limits
- Poor audio quality or heavy background noise may produce inaccurate transcription
- Missing or malformed audio input returns an error
- Payment failure on Base chain (insufficient USDC) results in 402 rejection
- Network timeout for longer audio files

## How this service works

400+ paid API endpoints. Social media, AI, crypto, finance, weather, ecommerce, tools. Pay per request with USDC on Base chain.

## Output

Returns a transcription result object containing the converted text from the submitted audio input, representing the spoken words as a text string.

## Response schema (JSON Schema)

```json
{
 "type": "object",
 "example": {
  "result": {
   "data": "example response"
  }
 },
 "properties": {
  "result": {
   "type": "object",
   "description": "API response data"
  }
 }
}
```

## More

- Live health (JSON, refreshed every minute): https://www.zero.xyz/c/kiteendpoints-ai-speech-to-text-0f6a6303/health.json
- [Zero catalog index](https://www.zero.xyz/llms.txt)
- [Other services from kite-endpoints.vercel.app](https://www.zero.xyz/host/kite-endpoints.vercel.app/llms.txt)
