# dt0ur.online Text-to-Speech (SpeechT5)

> dt0ur.online Text-to-Speech (SpeechT5) is a paid API for AI agents from dt0ur.online, paid per call via x402, $0.15/call, status down (last checked 2026-10-03).

Synthesizes natural-sounding human speech audio from a text string using SpeechT5 neural TTS models with selectable speaker voice profiles.

## Facts

- Endpoint: POST https://dt0ur.online/api/inference/tts?utm_source=zero.xyz
- Price: $0.15/call
- Payment: x402
- Status: down
- Last checked: 2026-10-03
- Activations on Zero: 0
- Tags: x402
- Canonical page: https://www.zero.xyz/c/dt0ur-online-text-to-speech-speecht5-20065c29
- Structured record (JSON): https://api.zero.xyz/v1/capabilities/cap_JLM4JsS8FisxtV7bDiChS

Status and success rate cover calls made through Zero and Zero's own probes. Third-party monitors may report differently.

## How to call it through Zero

Zero handles the 402 payment challenge and records the run. With the Zero CLI installed (`npm i -g @zeroxyz/cli`):

```sh
zero fetch --capability dt0ur-online-text-to-speech-speecht5-20065c29 -d '<json body>'
```

Example prompt: Can you synthesize this paragraph into spoken audio using the default voice: 'Welcome to our platform. We're glad you're here and hope you enjoy the experience.'

## When to prefer this

Choose this endpoint when you need on-demand neural speech synthesis from text with multi-speaker support and pay-per-call pricing via x402. It is well-suited for agents that need to generate spoken audio dynamically without a long-term TTS subscription, especially when integrated into agentic pipelines that already use x402 micropayments.

## Known failure modes

- Empty or missing text field returns a validation error
- Unsupported or unknown voice profile identifier may fall back to default or return an error
- Very long text inputs may exceed processing limits or increase latency
- Payment not included or insufficient USDC causes 402 Payment Required response
- Service unavailability results in 5xx error with no audio output

## How this service works

Synthesizes natural human speech from text prompts using SpeechT5 neural voice generation models with multi-speaker support.

## Output

Returns a synthesized audio file or audio stream representing the input text spoken aloud in the selected voice profile, generated by SpeechT5 neural TTS models. The audio is human-like in cadence and intonation.

## More

- Live health (JSON, refreshed every minute): https://www.zero.xyz/c/dt0ur-online-text-to-speech-speecht5-20065c29/health.json
- [Zero catalog index](https://www.zero.xyz/llms.txt)
- [Other services from dt0ur.online](https://www.zero.xyz/host/dt0ur.online/llms.txt)
