29 Services from api.zeroreader.com
Runs chat completions using Meta's Llama 3.2 1B model — the smallest and fastest Llama variant, optimized for low-latency, low-cost inference
$0.001/messageRuns chat completions against Qwen 2.5 Coder 32B, a 32-billion-parameter model specialized in code generation and programming tasks.
$0.006/messageProvides chain-of-thought reasoning via DeepSeek R1 32B for math, logic, and code tasks through an OpenAI-compatible chat completions interface
$0.015/messageRuns inference against Meta's Llama 3.3 70B flagship open-source LLM via Cloudflare Workers AI, returning a chat completion response in OpenAI-compatible format
$0.008/messageConverts input text to spoken audio using MeloTTS synthesis, returning a WAV audio file
$0.005/messageRe-ranks a list of documents against a query to return relevance scores for better search result ordering.
$0.001/messageGenerates fast, lightweight English text embeddings using the BGE Small EN v1.5 model, returning dense float vectors for semantic similarity and retrieval tasks.
$0.001/messageGenerates dense vector embeddings for text using the BGE-M3 multilingual model, supporting 100+ languages
$0.001/messageRuns chat completions using the SEA-LION 27B model, a large language model specialized in Southeast Asian languages
$0.004/messageRuns chat completions using Alibaba's QwQ 32B reasoning-focused language model via a pay-per-call API
$0.006/messageGenerates images from text prompts in approximately 2 seconds using the FLUX.1 Schnell model, returning a PNG image.
$0.01/messageRuns chat completions using Qwen3 30B MoE (3B active parameters) — a fast, capable mixture-of-experts LLM served via pay-per-call x402 micropayment
$0.003/messageRuns chat completions using Meta's Llama 4 Scout 17B mixture-of-experts model via ZeroReader's API, paid per call with USDC.
$0.005/messageRuns multimodal chat completions using Meta's Llama 3.2 11B Vision model, capable of understanding both images and text
$0.005/messageTranscribes audio input into text using OpenAI Whisper, supporting 90+ languages.
$0.008/messageRuns fast, ultra-cheap chat completions using IBM Granite 4.0 Micro via a pay-per-call API
$0.001/messageRuns chat completions using Mistral 7B, a European open-source model optimized for structured output, via a pay-per-call API
$0.002/messageRuns chat completions using Meta's Llama 3.2 3B model, offering a fast and cost-effective balance of speed and quality for straightforward text generation tasks.
$0.002/messageGenerates high-quality English text embeddings using the BGE Large EN v1.5 model, returning dense vector representations of input text or arrays of texts.
$0.002/messageRuns chat completions using Mistral Small 3.1 24B, a strong mid-size European LLM, via a pay-per-call x402 API
$0.004/messageTranslates text between 100 languages using Meta's M2M100 1.2B model, defaulting to English-to-Japanese translation
$0.002/messageClassifies input text as POSITIVE or NEGATIVE with a confidence score using a DistilBERT model.
$0.001/messageRuns chat completions using Zhipu AI's GLM-4 Flash multilingual model, optimized for Japanese and other Asian languages
$0.002/messageGenerates compact vector embeddings from text using Alibaba's Qwen3 0.6B embedding model via the ZeroReader AI API
$0.001/messageRuns fast, cost-effective chat completions using Meta's Llama 3.1 8B model via a pay-per-call x402 API
$0.002/messageRuns chat completions using OpenAI's open-source 20B language model via a pay-per-call x402 API
$0.002/messageRuns chat completions against OpenAI's largest open-source 120B model via a pay-per-call API
$0.01/messageRuns chat completions using Google's Gemma 3 12B open model, served via ZeroReader's pay-per-call API with x402 micropayment support
$0.004/messageTranscribes audio to text using OpenAI's Whisper Large v3 Turbo model, offering high accuracy and fast inference
$0.01/message