Speech synthesis APIs

Compare Text-to-Speech Models and API Pricing

Explore speech synthesis models for voice agents, narration, accessibility, and conversational products. Per-character and per-minute rates remain separate comparison groups.

Current catalog

0

text-to-speech models with verified catalog metadata

Pricing content refreshed August 16, 2026

Available Text-to-Speech Models

Comparable models are ordered by their current normalized price.

0 models

Catalog data is being refreshed

Model comparisons will appear when the public catalog API is available.

How to choose

Voice agents
Narration
Accessibility
Conversational applications

Pricing methodology

OneInfer compares only positive prices with the same billing unit. Per-minute, per-character, per-token, per-image, per-video, and per-second rates remain separate. Prices can change, so the current model page and console remain the source of truth.

Frequently asked questions

How are text-to-speech prices compared?

OneInfer compares models only when they use the same billing unit, such as price per million characters. Different units are not combined into a misleading ranking.

Does every TTS model support voice cloning?

No. Voice cloning is shown only when it is explicitly provided in verified model metadata.