Audio and voice APIs

Compare Audio AI Models and API Pricing

Explore audio models for transcription, speech synthesis, voice applications, and generative audio. Because providers use different billing units, OneInfer ranks prices only within the same comparable unit.

Current catalog

0

audio models with verified catalog metadata

Pricing content refreshed August 16, 2026

Available Audio Models

Comparable models are ordered by their current normalized price.

0 models

Catalog data is being refreshed

Model comparisons will appear when the public catalog API is available.

How to choose

Voice agents
Transcription
Speech synthesis
Audio generation

Pricing methodology

OneInfer compares only positive prices with the same billing unit. Per-minute, per-character, per-token, per-image, per-video, and per-second rates remain separate. Prices can change, so the current model page and console remain the source of truth.

Frequently asked questions

What is the cheapest audio model on OneInfer?

The lowest-priced model is calculated from the current catalog within each comparable billing unit. Per-minute, per-character, and token-based prices are not mixed.

Why are some audio models not ranked together?

Audio providers charge by different units. A transcription model billed per minute cannot be honestly ranked against a speech model billed per character without a common workload assumption.