Transcription APIs

Compare Speech-to-Text APIs and Transcription Pricing

Find automatic speech recognition models for meetings, podcasts, voice notes, call analytics, and real-time transcription. Per-minute prices are ranked separately from token-based audio prices.

Current catalog

0

speech-to-text models with verified catalog metadata

Pricing content refreshed August 16, 2026

Available Speech-to-Text Models

Comparable models are ordered by their current normalized price.

0 models

Catalog data is being refreshed

Model comparisons will appear when the public catalog API is available.

How to choose

Meeting transcription
Podcast transcription
Call analytics
Voice interfaces

Pricing methodology

OneInfer compares only positive prices with the same billing unit. Per-minute, per-character, per-token, per-image, per-video, and per-second rates remain separate. Prices can change, so the current model page and console remain the source of truth.

Frequently asked questions

How is the cheapest speech-to-text API selected?

OneInfer compares current positive per-audio-minute prices when providers expose that unit. Models using another billing unit are shown but not included in that ranking.

Do all transcription models provide timestamps?

No. Timestamp support is a model capability and should be confirmed on the individual model page.