Transcription APIs
Compare Speech-to-Text APIs and Transcription Pricing
Find automatic speech recognition models for meetings, podcasts, voice notes, call analytics, and real-time transcription. Per-minute prices are ranked separately from token-based audio prices.
Current catalog
0
speech-to-text models with verified catalog metadata
Pricing content refreshed August 16, 2026
Available Speech-to-Text Models
Comparable models are ordered by their current normalized price.
0 models
Catalog data is being refreshed
Model comparisons will appear when the public catalog API is available.
How to choose
Pricing methodology
OneInfer compares only positive prices with the same billing unit. Per-minute, per-character, per-token, per-image, per-video, and per-second rates remain separate. Prices can change, so the current model page and console remain the source of truth.
Frequently asked questions
How is the cheapest speech-to-text API selected?
OneInfer compares current positive per-audio-minute prices when providers expose that unit. Models using another billing unit are shown but not included in that ranking.
Do all transcription models provide timestamps?
No. Timestamp support is a model capability and should be confirmed on the individual model page.