Embedding and retrieval APIs
Compare Embedding Models and API Pricing
Find embedding models for semantic search, retrieval-augmented generation, recommendations, clustering, and document processing through a consistent API.
Current catalog
0
embedding models with verified catalog metadata
Pricing content refreshed August 16, 2026
Available Embedding Models
Comparable models are ordered by their current normalized price.
0 models
Catalog data is being refreshed
Model comparisons will appear when the public catalog API is available.
How to choose
Pricing methodology
OneInfer compares only positive prices with the same billing unit. Per-minute, per-character, per-token, per-image, per-video, and per-second rates remain separate. Prices can change, so the current model page and console remain the source of truth.
Frequently asked questions
How are embedding model prices compared?
Embedding models are compared using their positive input-token price per one million tokens when that pricing is available.
Can I use these models with my existing vector database?
Yes. Store the returned vectors in your preferred vector database and use them for retrieval or similarity search.