Embedding model for semantic search, retrieval, clustering, and ranking pipelines
Gemini Embedding 001 by google is offered by 6 providers on this page. Public prices are shown for 4 of them. Its official list price is $0.15 per 1M input tokens and $0 per 1M output tokens.
The context window is 2,048 tokens at the reference host, but hosts report different limits, from 2,048 to 8,192, so the usable window depends on the provider you pick. Its training knowledge cuts off at 2025-05.
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| GoogleOfficial | First-party | $0.15 | $0 | — | — | 2,048 | 1 | |
| Merge Gateway | Gateway | $0.15 | $0 | — | — | 2,048 | — | |
| Tempr Gateway | Gateway | $0.15 | $0 | — | — | 2,048 | 1 | |
| VertexOfficial | Cloud | $0.15 | $0 | — | — | 2,048 | 1 | |
| Vercel AI Gateway | Cloud | — | — | — | — | 8,192 ⚠ | 1,536 | hostTable.undisclosed |
| SAP AI Core gemini-embedding | Cloud | — | — | — | — | 2,048 | 1 | hostTable.undisclosed |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.