← Model list
Hermes Medium
misc·misc/hermes-medium·GA·Closed·hermes series
Compact GPT model for low-latency assistance and high-volume workloads
At a glance
Reference price$0.315 / $1.26 per 1M
Cache read $0.158 · NanoGPT
Lowest paid$0.315 / $1.26
NanoGPT Gateway
Context204,800
Output limit131,072
Capabilities
✓ Reasoning✓ Tool use✓ Structured output? TemperatureAttachments
Modalities
Text
Knowledge cutoff—
Released / updated2026-05-11 / 2026-05-11
Quality & performance
Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.
Available at 1 providers1 with public prices
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| NanoGPT | Gateway | $0.315 | $1.26 | $0.158 | — | 204,800 | 131,072 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1NanoGPT$107.10
The cheapest paid channel is the only channel.
Benchmark
No upstream benchmark data for this model. For quality, see the Artificial Analysis intelligence score above.
Reasoning control
Upstream provides no control info
1 / 1 providers expose no reasoning control (reasoning_options: []).
Related models
Hermes 3 Llama 3.1 405bsame series$1.10 / $3.00Hermes Highsame series$5.00 / $25.00Hermes Lowsame series$0.25 / $1.50DeepSeek V4 Flashcheaper alternative$0.14 / $0.28GPT-5 Nanocheaper alternative$0.05 / $0.40GLM-4.7-Flashcheaper alternative$0 / $0MiMo-V2.5cheaper alternative$0.14 / $0.28
Price historyone sample accumulated per data sync
Input listOutput listMin blended