← Model list

Llama 3.1 8B (decentralized)

meta·meta/meta-llama-3-1-8b-instruct-fp8·GA·Open weights·llama series

Compact GPT model for low-latency assistance and high-volume workloads

At a glance
Reference price$0.02 / $0.03 per 1M
Cache read $0.01 · NanoGPT
Lowest paid$0.02 / $0.03
NanoGPT Gateway
Context128,000
Output limit16,384
Capabilities
ReasoningTool useStructured output? TemperatureAttachments
Modalities
Text
Knowledge cutoff
Released / updated2024-01-01 / 2024-07-23

Quality & performance

Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.

Available at 1 providers1 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
NanoGPT
Meta-Llama-3-1-8B-Instruct-FP8
Gateway$0.02$0.03$0.01128,00016,384

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Your usage cost

1NanoGPT$4.30
The cheapest paid channel is the only channel.

Related models

Price historyone sample accumulated per data sync

Input listOutput listMin blended
$0.03$02026-08-052026-08-132026-08-05 · Input list $0.022026-08-06 · Input list $0.022026-08-07 · Input list $0.022026-08-08 · Input list $0.022026-08-09 · Input list $0.022026-08-10 · Input list $0.022026-08-11 · Input list $0.022026-08-12 · Input list $0.022026-08-13 · Input list $0.022026-08-05 · Output list $0.032026-08-06 · Output list $0.032026-08-07 · Output list $0.032026-08-08 · Output list $0.032026-08-09 · Output list $0.032026-08-10 · Output list $0.032026-08-11 · Output list $0.032026-08-12 · Output list $0.032026-08-13 · Output list $0.032026-08-05 · Min blended $0.0222026-08-06 · Min blended $0.0222026-08-07 · Min blended $0.0222026-08-08 · Min blended $0.0222026-08-09 · Min blended $0.0222026-08-10 · Min blended $0.0222026-08-11 · Min blended $0.0222026-08-12 · Min blended $0.0222026-08-13 · Min blended $0.022