groq/llama-3.1-8b-instant
Groq textgroq/llama-3.1-8b-instant
Input
$0.0500
per 1M input tokens
Output
$0.0800
per 1M output tokens
Cache read
n/a
per 1M cached read tokens
Cache write
n/a
per 1M cache write tokens
Context window
131,072
Max output
131,072
Effective date
Aug 14, 2026
Estimate a workload
Enter token counts to see the cost at this model's rates.
Estimated cost:
$0.00
Price history
blended $ per 1M tokens (3:1)| Effective | Input | Output | Blended |
|---|---|---|---|
| Aug 14, 2026 | $0.05 | $0.08 | $0.06 |
| Jun 15, 2026 | $0.05 | $0.08 | $0.06 |
Other Groq models
groq/meta-llama/llama-prompt-guard-2-22m
· $0.0300/Mtok out
groq/meta-llama/llama-prompt-guard-2-86m
· $0.0400/Mtok out
groq/gemma-7b-it
· $0.0800/Mtok out
groq/openai/gpt-oss-20b
· $0.3000/Mtok out
groq/openai/gpt-oss-safeguard-20b
· $0.3000/Mtok out
groq/meta-llama/llama-4-scout-17b-16e-instruct
· $0.3400/Mtok out
Source: litellm. Confirm against the provider's official pricing before relying on these figures.