Lowest input price
Llama 3.1 8B
$0.05 / 1M
Lowest listed cost for 1M input tokens.
View model pricingProvider pricing hub
Pricing summary
Models tracked
6
Cached-price models
0
Context window range
131,072 tokens
Latest verification
Jul 20, 2026
Model pricing
Swipe sideways to see all columns.
| Model | Input price | Output price | Cached input | Context window | Capabilities | Verification | Action |
|---|---|---|---|---|---|---|---|
GPT OSS 20B GPT OSS | $0.075 / 1M tokens | $0.30 / 1M tokens | Not available | 131,072 tokens | text, reasoning | GroqGPT OSS 20B 128kStable Official sourceVerified Jul 13, 2026. Estimates vary by usage and provider pricing conditions. | Open model |
Safety GPT OSS 20B GPT OSS | $0.075 / 1M tokens | $0.30 / 1M tokens | Not available | 131,072 tokens | text, reasoning | GroqGPT OSS Safeguard 20BStable Official sourceVerified Jul 13, 2026. Estimates vary by usage and provider pricing conditions. | Open model |
GPT OSS 120B GPT OSS | $0.15 / 1M tokens | $0.60 / 1M tokens | Not available | 131,072 tokens | text, reasoning | GroqGPT OSS 120B 128kStable Official sourceVerified Jul 13, 2026. Estimates vary by usage and provider pricing conditions. | Open model |
Llama 3.3 70B Llama 3.3 | $0.59 / 1M tokens | $0.79 / 1M tokens | Not available | 131,072 tokens | text | GroqLlama 3.3 70B Versatile 128kStable Official sourceVerified Jul 20, 2026. Estimates vary by usage and provider pricing conditions. | Open model |
Llama 3.1 8B Llama 3.1 | $0.05 / 1M tokens | $0.08 / 1M tokens | Not available | 131,072 tokens | text | GroqLlama 3.1 8B Instant 128kStable Official sourceVerified Jul 20, 2026. Estimates vary by usage and provider pricing conditions. | Open model |
Qwen/Qwen3.6-27B Qwen 3.6 | $0.60 / 1M tokens | $3.00 / 1M tokens | Not available | 131,072 tokens | text, reasoning | GroqQwen 3.6 27B 131kStable Official sourceVerified Jul 20, 2026. Estimates vary by usage and provider pricing conditions. | Open model |
Lowest input price
$0.05 / 1M
Lowest listed cost for 1M input tokens.
View model pricingLowest output price
$0.08 / 1M
Lowest listed cost for 1M output tokens.
View model pricingLowest combined token price
$0.13 total
Cost of 1M input tokens plus 1M output tokens.
View model pricingCalculator
Run the calculator to see projected cost and usage volume.
Enter your usage details, then select Calculate estimate to see your projected cost.
Estimated cost = input usage cost + output usage cost + supported optional charges.
Reviewed starting points
Reasoning
Listed in the reviewed reasoning catalog category.
Use calculatorLowest combined price
Lowest combined verified token price among compatible generation models.
Use calculatorVision
Reviewed metadata lists image input support.
View model detailsSafety
Listed in a reviewed safety or moderation catalog category.
Use calculatorModel catalog
Calculator actions appear only for exact model IDs with compatible verified token pricing.
8 of 8 models
openai/gpt-oss-20bGPT OSS
OpenAI open-weight reasoning model hosted on GroqCloud with browser search and code execution support.
Capabilities
Endpoints
Deployment
Verified token pricing
$0.075 input / $0.3 output per 1M tokens
Use calculatoropenai/gpt-oss-120bGPT OSS
OpenAI flagship open-weight 120B model hosted on GroqCloud with reasoning and built-in tool support.
Capabilities
Endpoints
Deployment
Verified token pricing
$0.15 input / $0.6 output per 1M tokens
Use calculatorqwen/qwen3-32bQwen3
Preview Qwen3 reasoning model hosted on GroqCloud.
Capabilities
Endpoints
Deployment
Catalog details only
No compatible public calculator price is listed.
qwen/qwen3.6-27bQwen 3.6
Preview Qwen 3.6 27B model hosted on GroqCloud.
Capabilities
Endpoints
Deployment
Verified token pricing
$0.6 input / $3 output per 1M tokens
Use calculatormeta-llama/llama-4-scout-17b-16e-instructLlama 4
Preview multimodal Llama 4 Scout model hosted on GroqCloud.
Capabilities
Endpoints
Deployment
Catalog details only
No compatible public calculator price is listed.
llama-3.3-70b-versatileLlama 3.3
Production Llama 3.3 70B general-purpose model hosted on GroqCloud.
Capabilities
Endpoints
Deployment
Verified token pricing
$0.59 input / $0.79 output per 1M tokens
Use calculatorllama-3.1-8b-instantLlama 3.1
Production low-latency Llama 3.1 8B model hosted on GroqCloud.
Capabilities
Endpoints
Deployment
Verified token pricing
$0.05 input / $0.08 output per 1M tokens
Use calculatoropenai/gpt-oss-safeguard-20bGPT OSS
Preview safety-focused GPT OSS model hosted on GroqCloud.
Capabilities
Endpoints
Deployment
Verified token pricing
$0.075 input / $0.3 output per 1M tokens
Use calculatorPricing sources
Open Groq pricing references for current provider terms, tiers, and availability notes.
Open sourceReview CostRivo's cross-provider pricing table, verification dates, source links, and lifecycle labels.
Open sourceFAQ
The calculator multiplies input, output, and cached input tokens by the selected model pricing, then scales the result by request volume.
Lowest input price: Llama 3.1 8B at $0.05 per 1M tokens; Lowest output price: Llama 3.1 8B at $0.08 per 1M tokens; Lowest combined token price: Llama 3.1 8B at $0.13 for 1M input plus 1M output tokens. Each criterion is calculated independently from eligible verified token prices.
Cached input pricing is not listed for the current provider models in this data set.
The most recent model verification shown by CostRivo is Jul 20, 2026. Individual model rows retain their own source and verification details.
Pricing updates
Get notified when AI model prices change, new providers are added, product updates ship, launch notes go out, or Costrivo introduces future premium planning features.
No spam. Pricing and product updates only. We only store your email, this page, and signup time.