Skip to main content

Groq model cost planner

Llama 3.3 70B

Production Llama 3.3 70B general-purpose model hosted on GroqCloud.

Pricing details

Llama 3.3 70B pricing

These values come from the central model pricing data used by the calculator.

Input price

$0.59 / 1M tokens

Output price

$0.79 / 1M tokens

Pricing unit

Per 1M tokens

Verified provider pricing

GroqLlama 3.3 70B Versatile 128k

Stable
Last verified Jul 20, 2026Official source

Estimates vary by tier, modality, caching, batch usage, token thresholds, and other provider conditions.

Pricing notes
Production Llama 3.3 70B general-purpose model hosted on GroqCloud.

Deterministic cost examples

These shared examples apply the listed token prices directly. They illustrate arithmetic only and are not claims about a typical user workload.

1M input tokens

$0.59

Input-only example at the listed standard token price.

1M output tokens

$0.79

Output-only example at the listed standard token price.

Example 1M-token mix

$0.64

Defined as 750K input tokens plus 250K output tokens.

Calculator

Calculate Llama 3.3 70B API cost

Groq and Llama 3.3 70B are preselected so the estimate starts on this exact model.

Model selection

Choose the provider and model you want to estimate.
GroqLlama 3.3 70B Versatile 128k

Input price

$0.59 / 1M tokens

Output price

$0.79 / 1M tokens

GroqLlama 3.3 70B Versatile 128kStable
Official source

Verified Jul 20, 2026. Estimates vary by usage and provider pricing conditions.

Usage assumptions

Estimate traffic and token usage for an average request.
Active seats, customers, or internal users.
Average AI calls per user each day.
Prompt, history, and retrieved context per request.
Generated answer length; SaaS founders should test long replies.
Use 30 for always-on products or fewer for batch jobs.

Estimated results

Run the calculator to see projected cost and usage volume.

Enter your usage details, then select Calculate estimate to see your projected cost.

Estimated cost = input usage cost + output usage cost + supported optional charges.

Verified model facts

Llama 3.3 70B model details

Only fields present in canonical pricing, identity, or reviewed catalog data are shown.
Provider
Groq
Model family
Llama 3.3
Lifecycle
Stable
API availability
Available
Catalog category
Generative
Context window
131,072 tokens
Max output tokens
32,768 tokens
Input modalities
Text
Output modalities
Text
Capabilities
Text generation, Function calling, Structured outputs
Service tier
Standard
Context tier
Default
Pricing region
global

Use-case calculators

Plan workflows with Llama 3.3 70B

These calculator pages are selected from canonical model capabilities and provider context.

Related models

Other Groq models

Run the same assumptions against another model from this provider.

Alternatives

Capability-linked lower-price alternatives

Every candidate shares at least one listed capability, uses the same pricing unit, and has a lower canonical input or output price. Price does not establish equivalent quality.

Groq

Llama 3.1 8B

Lower input and output prices. Input $0.05 / 1M tokens, output $0.08 / 1M tokens.

Shared listed capabilities: Text.

Compare cost

Groq

GPT OSS 20B

Lower input and output prices. Input $0.075 / 1M tokens, output $0.30 / 1M tokens.

Shared listed capabilities: Text.

Compare cost

Groq

Safety GPT OSS 20B

Lower input and output prices. Input $0.075 / 1M tokens, output $0.30 / 1M tokens.

Shared listed capabilities: Text.

Compare cost

Entity relationships

Provider, calculator, and exact comparisons

Continue to this model's canonical provider, provider calculator, or an exact related comparison.