Skip to content

PricingSaaS joins Willingness to PayRead the announcement →

Pricing change·Apr 1 → Jul 4, 2025·Q2 2025·AI & Machine Learning

Groq Introduces Llama 4 Maverick and Qwen3 Models Amid Removal of Multiple High-Cost Plans and End of 50% Batch Processing Discount

GroqPricingPlan removedNew planLimit increaseDiscount removed

What changed

Plan removed

Llama 3.2 11B Vision 8k (Preview) model removed (was $0.18 input/$0.18 output)

Plan removed

DeepSeek R1 Distill Qwen 32B model removed (was $0.69 input/$0.69 output)

Plan removed

Qwen 2.5 32B Instruct model removed (was $0.79 input/$0.79 output)

New plan

Llama 4 Maverick (17Bx128E) model added with 562 tokens/sec speed, $0.20 input/$0.60 output per M tokens

Plan removed

Llama 3.2 3B (Preview) model removed (was $0.06 input/$0.06 output)

New plan

Qwen3 32B model added with 662 tokens/sec speed, $0.29 input/$0.59 output per M tokens

Plan removed

Qwen 2.5 Coder 32B Instruct model removed (was $0.79 input/$0.79 output)

Plan removed

Llama 3.3 70B SpecDec model removed (was $0.59 input/$0.99 output)

Limit increase

Whisper Large v3 Turbo speed factor increased from 216x to 228x

Plan removed

Llama 3.2 90B Vision 8k (Preview) model removed (was $0.90 input/$0.90 output)

Limit increase

Llama 3.1 8B Instant speed increased from 750 to 840 tokens/sec

Plan removed

Llama 3.2 1B (Preview) model removed (was $0.04 input/$0.04 output)

New plan

Llama 4 Scout (17Bx16E) model added with 594 tokens/sec speed, $0.11 input/$0.34 output per M tokens

New plan

Llama Guard 4 12B model added with 325 tokens/sec speed, $0.20 input/$0.20 output per M tokens

Limit increase

Whisper V3 Large speed factor increased from 189x to 217x

Limit increase

Llama 3 8B speed increased from 1250 to 1345 tokens/sec

Limit increase

DeepSeek R1 Distill Llama 70B speed increased from 275 to 400 tokens/sec

Discount removed

50% batch processing discount promotion ended (was doubled from 25% through end of April 2025)

Limit increase

Llama 3.3 70B Versatile speed increased from 275 to 394 tokens/sec

Before and after

Apr 1, 2025 vs Jul 4, 2025
Pulse.
100%
Change 1 of 24 · plan removed

Llama 3.2 11B Vision 8k (Preview) model removed (was $0.18 input/$0.18 output)

Sign in to view the full comparison

See the before/after screenshots and every marked-up change.