Skip to content

PricingSaaS joins Willingness to PayRead the announcement →

Pricing change·Dec 29 → Mar 30, 2026·Q1 2026·Marketing & Content

Fireworks hikes H100 GPU pricing 50% while trimming AI model lineup

FireworksPricingFeature removedPrice increase

What changed

Feature removed

Streaming ASR v1/v2 removed from STT pricing. DeepSeek V3, Kimi K2 Instruct/Thinking, and gpt-oss-20b lost cached pricing. Caching footnote updated to 50% universal rate.

Feature removed

DeepSeek R1 0528 ($1.35 input/$0.68 cached/$5.4 output), Qwen3 235B Family ($0.22 input/$0.88 output), and Qwen3 Coder 480B ($0.45 input/$0.23 cached/$1.80 output) removed from serverless pricing table.

Price increase

H100 80 GB GPU on-demand price increased from $4.00/hr to $6.00/hr (50% increase). MiniMax M2 family cached token price decreased from $0.15 to $0.03 per 1M tokens.

Before and after

Dec 29, 2025 vs Mar 30, 2026
Pulse.
100%
Change 1 of 3 · feature removed

Streaming ASR v1/v2 removed from STT pricing. DeepSeek V3, Kimi K2 Instruct/Thinking, and gpt-oss-20b lost cached pricing. Caching footnote updated to 50% universal rate.

Sign in to view the full comparison

See the before/after screenshots and every marked-up change.