Fireworks Announces Major 30% Price Reduction on H200 141 GB GPU Hourly Rate, Now Just $6.99
What changed
Developer plan tier removed from main pricing section (replaced with product-based presentation)
Qwen3 30B model added at $0.15 input / $0.60 output per 1M tokens
H200 141 GB GPU hourly rate decreased from $9.99 to $6.99 (30% reduction)
Fine-tuning pricing restructured: MoE tier pricing ($2.00-$6.00 per 1M tokens) removed and replaced with DeepSeek R1/V3 specific pricing at $10.00 per 1M training tokens
DeepSeek R1 0528 (Fast) model added at $3.00 input / $8.00 output per 1M tokens
Enterprise plan tier removed from main pricing section (replaced with product-based presentation)
Qwen3 235B model added at $0.22 input / $0.88 output per 1M tokens
Yi Large model removed (was priced at $3.00 per 1M tokens)
B200 180 GB GPU added at $11.99 per hour
Before and after
Apr 9, 2025 vs Jul 6, 2025Developer plan tier removed from main pricing section (replaced with product-based presentation)
Sign in to view the full comparison
See the before/after screenshots and every marked-up change.
More from Fireworks
Full history →- Week 40, 2026·packagingFireworks adds GLM 5.3 Flash with 200K context, per-token training pricing
- Week 37, 2026·improvedFireworks rebuilds Serverless Training API catalog: drops Qwen 3.5 9B, adds DeepSeek V4 Flash and Muse Glimmer 30B
- Week 34, 2026·increaseFireworks: On-Demand GPU pricing up 11-30% from Sep 1; region surcharge to 1.5x
- Week 33, 2026·improvedFireworks: New GB300 GPU tier ($18/hr) + region premium pricing
- Week 32, 2026·packagingFireworks: New Serverless Training API added (Qwen 3.5 9B, 3.6 27B, Kimi K3)
