Skip to content

PricingSaaS joins Willingness to PayRead the announcement →

Pricing change·Jun 15 → Jun 22, 2026·Week 26, 2026·AI & Machine Learning

Together AI slashes GPU and AI inference prices up to 25%

Together AIPricingPrice decrease

What changed

Price decrease

GPU Clusters on-demand rates cut across all hardware: H100 $5.49→$4.79, H200 $6.79→$5.99, B200 $9.95→$8.19/hour. Reserved rates also reduced 12–25% across H100, H200, and B200 tiers.

Price decrease

DeepSeek V4 Pro serverless inference prices reduced: input from $2.10 to $1.74/1M tokens (~17% cut) and output from $4.40 to $3.48/1M tokens (~21% cut). Cached input unchanged at $0.20.

Before and after

Jun 15, 2026 vs Jun 22, 2026
Pulse.
100%
Change 1 of 2 · price decreased

GPU Clusters on-demand rates cut across all hardware: H100 $5.49→$4.79, H200 $6.79→$5.99, B200 $9.95→$8.19/hour. Reserved rates also reduced 12–25% across H100, H200, and B200 tiers.

Sign in to view the full comparison

See the before/after screenshots and every marked-up change.