Skip to content

PricingSaaS joins Willingness to PayRead the announcement →

Pricing change·May 25 → Jun 1, 2026·Week 23, 2026·AI & Machine Learning

Together AI cuts Qwen3 inference prices 50%, adds caching

Together AIPricingPrice decrease

What changed

Price decrease

Qwen3.7-Max serverless inference pricing decreased: input from $2.50 to $1.25/1M tokens (50% cut), output from $7.50 to $3.75/1M tokens (50% cut); new cached input price of $0.13/1M tokens added.

Before and after

May 25, 2026 vs Jun 1, 2026
Pulse.
100%
What changed

Qwen3.7-Max serverless inference pricing decreased: input from $2.50 to $1.25/1M tokens (50% cut), output from $7.50 to $3.75/1M tokens (50% cut); new cached input price of $0.13/1M tokens added.

Sign in to view the full comparison

See the before/after screenshots and every marked-up change.