Skip to content

PricingSaaS joins Willingness to PayRead the announcement →

Pricing change·Aug 28 → Aug 29, 2026·Day 241, 2026·AI & Machine Learning

Together AI cuts Qwen3.8-2.4T-A95B inference prices and adds GLM-5.3

Qwen3.8-2.4T-A95B serverless inference pricing decreased: input $2.50 to $2.00, output $6.25 to $6.00, cached input $0.50 to $0.25 per 1M tokens.; New GLM-5.3 model added to serverless inference pricing at $1.40 input / $4.40 output / $0.26 cached input per 1M tokens.

Together AIPricingPrice decreaseFeature added

What changed

Price decrease

Qwen3.8-2.4T-A95B serverless inference pricing decreased: input $2.50 to $2.00, output $6.25 to $6.00, cached input $0.50 to $0.25 per 1M tokens.

Feature added

New GLM-5.3 model added to serverless inference pricing at $1.40 input / $4.40 output / $0.26 cached input per 1M tokens.

Before and after

Aug 28, 2026 vs Aug 29, 2026
Pulse.
100%
Change 1 of 2 · price decreased

Qwen3.8-2.4T-A95B serverless inference pricing decreased: input $2.50 to $2.00, output $6.25 to $6.00, cached input $0.50 to $0.25 per 1M tokens.

Sign in to view the full comparison

See the before/after screenshots and every marked-up change.