Pricing change·Sep 22 → Sep 23, 2026·Day 266, 2026·AI & Machine Learning
Together AI cuts Qwen3.7-Max and Qwen3.8 Flash inference prices by 40%
Qwen3.8 Flash and Qwen3.7-Max serverless inference pricing decreased per 1M tokens: Qwen3.8 Flash $0.15/$0.47 to $0.09/$0.28; Qwen3.7-Max $2.50/$7.50 to $1.50/$4.50 (cached $0.25 to $0.30).
What changed
Price decrease
Qwen3.8 Flash and Qwen3.7-Max serverless inference pricing decreased per 1M tokens: Qwen3.8 Flash $0.15/$0.47 to $0.09/$0.28; Qwen3.7-Max $2.50/$7.50 to $1.50/$4.50 (cached $0.25 to $0.30).
Before and after
Sep 22, 2026 vs Sep 23, 2026Pulse.
100%
What changed
Qwen3.8 Flash and Qwen3.7-Max serverless inference pricing decreased per 1M tokens: Qwen3.8 Flash $0.15/$0.47 to $0.09/$0.28; Qwen3.7-Max $2.50/$7.50 to $1.50/$4.50 (cached $0.25 to $0.30).
Sign in to view the full comparison
See the before/after screenshots and every marked-up change.
Covered on Pulse
More from Together AI
Full history →- Week 41, 2026·improvedTogether AI adds Tev1 4B Experimental at $0.04 input with free output
- Week 39, 2026·increaseTogether AI: Qwen3.7-Max inference price increased 25% ($2.00→$2.50)
- Week 38, 2026·improvedTogether AI: Qwen3.7-Max price hike + new DeepSeek V4.1 Flash model
- Week 34, 2026·improvedTogether AI unveils Qwen3.8-2.4T-A95B, Muse Glimmer, DeepSeek V4 Pro pricing on ...
- Week 32, 2026·increaseTogether AI: HGX H100 GPU price increase + Kimi K3 & Inkling Small models added
