Together AI: Qwen3.7-Max price hike + new DeepSeek V4.1 Flash model
Qwen3.7-Max increased from $1.25/$0.13(cached) input and $3.75 output to $2.00/$0.25(cached) input and $6.00 output; new DeepSeek V4.1 Flash model added at $0.30/$0.006(cached) input and $1.20 output per 1M tokens
What changed
New model DeepSeek V4.1 Flash added to the serverless model catalog at $0.30/1M input tokens, $1.20/1M output tokens, and $0.006/1M cached tokens
Qwen3.7-Max price increased from $1.25/1M input ($0.13 cached) and $3.75/1M output to $2.00/1M input ($0.25 cached) and $6.00/1M output
Before and after
Sep 7, 2026 vs Sep 14, 2026New model DeepSeek V4.1 Flash added to the serverless model catalog at $0.30/1M input tokens, $1.20/1M output tokens, and $0.006/1M cached tokens
Sign in to view the full comparison
See the before/after screenshots and every marked-up change.
Covered on Pulse
More from Together AI
Full history →- Week 41, 2026·improvedTogether AI adds Tev1 4B Experimental at $0.04 input with free output
- Week 39, 2026·increaseTogether AI: Qwen3.7-Max inference price increased 25% ($2.00→$2.50)
- Week 34, 2026·improvedTogether AI unveils Qwen3.8-2.4T-A95B, Muse Glimmer, DeepSeek V4 Pro pricing on ...
- Week 32, 2026·increaseTogether AI: HGX H100 GPU price increase + Kimi K3 & Inkling Small models added
- Week 31, 2026·packagingTogether AI: LFM2.5-8B-A1B model removed from Serverless Inference
