Together AI hikes Qwen3.7-Max pricing 60% and cuts LFM2.5-8B model
LFM2.5-8B-A1B model was removed from the Serverless Inference chat model catalog; it was priced at $0.03 input / $0.12 output per 1M tokens.; Qwen3.7-Max serverless inference pricing increased: input $1.25 -> $2.00 per 1M tokens (+60%), cached input $0.13 -> $0.25 (+92%), output $3.75 -> $6.00 per 1M tokens (+60%).
What changed
LFM2.5-8B-A1B model was removed from the Serverless Inference chat model catalog; it was priced at $0.03 input / $0.12 output per 1M tokens.
Qwen3.7-Max serverless inference pricing increased: input $1.25 -> $2.00 per 1M tokens (+60%), cached input $0.13 -> $0.25 (+92%), output $3.75 -> $6.00 per 1M tokens (+60%).
Before and after
Sep 10, 2026 vs Sep 11, 2026LFM2.5-8B-A1B model was removed from the Serverless Inference chat model catalog; it was priced at $0.03 input / $0.12 output per 1M tokens.
Sign in to view the full comparison
See the before/after screenshots and every marked-up change.
More from Together AI
Full history →- Week 41, 2026·improvedTogether AI adds Tev1 4B Experimental at $0.04 input with free output
- Week 39, 2026·increaseTogether AI: Qwen3.7-Max inference price increased 25% ($2.00→$2.50)
- Week 38, 2026·improvedTogether AI: Qwen3.7-Max price hike + new DeepSeek V4.1 Flash model
- Week 34, 2026·improvedTogether AI unveils Qwen3.8-2.4T-A95B, Muse Glimmer, DeepSeek V4 Pro pricing on ...
- Week 32, 2026·increaseTogether AI: HGX H100 GPU price increase + Kimi K3 & Inkling Small models added
