GPT-OSS 20B fine-tuning $0.40->$2.50/1M tok; Nemotron 3.5 Lightning $0.80->$0.40; Qwen3.8 27B $4.00->$2.50; added NVIDIA B200 ($9.65/hr) & AMD MI355x ($7.40/hr) self-serve GPUs. 4 new fine-tuning models added (DeepSeek V4 Flash 0731 QAT, Llama 3.3 70B Instruct, GLM 5.2 FP8 QAT, GLM 5.2 FP8). fine-tuning context windows reduced for DeepSeek V4 Flash 0731 (1M->320K), Gemma 4 31B it (262K->32.7K), GPT-OSS 120B (131K->32.7K), Qwen3.6 35B A3B (256K->131K); prices unchanged.
GPT-OSS 20B fine-tuning price $0.40->$2.50/1M tok; Nemotron 3.5 Lightning $0.80->$0.40; Qwen3.8 27B $4.00->$2.50. Added NVIDIA B200 ($9.65/hr) & AMD MI355x ($7.40/hr) self-serve GPUs.
Fine-tuning model context windows shrunk: DeepSeek V4 Flash 0731 1M->320K, Gemma 4 31B it 262K->32.7K, GPT-OSS 120B 131K->32.7K, Qwen3.6 35B A3B 256K->131K tokens (prices unchanged).
Added 4 new Serverless Fine-Tuning models: DeepSeek V4 Flash 0731 QAT ($5.00/1M tok), Llama 3.3 70B Instruct ($6.00/1M tok), GLM 5.2 FP8 QAT ($10.00/1M tok), GLM 5.2 FP8 ($10.00/1M tok).
GPT-OSS 20B fine-tuning price $0.40->$2.50/1M tok; Nemotron 3.5 Lightning $0.80->$0.40; Qwen3.8 27B $4.00->$2.50. Added NVIDIA B200 ($9.65/hr) & AMD MI355x ($7.40/hr) self-serve GPUs.
See the before/after screenshots and every marked-up change.