Skip to content

PricingSaaS joins Willingness to PayRead the announcement →

Product change·Feb 9 → Feb 18, 2026·Week 8, 2026·Marketing & Content

Fireworks overhauls AI model lineup with GLM-5 and MiniMax M2 additions

FireworksProductFeature removedFeature change

What changed

Feature removed

GLM-4.6 removed from serverless inference pricing [was $0.55/1M input, $2.19/1M output tokens]

Feature removed

Qwen3 Coder 480B removed from serverless inference pricing [was $0.45/1M input, $1.80/1M output tokens]

Feature removed

MiniMax M2 family cached input pricing added at $0.03/1M tokens [previously no cached pricing; model listing consolidated from M2/M2.1 to M2 family]

Feature removed

Qwen3 235B Family removed from serverless inference pricing [was $0.22/1M input, $0.88/1M output tokens]

Feature removed

GLM-5 added to serverless inference pricing at $1.00/1M input, $0.20 cached input, $3.20/1M output tokens

Feature change

DeepSeek R1 0528 removed from serverless inference pricing [was $1.35/1M input, $5.40/1M output tokens]

Before and after

Feb 9, 2026 vs Feb 18, 2026
Pulse.
100%
Change 1 of 6 · feature removed

GLM-4.6 removed from serverless inference pricing [was $0.55/1M input, $2.19/1M output tokens]

Sign in to view the full comparison

See the before/after screenshots and every marked-up change.