Skip to content

PricingSaaS joins Willingness to PayRead the announcement →

Pricing change·Sep 10 → Sep 11, 2026·Day 254, 2026·AI & Machine Learning

Together AI hikes Qwen3.7-Max pricing 60% and cuts LFM2.5-8B model

LFM2.5-8B-A1B model was removed from the Serverless Inference chat model catalog; it was priced at $0.03 input / $0.12 output per 1M tokens.; Qwen3.7-Max serverless inference pricing increased: input $1.25 -> $2.00 per 1M tokens (+60%), cached input $0.13 -> $0.25 (+92%), output $3.75 -> $6.00 per 1M tokens (+60%).

Together AIPricingFeature removedPrice increase

What changed

Feature removed

LFM2.5-8B-A1B model was removed from the Serverless Inference chat model catalog; it was priced at $0.03 input / $0.12 output per 1M tokens.

Price increase

Qwen3.7-Max serverless inference pricing increased: input $1.25 -> $2.00 per 1M tokens (+60%), cached input $0.13 -> $0.25 (+92%), output $3.75 -> $6.00 per 1M tokens (+60%).

Before and after

Sep 10, 2026 vs Sep 11, 2026
Pulse.
100%
Change 1 of 2 · feature removed

LFM2.5-8B-A1B model was removed from the Serverless Inference chat model catalog; it was priced at $0.03 input / $0.12 output per 1M tokens.

Sign in to view the full comparison

See the before/after screenshots and every marked-up change.