Skip to content

PricingSaaS joins Willingness to PayRead the announcement →

Packaging change·Jul 6 → Jul 13, 2026·Week 29, 2026·AI & Machine Learning

Together AI: New Provisioned Throughput pricing + GPU rate cuts

Added Provisioned Throughput (PTU) reserved-capacity pricing ($0.05/PTU/min, MiniMax M3 & GLM-5.2); Dedicated Inference on-demand GPU rates decreased (H100 $6.49->$5.49/hr, HGX B200 $11.95->$8.99/hr)

Together AIPackagingPackage change

What changed

Package change

Added Provisioned Throughput (PTU) pricing ($0.05/PTU/min, MiniMax M3 & GLM-5.2); Dedicated Inference on-demand GPU rates cut (H100 $6.49->$5.49/hr, HGX B200 $11.95->$8.99/hr)

Before and after

Jul 6, 2026 vs Jul 13, 2026
Pulse.
100%
What changed

Added Provisioned Throughput (PTU) pricing ($0.05/PTU/min, MiniMax M3 & GLM-5.2); Dedicated Inference on-demand GPU rates cut (H100 $6.49->$5.49/hr, HGX B200 $11.95->$8.99/hr)

Sign in to view the full comparison

See the before/after screenshots and every marked-up change.