Skip to content

PricingSaaS joins Willingness to PayRead the announcement →

Pricing change·Jul 27 → Aug 3, 2026·Week 32, 2026·AI & Machine Learning

Together AI: HGX H100 GPU price increase + Kimi K3 & Inkling Small models added

NVIDIA HGX H100 GPU cluster reserved pricing increased across all commitment tiers (31-90 days $3.59→$3.69/GPU/hr, 91-180 days $3.29→$3.45/GPU/hr, 181+ days $3.09→$3.19/GPU/hr; on-demand $3.99/hr unchanged) Added Kimi K3 ($3.00/$15.00 per 1M input/output tokens, $0.30 cached) and Inkling Small ($0.50/$1.20 per 1M tokens, $0.10 cached) to the serverless inference model catalog

Together AIPricingFeature addedPrice increase

What changed

Feature added

Added two new serverless inference models to the model catalog: Kimi K3 ($3.00/$15.00 per 1M input/output tokens, $0.30 cached input) and Inkling Small ($0.50/$1.20 per 1M input/output tokens, $0.10 cached input)

Price increase

NVIDIA HGX H100 GPU cluster reserved pricing increased across all commitment tiers: 31-90 days $3.59→$3.69/GPU/hr, 91-180 days $3.29→$3.45/GPU/hr, 181+ days $3.09→$3.19/GPU/hr (on-demand rate $3.99/hr unchanged)

Before and after

Jul 27, 2026 vs Aug 3, 2026
Pulse.
100%
Change 1 of 2 · feature added

Added two new serverless inference models to the model catalog: Kimi K3 ($3.00/$15.00 per 1M input/output tokens, $0.30 cached input) and Inkling Small ($0.50/$1.20 per 1M input/output tokens, $0.10 cached input)

Sign in to view the full comparison

See the before/after screenshots and every marked-up change.