Product change·Jul 16 → Jul 17, 2026·Day 198, 2026·AI & Machine Learning
Together AI: Added Inkling & LFM2.5-8B-A1B models to Serverless Inference
Two new models added to Serverless Inference catalog — Inkling ($1.20/1M input, $4.05/1M output, $0.17 cached) and LFM2.5-8B-A1B ($0.03/1M input, $0.12/1M output).
What changed
Feature added
Two new models added to the Serverless Inference catalog: Inkling ($1.20/1M input, $4.05/1M output, $0.17 cached) and LFM2.5-8B-A1B ($0.03/1M input, $0.12/1M output)
Before and after
Jul 16, 2026 vs Jul 17, 2026Pulse.
100%
What changed
Two new models added to the Serverless Inference catalog: Inkling ($1.20/1M input, $4.05/1M output, $0.17 cached) and LFM2.5-8B-A1B ($0.03/1M input, $0.12/1M output)
Sign in to view the full comparison
See the before/after screenshots and every marked-up change.
More from Together AI
Full history →- Week 41, 2026·improvedTogether AI adds Tev1 4B Experimental at $0.04 input with free output
- Week 39, 2026·increaseTogether AI: Qwen3.7-Max inference price increased 25% ($2.00→$2.50)
- Week 38, 2026·increaseTogether AI: Qwen3.7-Max price hike + new DeepSeek V4.1 Flash model
- Week 34, 2026·improvedTogether AI unveils Qwen3.8-2.4T-A95B, Muse Glimmer, DeepSeek V4 Pro pricing on ...
- Week 32, 2026·improvedTogether AI: HGX H100 GPU price increase + Kimi K3 & Inkling Small models added
