New serverless models
UpdateUnverifiedAdded Sep 24, 2026
The following models are now available on serverless : deepseek-ai/DeepSeek-V4.1-Flash : 1,000,000 context length, FP8 quantization, function calling, and structured outputs. Pricing: $0.30 input / $1.20 output / $0.006 cached input (per 1M tokens).
Topics: Pricing
More Together AI platform releases
Every Together AI platform release| Date | Release | Type |
|---|---|---|
| Sep 10 | Preemptible compute for GPU clustersunverified Preview | Preview |
| Sep 10 | What's new:unverified Update | Update |
| Sep 10 | Together CLI v2.33.2unverified Beta | Beta |
| Sep 14 | New serverless modelsunverified Update | Update |
| Sep 8 | New models available for fine-tuningunverified Update | Update |
| Sep 15 | Rollouts for dedicated model inferenceunverified Beta | Beta |
| Sep 15 | Together CLI v2.34.0unverified Beta | Beta |
| Sep 15 | Model deprecationsunverified Deprecation | Deprecation |
Also shipped on Sep 11, 2026
Together AI in September 2026Sources: each vendor's own release notes, changelogs and GitHub releases, read daily to monthly by how often it posts. Logos via logo.dev; trademarks belong to their owners.