Skip to content

New serverless models

UpdateUnverifiedAdded Sep 24, 2026

The following models are now available on serverless : deepseek-ai/DeepSeek-V4.1-Flash : 1,000,000 context length, FP8 quantization, function calling, and structured outputs. Pricing: $0.30 input / $1.20 output / $0.006 cached input (per 1M tokens).

Topics: Pricing

Together AI's release notes

More Together AI platform releases

Every Together AI platform release
DateRelease
Sep 10Preemptible compute for GPU clustersunverified
Preview
Sep 10What's new:unverified
Update
Sep 10Together CLI v2.33.2unverified
Beta
Sep 14New serverless modelsunverified
Update
Sep 8New models available for fine-tuningunverified
Update
Sep 15Rollouts for dedicated model inferenceunverified
Beta
Sep 15Together CLI v2.34.0unverified
Beta
Sep 15Model deprecationsunverified
Deprecation

Also shipped on Sep 11, 2026

Together AI in September 2026

Sources: each vendor's own release notes, changelogs and GitHub releases, read daily to monthly by how often it posts. Logos via logo.dev; trademarks belong to their owners.

New releases by email

Tuesday mornings: the week's data, AI and developer-tools releases, only in weeks when something shipped.

Double opt-in. Unsubscribe any time.