Skip to content

Rollouts for dedicated model inference

BetaUnverifiedAdded Sep 24, 2026

Rollouts shift live traffic from one deployment to another under the same endpoint, without changing the endpoint URL. Pick a canary, blue-green, or rolling strategy to determine how traffic moves, and optionally gate a canary rollout on live metrics so it pauses automatically if the new deployment regresses.

Topics: Observability, Developer tools

Together AI's release notes

More Together AI platform releases

Every Together AI platform release
DateRelease
Sep 15Together CLI v2.34.0unverified
Beta
Sep 15Model deprecationsunverified
Deprecation
Sep 16Automatic idle shutdown for dedicated deploymentsunverified
Update
Sep 14New serverless modelsunverified
Update
Sep 11New serverless modelsunverified
Update
Sep 10Preemptible compute for GPU clustersunverified
Preview
Sep 10What's new:unverified
Update
Sep 10Together CLI v2.33.2unverified
Beta

Also shipped on Sep 15, 2026

Together AI in September 2026

Sources: each vendor's own release notes, changelogs and GitHub releases, read daily to monthly by how often it posts. Logos via logo.dev; trademarks belong to their owners.

New releases by email

Tuesday mornings: the week's data, AI and developer-tools releases, only in weeks when something shipped.

Double opt-in. Unsubscribe any time.