Skip to content

July 2026

UpdateUnverifiedAdded Sep 25, 2026

Pairwise scoring When you compare two experiments in diff mode, you can now record which one produced the better output for each row. Braintrust aggregates these head-to-head judgments into a win rate, so you can run A-vs-B human evaluations without configuring a separate scoring function. See Pairwise scoring for details.

Topics: Governance, Observability

's release notes

More releases

Every release
DateRelease
Jul 13June 2026unverified
Update
Jun 12May 2026unverified
Update
Jun 5March 2026unverified
Update
Jun 5February 2026unverified
Update
Sep 11September 2026unverified
Update
Sep 11August 2026unverified
Update
Apr 29April 2026unverified
Update
Jan 10January 2026unverified
Update

Also shipped on Jul 22, 2026

in July 2026

Sources: each vendor's own release notes, changelogs and GitHub releases, read daily to monthly by how often it posts. Logos via logo.dev; trademarks belong to their owners.

New releases by email

Tuesday mornings: the week's data, AI and developer-tools releases, only in weeks when something shipped.

Double opt-in. Unsubscribe any time.