Transformers Release v5.19.0
Update4 days agoUnverifiedAdded Oct 10, 2026
Release v5.19.0 New Model additions EmbeddingGemma2 EmbeddingGemma 2 is a multimodal embedding model from Google built on the Gemma 4 architecture. It encodes text, images, audio, and video, individually or combined in one input, into a shared 768-dimensional vector space for cross-modal retrieval, semantic similarity, clustering, and classification.
Topics: Vector search, SQL, Governance, Notebooks
More Transformers releases
Every Transformers releaseTransformers Release 5.18.0
New Model additions Nemotron 3 Diarization Nemotron 3 Diarization is an open-weight streaming speaker diarization model designed to determine "who spoke when" in real-world audio. It supports both streaming and offline inference, handles up to eight speakers, and orders speaker outputs by each speaker's first arrival in the input audio.
Transformers Release 5.17.0
Release v5.17.0 New Model additions HYV4 Hy4-Preview is a 780B-parameter mixture-of-experts language model that activates 49B parameters per token. Each MoE layer holds 256 routed experts plus one always-active shared expert and routes every token to 8 of them. The context window is 1M tokens.
Transformers Release v5.16.1
Release v5.16.1 This is a special release as we include GLM! (and a few small fixes) GLM-5.3-Flash GLM-5.3-Flash, the first natively multimodal model in the GLM-5 series.
Transformers Release: v5.16.0
Release v5.16.0 New Model additions Qwen4-Exp Qwen4-Exp builds on Qwen3.5's hybrid text and multimodal architecture with three key components: GatedResidual (GR), Qwen Sparse Attention (QSA), and Per-Layer Embedding (PLE). GR is a Qwen-developed residual architecture that combines Hyper-Connection with GatedNorm.
Transformers Patch release: v5.15.1
Patch release v5.15.1 This patch most notably solves a few issues with DFlash and MTP candidate generators, as well as an issue where images could sometimes not be processed on accelerator if using Lanczos filter.
Transformers Release: v5.15.0
Release v5.15.0 New Model additions Meta Muse Glimmer Muse Glimmer, released today, is Meta’s new multimodal model, especially designed for agentic use cases. Distilled from Muse to 30B parameters, and released under the Apache 2.0 license, it can be deployed to local setups for privacy-aware applications such as coding, document analysis, personal...
Transformers Patch release: v5.14.1
Patch release v5.14.1 This patch solves a few issues which appeared when integrating Inkling model, most notably an issue affecting models using EncoderDecoderCache during assisted generation. It also fixes an issue that could appear during prefill with StaticCache and sdpa without padding for Inkling which uses a position_bias.
Transformers Release v5.14.0
Release v5.14.0 New Model additions Inkling (fresh from Thinking Machines): 975B total, 41B active Add Inkling model #47347 by @molbap @Cyrilvallez @eustlb and @zucchini-nlp Inkling is a general-purpose multimodal model that accepts text, image and audio inputs and generates text outputs.
Also shipped on Oct 6, 2026
Hugging Face in October 2026WAF Release - 2026-10-06
This release introduces a new detection to mitigate a heap-based buffer overflow vulnerability in F5 BIG-IP, and enhances existing command injection protections by incorporating tested beta logic into the baseline rule. Key Findings CVE-2026-94127: A heap-based buffer overflow vulnerability in F5 BIG-IP.
Metric view window measures are generally available
Window measures in metric views are now generally available. Window measures define rolling window, cumulative, or semiadditive aggregations, such as moving averages, period-over-period changes, and running totals. See Window measures .
Conversational analytics in BigQuery now supports the AI.CAUSAL_EFFECT function to quantify the impact of...
Conversational analytics in BigQuery now supports the AI.CAUSAL_EFFECT function to quantify the impact of specific interventions on time series data. This feature is in Preview .
Gemini Enterprise: Control access to Antigravity features
Administrators can control whether users in their organization have access to the following Google Antigravity features: Boost : Uses a tiered multi-agent hierarchy to solve complex algorithmic challenges and deep debugging tasks with the /boost command. This feature is generally available (GA) in Antigravity.
Gemini Nano Banana 2.1
Gemini Nano Banana 2.1 ( gemini-nano-banana-2.1 ) is available in General Availability (GA) . Gemini Nano Banana 2.1 is optimized for high-speed multimodal image generation and editing, offering improved visual quality, prompt adherence, and text rendering across 1K , 2K , and 4K output resolutions.
Genie Code CLI is in Beta
Genie Code CLI is a coding agent that runs in your terminal, tuned for data and AI work on Databricks. It works with your local files and uses the Databricks CLI to discover data, answer questions, and build and deploy assets such as pipelines, models, and apps. Model access is provided and governed through Unity Gateway. See Genie Code CLI .