Skip to content

Transformers Release: v5.16.0

UpdateVerifiedAdded Sep 22, 2026

Release v5.16.0 New Model additions Qwen4-Exp Qwen4-Exp builds on Qwen3.5's hybrid text and multimodal architecture with three key components: GatedResidual (GR), Qwen Sparse Attention (QSA), and Per-Layer Embedding (PLE). GR is a Qwen-developed residual architecture that combines Hyper-Connection with GatedNorm.

Topics: Vector search, SQL, Streaming, Governance

Hugging Face's release notes

Summaries of vendors' own notes. Product names and logos belong to their owners; logos via logo.dev.

More Transformers releases

Every Transformers release

Also shipped on Aug 26, 2026

Hugging Face in August 2026

Weekly: the week's data and AI releases, Tuesday mornings.