Published event
ArtificialIntelligence ModelRelease 1 source(s)

Release v5.10.1

Updated September 26, 2026 · 2:47 PM · source date June 3, 2026

Summary

Release v5.10.1 huggingface / transformers Public Notifications You must be signed in to change notification settings Fork 34.7k Star 167k Release v5.10.1 ArthurZucker released this 03 Jun 15:37 · 1107 commits to main since this release v5.10.1 90c3ae5 Release v5.10.1 v5.10.0 was yanked as we publish on a corrupted branch. Sorry everyone, this happens when we rush a release!!!

Why it matters

This ModelRelease is relevant to the technology intelligence record because it involves GitHub, Google, DeepSeek, AMD. The source article should remain the factual reference for follow-up coverage.

Key facts
  • huggingface / transformers Public Notifications You must be signed in to change notification settings Fork 34.7k Star 167k Release v5.10.1 ArthurZucker released this 03 Jun 15:37 · 1107 commits to main since this release v5.10.1 90c3ae5 Release v5.10.1 v5.10.0 was yanked as we publish on a corrupted branch.
  • Sorry everyone, this happens when we rush a release!!!
  • New Model additions Gemma4 unified+ Gemma4 MTP Gemma 4 12B Unified is an encoder-free multimodal model with pretrained and instruction-tuned variants.
  • Unlike standard Gemma 4 , which uses dedicated encoder towers, Gemma 4 12B Unified projects raw inputs directly into the language model's embedding space through lightweight linear pipelines.
  • This results in a simpler architecture while maintaining strong multimodal performance.
  • Key differences from standard Gemma 4: No Vision Tower : Raw pixel patches are projected directly into LM space via a Dense + LayerNorm pipeline with factorized 2D positional embeddings, replacing the vision encoder.
Entities in this story
Related events