Transformers 5.19 adds EmbeddingGemma 2 and expands MoE expert parallelism
Hugging Face’s October 6 release adds Google’s multimodal EmbeddingGemma 2 architecture, makes token-dispatch expert parallelism the default for Qwen3 MoE and Mellum, and repairs quantized-cache behavior.