Release v5.12.0
- transformers: 12 events in the last 90 days
- transformers: 12th Release in the last 90 days
- Previous: 1 days earlier · Release v5.11.0
What happened
Release v5.12.0 New Model additions MiniMax-M3-VL MiniMax-M3-VL is the vision-language member of the MiniMax-M3 family that pairs a CLIP-style vision tower with 3D rotary position embeddings with the MiniMax-M3 text backbone. It uses a mixed dense/sparse Mixture-of-Experts decoder with SwiGLU-OAI gated experts and a lightning indexer for block-sparse attention. The model processes images through a Conv3d patch embed…
Summary assembled by rule from the sources below