← 返回事件
持续讨论AI发版0.33.3-rc2

v0.33.3-rc2: gemma4: image and audio input support

图:Ollama Releases

发生了什么

Safetensors gemma4 imports served by the MLX engine now answer image and audio chats. Images run through both vision architectures: the transformer tower (26B, 31B, e-series) and the 12B's encoder-free unified embedder. Audio arrives through the same intake the ollama API already accepts for gemma4 GGUFs — WAV bytes in the images field, OpenAI input_audio parts, and /v1/audio/transcriptions uploads — with the e2b/e4…

摘要按规则整理自下方来源原文

为什么在扩散

来源

发布公告