v0.32.15
- Ollama: 12 events in the last 90 days
- Ollama: 12th Release in the last 90 days
- Previous: earlier the same day · v0.32.15-rc2
What happened
What's Changed New desktop onboarding flow on first launch Caches resolved model metadata between requests, cutting time-to-first-token by roughly half (TTFT dropped from ~995 ms to ~524 ms in benchmarks) Fixes a bug where chat and generate could wedge after a mid-stream parser error Qwen 3.8 system messages are now normalized so non-leading system messages are handled consistently MLX and llama.cpp dependency updat…
Summary assembled by rule from the sources below