← Back to events
CoolingAIRelease0.32.15

v0.32.15

What happened

What's Changed New desktop onboarding flow on first launch Caches resolved model metadata between requests, cutting time-to-first-token by roughly half (TTFT dropped from ~995 ms to ~524 ms in benchmarks) Fixes a bug where chat and generate could wedge after a mid-stream parser error Qwen 3.8 system messages are now normalized so non-leading system messages are handled consistently MLX and llama.cpp dependency updat…

Summary assembled by rule from the sources below

Why it's spreading

Sources

Release