← 返回事件
降温中AI发版0.32.15

v0.32.15

发生了什么

What's Changed New desktop onboarding flow on first launch Caches resolved model metadata between requests, cutting time-to-first-token by roughly half (TTFT dropped from ~995 ms to ~524 ms in benchmarks) Fixes a bug where chat and generate could wedge after a mid-stream parser error Qwen 3.8 system messages are now normalized so non-leading system messages are handled consistently MLX and llama.cpp dependency updat…

摘要按规则整理自下方来源原文

为什么在扩散

来源

发布公告