v0.34.1
Summary
What's Changed app: fix ChatGPT model selector spacing mlxrunner: Evict prefix cache snapshots from the active conversation mlxrunner: check system free memory and wait for evicted runners before loading the next MLX model llm: raise token repeat limit to 100 and return error instead of incomplete result mlx: scope array lifetimes instead of pinning and sweeping llm: keep gemma3n projector off the CPU app: refresh Apps layout and command copy feedback MLX and llama.cpp updates Full Changelog : v0.34.0...v0.34.1-rc1
Lotu Radar provides attributed news summaries and links to the original publisher. Full reporting and copyright remain with the source.