What's Changed
- app: fix ChatGPT model selector spacing
- mlxrunner: Evict prefix cache snapshots from the active conversation
- mlxrunner: check system free memory and wait for evicted runners before loading the next MLX model
- llm: raise token repeat limit to 100 and return error instead of incomplete result
- mlx: scope array lifetimes instead of pinning and sweeping
- llm: keep gemma3n projector off the CPU
- app: refresh Apps layout and command copy feedback
- MLX and llama.cpp updates
Full Changelog: v0.34.0...v0.34.1-rc1
SOCIAL SHARE CARD GENERATOR