🚀 New Features
- Integrations tab — launch the OpenCode and GitHub Copilot CLI agents in one click, backed by the local API server
- Artifacts — a live preview panel for HTML/CSS/JS code with copy, download and print buttons
- New MLX-VLM support — faster, leaner models:
- EAGLE-3 speculative decoding for Gemma 4
- MTP for Qwen 3.5/3.6 and DeepSeek V4
- TurboQuant KV cache with RHT-correct fast paths for a smaller memory footprint
🔧 Improvements & Fixes
- Quit now properly terminates
llamacpp-upstreamchild processes — no more stuckllama-serverinstances (#31) - Redesigned the hub model card for clearer, better-organized model info
- The send button no longer stays disabled on the new-chat screen while a local model is still loading
- Added Hermes Desktop to the "Launch With" table in the docs
- 5+ more stability and UI fixes, including dark theme
🙏 Contributors
Thanks to @evil-doer, @Vect0rM, @Albert-Atomic, @dtorey-d, @yanalialiuk for their contributions to this release!