Unsloth Desktop is here! The first desktop app to run and train AI models locally. Research, export and deploy from the same open-source app on Windows, macOS and Linux.
🦥 Download Unsloth Desktop for Linux, Windows, MacOS
Here's what you can do with Unsloth Desktop:
- Get up to 50% more accurate tool calling with self-healing calls and sandboxed code execution.
- Run Muse Glimmer 30B, Kimi K3, Qwen3.8, DeepSeek-V4 Flash 0731, Gemma 4, and more.
- Generate videos with MiniMax-H3, and create images and videos with other diffusion models at up to 2× faster inference on supported workflows.
- Use unlimited private web search, Deep Research, RAG and MCP.
- Export models to NVFP4, GGUF and other formats.
- Access Unsloth remotely through Cloudflare HTTPS.
- Run on CPU or multiple GPUs across NVIDIA, AMD, Intel and Mac.
- Train models without code, using less time and VRAM.
- Use local models through Unsloth's OpenAI-compatible API, or connect OpenAI and Anthropic models.
Tools, private research + APIs
Self-healing tool calling repairs malformed calls instead of dropping them. Models can run Python and Bash inside sandboxed environments, so they can test code, create files and verify their work.
Use unlimited private web search, let Deep Research plan and produce cited reports, or bring your own files into RAG. You can also connect MCP tools for workflows that need external apps, data or actions.
Local models can be served through Unsloth's OpenAI-compatible API for agents and other clients. Inside Desktop, you can also connect OpenAI and Anthropic as cloud model providers.
Muse Glimmer 30B + latest models
Run Muse Glimmer 30B locally for chat, agents, tools and APIs, alongside Kimi K3, Qwen3.8, DeepSeek-V4 Flash 0731 and Gemma 4. Download and manage them in one place through Unsloth Desktop.
MiniMax-H3 + image and video diffusion
Run MiniMax-H3 locally for video generation. Create images and videos locally, edit existing images and train supported diffusion models. Use LoRAs, reference images and ControlNet where available, with up to 2× faster inference on supported workflows.
No-code training, export + remote deployment
Pick a model and dataset, adjust the settings and start training. You can train supported LLMs, diffusion models, TTS models and embedding models without writing code. On supported LLM workloads, training is up to 2× faster and uses up to 70% less VRAM.
Export your trained models to NVFP4, GGUF and other supported formats. You can also securely deploy and access models remotely: turn on Remote access to publish Unsloth through a Cloudflare HTTPS link, then use the app and its local APIs from another device.
Hardware + platform support
Unsloth Desktop runs on Windows, macOS and Linux. Hardware support spans CPU and multi-GPU systems, NVIDIA and AMD GPUs, Intel hardware, and Mac.
CPU support includes Chat and Data Recipes. Training and inference options vary by model and backend.
Download Unsloth Desktop
Unsloth Desktop is free and open source. Download it for:
- Windows
- macOS
- Linux