# unsloth v0.1.702-beta - Product: unsloth (https://whatsnew.fyi/product/unsloth) - Vendor: unslothai - Date: 2026-08-13 - Version: v0.1.702-beta - Original notes: https://github.com/unslothai/unsloth/releases/tag/v0.1.702-beta - Permalink: https://whatsnew.fyi/product/unsloth/releases/v0.1.702-beta - Labels: Pre-release What's New is an index, not a publisher: every entry below links to the vendor's own release notes, which are the authoritative source. Entries are labelled where they are hand-curated sample data, pre-releases, or drawn from a secondary source such as a developer blog. Reuse: the summaries, labels and curation here are © What's New. Quote freely with attribution and a link back; wholesale republication of the corpus is not permitted — terms: https://whatsnew.fyi/terms. The vendors' own release notes remain their publishers'. --- - **added** — Added tool calling and web search for all external providers - **fixed** — Fixed bypass permissions not working for sandboxing - **changed** — Made VRAM usage tunable in UI - **changed** — Improved inference speed by 10% and reduced VRAM usage - **changed** — Improved AMD RDNA3 and RDNA4 support - **changed** — Improved Strix Halo support - **changed** — Improved Mac support - **fixed** — Fixed image diffusion and video generation - **added** — Added ability to login with Codex subscription - **fixed** — Fixed slow Windows downloading due to throttling - **fixed** — Fixed Mac asking to download command line tools - **fixed** — Fixed AMD Strix Halo not being detected [Unsloth Desktop](https://unsloth.ai/) is here! The first desktop app to run and train AI models locally. Research, export and deploy from the same open-source app on Windows, macOS and Linux. ##### v0.1.702-beta Update (August 13th) - Added tool calling / web search & more for all external providers - Fixed bypass permissions not working for sandboxing - UI and UX fixes - VRAM usage is now tunable - 10% faster inference + reduced VRAM usage and other perf fixes - Much better AMD RDNA3,4 + Strix Halo, Mac support - Image diffusion, video generation fixes - Can login with Codex subscription - Many bug fixes unsloth desktop **[🦥 Download Unsloth Desktop for Linux, Windows, MacOS](https://unsloth.ai/download)** Here's what you can do with Unsloth Desktop: - Get **up to 50% more accurate tool calling** with self-healing calls and sandboxed code execution. - Run **Muse Glimmer 30B, Kimi K3, Qwen3.8, DeepSeek-V4 Flash 0731, Gemma 4**, and more. - Generate videos with **MiniMax-H3**, and create images and videos with other diffusion models at up to **2× faster inference** on supported workflows. - Use unlimited private **web search, Deep Research, RAG and MCP**. - Export models to **NVFP4, GGUF** and other formats. - Access Unsloth remotely through **Cloudflare HTTPS**. - Run on **CPU or multiple GPUs** across **NVIDIA, AMD, Intel and Mac**. - Train models without code, using less time and VRAM. - Use local models through Unsloth's **OpenAI-compatible API**, or connect **OpenAI and Anthropic** models. ##### Tools, private research + APIs [Self-healing tool calling](https://unsloth.ai/docs/new/studio/chat#auto-healing-tool-calling) repairs malformed calls instead of dropping them. Models can run Python and Bash inside sandboxed environments, so they can test code, create files and verify their work. Use unlimited private web search, let [**Deep Research**](https://unsloth.ai/docs/new/studio/chat) plan and produce cited reports, or bring your own files into **RAG**. You can also connect MCP tools for workflows that need external apps, data or actions. Local models can be served through Unsloth's [**OpenAI-compatible API**](https://unsloth.ai/docs/basics/api) for agents and other clients. Inside Desktop, you can also connect OpenAI and Anthropic as cloud model providers. ##### Muse Glimmer 30B + latest models Run [**Muse Glimmer 30B**](https://unsloth.ai/docs/models/muse-glimmer) locally for chat, agents, tools and APIs, alongside **Kimi K3, Qwen3.8, DeepSeek-V4 Flash 0731 and Gemma 4**. Download and manage them in one place through Unsloth Desktop. ##### MiniMax-H3 + image and video diffusion Run [**MiniMax-H3**](https://huggingface.co/unsloth/MiniMax-H3-GGUF) locally for video generation. Create images and videos locally, edit existing images and train supported diffusion models. Use LoRAs, reference images and ControlNet where available, with up to **2× faster inference** on supported workflows. ##### No-code training, export + remote deployment Pick a model and dataset, adjust the settings and start training. You can train supported LLMs, diffusion models, TTS models and embedding models without writing code. On supported LLM workloads, training is up to 2× faster and uses up to 70% less VRAM. [Export](https://unsloth.ai/docs/new/unsloth-studio/export) your trained models to **NVFP4, GGUF** and other supported formats. You can also securely deploy and access models remotely: turn on Remote access to publish Unsloth through a Cloudflare HTTPS link, then use the app and its local APIs from another device. ##### Hardware + platform support Unsloth Desktop runs on **Windows, macOS and Linux**. Hardware support spans **CPU and multi-GPU systems**, **NVIDIA and AMD GPUs**, **Intel hardware**, and **Mac**. CPU support includes Chat and _[Truncated at 4000 characters — full notes: https://github.com/unslothai/unsloth/releases/tag/v0.1.702-beta]_