{"id":6873,"date":"2026-09-29T02:21:02","date_gmt":"2026-09-28T17:21:02","guid":{"rendered":"https:\/\/localmodelwatch.tsuchitsuchi.com\/2026\/09\/29\/unsloth-desktop-v01900-beta-release\/"},"modified":"2026-09-29T02:21:02","modified_gmt":"2026-09-28T17:21:02","slug":"unsloth-desktop-v01900-beta-release","status":"publish","type":"post","link":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/2026\/09\/29\/unsloth-desktop-v01900-beta-release\/","title":{"rendered":"unsloth Desktop v0.1.900-beta Released with Laya and Speedups"},"content":{"rendered":"<p><!-- lmw:facts --><\/p>\n<h2>At a Glance<\/h2>\n<div class=\"lmw-table-scroll\" tabindex=\"0\" style=\"overflow-x:auto;-webkit-overflow-scrolling:touch;max-width:100%;\">\n<table style=\"width:max-content;min-width:100%;border-collapse:collapse;\">\n<thead>\n<tr>\n<th>Item<\/th>\n<th>Value<\/th>\n<\/tr>\n<\/thead>\n<tbody>\n<tr>\n<td>Repository<\/td>\n<td><a href=\"https:\/\/github.com\/unslothai\/unsloth\">unslothai\/unsloth<\/a><\/td>\n<\/tr>\n<tr>\n<td>Version<\/td>\n<td><a href=\"https:\/\/github.com\/unslothai\/unsloth\/releases\/tag\/v0.1.900-beta\">v0.1.900-beta<\/a><\/td>\n<\/tr>\n<tr>\n<td>Published<\/td>\n<td>2026-09-28<\/td>\n<\/tr>\n<tr>\n<td>Source type<\/td>\n<td>Primary source (the publisher itself)<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<\/div>\n<p><em>Values determined by this site&#8217;s code at collection time. Dates are JST.<\/em><\/p>\n<p><!-- \/lmw:facts --><\/p>\n<h2>Overview<\/h2>\n<p>unsloth (unslothai\/unsloth), known as a fine-tuning and inference tool for local LLMs, has updated its Desktop app to v0.1.900-beta. The highlight of this release is the ability to locally run and serve <strong>Laya<\/strong>, one of the &#8220;Decision Models&#8221; based on the open-source Jev. In addition, it brings a &#8220;Library&#8221; feature to manage documents and media together, a PDF\/Office file viewing feature, a Skills Editor to directly edit skills within the app, and massive speedups for image and video generation (approximately 4.5x faster for LTX-2.3 clips and 1.7 to 6.3x faster for VAE decoding). For those already using unsloth Desktop on a daily basis, the image and video generation speedups alone make this update worth taking.<\/p>\n<h2>Breaking Changes and Deprecations<\/h2>\n<p>This release notes no breaking changes or deprecations that would replace existing options, APIs, or settings.<\/p>\n<h2>Key Changes<\/h2>\n<p><strong>Local Execution of Decision Models via Laya<\/strong><\/p>\n<p>You can now run Laya locally, which is one of the &#8220;Decision Models&#8221; that answers questions with probabilities for yes\/no, multiple-choice, and scoring tasks. To use it, enable the Decision API via Settings &gt; API and select the model to use along with whether to run it on the CPU or GPU. If using the TypeSafe SDK, it can be accessed via the Jev-compatible <code>\/v1\/systemone<\/code> endpoint. When GPU is selected, it is reported to run natively using MLX on Apple Silicon. Details can be found in the related PR (<a href=\"https:\/\/github.com\/unslothai\/unsloth\/pull\/11603\">#11603<\/a>). For those who have been replacing classification and judgment tasks with existing generative models, this is worth considering as a replacement in terms of accuracy and speed.<\/p>\n<p><strong>Addition of Skills Editor<\/strong><\/p>\n<p>A Skills Editor has been added, allowing you to directly create, edit, and delete Skills within the Desktop app. Workflows that previously required editing files externally and loading them can now be completed entirely within the app.<\/p>\n<p><strong>Support for Downloading Models from ModelScope<\/strong><\/p>\n<p>For users in environments unable to access Hugging Face, downloading models via ModelScope is now supported. This increases options for acquiring models for users who could not use HF due to regional restrictions or other reasons.<\/p>\n<p><strong>Library + Document Viewer<\/strong><\/p>\n<p>A Library \/ Document Viewer tab has been added to view PDF, Word, Excel, and PowerPoint files directly inside unsloth. Chats, images, videos, and more can also be centrally managed here, and attached file card displays have been improved. Users can now create their own sidebar sections and reorder them by dragging. This directly impacts users who employ workflows involving reading and interacting with documents.<\/p>\n<p><strong>Improvements for Apple Silicon<\/strong><\/p>\n<p>Several improvements for Apple Silicon environments are included, such as batched serving, structured outputs, and TurboQuant KV cache. Those serving models on a Mac should check for changes in memory efficiency and response speed.<\/p>\n<p><strong>Model Allocation Settings per GPU<\/strong><\/p>\n<p>In multi-GPU environments, you can now set which parts of the model are allocated to each GPU. This is relevant for those distributing models across multi-GPU configurations.<\/p>\n<p><strong>Speedup of Image and Video Generation<\/strong><\/p>\n<p>LTX-2.3 clip generation is reported to be approximately 4.5x faster, driven by distilled sampling, compilation-related fixes, and host-provided FP8 weights. Image and video VAE decoding is also 1.7 to 6.3x faster, which is said to correspondingly improve overall generation speed in workflows. Furthermore, initial rendering for MiniMax-H3 is reported to be up to 1 minute faster. Users generating images and videos with LTX-2.3 and MiniMax-H3 should be able to notice the improvement in perceived speed.<\/p>\n<p><strong>Addition of New Themes<\/strong><\/p>\n<p>Several new food-themed options have been added.<\/p>\n<h2>Specifications and Supported Hardware<\/h2>\n<p>Among Decision Models, <strong>Laya<\/strong>, which is based on the open-source Jev, can now be executed locally. When GPU is selected, it is reported to operate natively using MLX on Apple Silicon. Execution on CPU is also supported, and the execution method can be selected from Settings &gt; API.<\/p>\n<p>For image and video generation, <strong>LTX-2.3<\/strong> clip generation is accelerated in combination with host-provided FP8 weights, implying that an environment capable of handling FP8 format models is assumed (detailed requirements are not explicitly stated in the release notes). The VAE decoding speedup applies generally to image and video generation, and reduced initial rendering times have also been reported for <strong>MiniMax-H3<\/strong>.<\/p>\n<p>As a model distribution channel, downloading via <strong>ModelScope<\/strong> is now supported in addition to Hugging Face. Users who cannot access HF due to regional or environmental constraints can acquire models through this route.<\/p>\n<p>For multi-GPU environments, a feature has been added to configure which parts of the model are assigned to each GPU, catering to setups where models are distributed across multiple GPUs.<\/p>\n<h2>How to Get It<\/h2>\n<p>Specific installation commands (such as <code>pip install<\/code> or <code>git pull<\/code>) are not mentioned in this material. If you are using the unsloth Desktop app, please get the latest version (v0.1.900-beta) through the in-app update function or from the releases page.<\/p>\n<ul>\n<li><a href=\"https:\/\/github.com\/unslothai\/unsloth\/releases\/tag\/v0.1.900-beta\"><a href=\"https:\/\/github.com\/unslothai\/unsloth\/releases\/tag\/v0.1.900-beta\">https:\/\/github.com\/unslothai\/unsloth\/releases\/tag\/v0.1.900-beta<\/a><\/a><\/li>\n<\/ul>\n<p>Note that newly added features such as the Decision API and Skills Editor need to be enabled and configured from the Settings screen after updating.<\/p>\n<p><!-- lmw:releases --><\/p>\n<h2>Releases Since Our Last Article<\/h2>\n<p><em>Compiled by Local Model Watch from the project&#8217;s GitHub releases: the versions between this release and the last one we covered, which did not get separate articles.<\/em> <em>Full history: <a href=\"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/engine-unsloth-en\/\">release tracker<\/a>.<\/em><\/p>\n<div class=\"lmw-table-scroll\" tabindex=\"0\" style=\"overflow-x:auto;-webkit-overflow-scrolling:touch;max-width:100%;\">\n<table style=\"width:max-content;min-width:100%;border-collapse:collapse;\">\n<thead>\n<tr>\n<th>Version<\/th>\n<th>Released<\/th>\n<th>Release notes<\/th>\n<\/tr>\n<\/thead>\n<tbody>\n<tr>\n<td>prebuilt-wheels-cu13 (Flash-Attention2, Causal-Conv1D, Mamba_SSM Binaries)<\/td>\n<td>2026-09-27<\/td>\n<td><a href=\"https:\/\/github.com\/unslothai\/unsloth\/releases\/tag\/prebuilt-wheels-cu13\">GitHub<\/a><\/td>\n<\/tr>\n<tr>\n<td>v0.1.815-beta (Qwen-Image-2.1 + Skills)<\/td>\n<td>2026-09-23<\/td>\n<td><a href=\"https:\/\/github.com\/unslothai\/unsloth\/releases\/tag\/v0.1.815-beta\">GitHub<\/a><\/td>\n<\/tr>\n<tr>\n<td>v0.1.814-beta (Qwen-Image-2.1 + Skills)<\/td>\n<td>2026-09-23<\/td>\n<td><a href=\"https:\/\/github.com\/unslothai\/unsloth\/releases\/tag\/v0.1.814-beta\">GitHub<\/a><\/td>\n<\/tr>\n<tr>\n<td>v0.1.813-beta (Qwen-Image-2.1 + Skills)<\/td>\n<td>2026-09-23<\/td>\n<td><a href=\"https:\/\/github.com\/unslothai\/unsloth\/releases\/tag\/v0.1.813-beta\">GitHub<\/a><\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<\/div>\n<p><!-- \/lmw:releases --><\/p>\n<p><!-- lmw:related --><\/p>\n<h2>Related Articles<\/h2>\n<ul>\n<li><a href=\"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/2026\/09\/23\/unsloth-qwen-image-2-1-agent-skills-update\/\">Unsloth Update: Qwen-Image-2.1 Support and Agent Skills Added<\/a><\/li>\n<li><a href=\"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/2026\/09\/19\/unsloth-v0-1-811-beta-released\/\">Unsloth v0.1.811-beta Released with AMD &amp; NVIDIA Docker Images<\/a><\/li>\n<li><a href=\"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/2026\/09\/18\/unsloth-v0-1-810-beta-released\/\">Unsloth v0.1.810-beta Released with Multi-User and AMD Support<\/a><\/li>\n<li><a href=\"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/2026\/09\/16\/unsloth-windows-arm64-release\/\">Unsloth Releases Windows ARM64 Binary Version<\/a><\/li>\n<\/ul>\n<p><!-- \/lmw:related --><\/p>\n<p><!-- lmw:next-steps --><\/p>\n<h2>What to Read Next<\/h2>\n<ul>\n<li><strong>Follow this tool<\/strong> \u2192 <a href=\"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/engine-unsloth-en\/\">Unsloth overview and release history (71 releases tracked)<\/a><\/li>\n<li><strong>Other quantization, model formats and fine-tuning<\/strong> \u2192 <a href=\"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/engine-ggml-en\/\">ggml<\/a> \/ <a href=\"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/engine-exllamav3-en\/\">ExLlamaV3<\/a><\/li>\n<li><strong>Engines mentioned in this article<\/strong> \u2192 <a href=\"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/engine-localai-en\/\">LocalAI<\/a><\/li>\n<\/ul>\n<p><!-- \/lmw:next-steps --><\/p>\n<h2>Sources<\/h2>\n<ul>\n<li><a href=\"https:\/\/github.com\/unslothai\/unsloth\/releases\/tag\/v0.1.900-beta\">https:\/\/github.com\/unslothai\/unsloth\/releases\/tag\/v0.1.900-beta<\/a><\/li>\n<li><a href=\"https:\/\/github.com\/unslothai\/unsloth\/pull\/11603\">https:\/\/github.com\/unslothai\/unsloth\/pull\/11603<\/a><\/li>\n<\/ul>\n","protected":false},"excerpt":{"rendered":"<p>unsloth Desktop v0.1.900-beta adds local Laya execution, a Skills Editor, Library and Document Viewer, and major image\/video generation speedups.<\/p>\n","protected":false},"author":1,"featured_media":6872,"comment_status":"closed","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[105],"tags":[2330,1507,2604,509,1547],"class_list":["post-6873","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-engines-and-tools","tag-laya-en","tag-localai-en","tag-ltx-2-3-en","tag-unsloth-en","tag-verified"],"lang":"en","translations":{"en":6873,"ja":6871},"_links":{"self":[{"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/posts\/6873","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/comments?post=6873"}],"version-history":[{"count":0,"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/posts\/6873\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/media\/6872"}],"wp:attachment":[{"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/media?parent=6873"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/categories?post=6873"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/tags?post=6873"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}