{"id":7404,"date":"2026-09-30T01:19:30","date_gmt":"2026-09-29T16:19:30","guid":{"rendered":"https:\/\/localmodelwatch.tsuchitsuchi.com\/2026\/09\/30\/longlive-plug-wan2-1-t2v-14b-cfg\/"},"modified":"2026-09-30T02:13:48","modified_gmt":"2026-09-29T17:13:48","slug":"longlive-plug-wan2-1-t2v-14b-cfg","status":"publish","type":"post","link":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/2026\/09\/30\/longlive-plug-wan2-1-t2v-14b-cfg\/","title":{"rendered":"LongLive-Plug-Wan2.1-T2V-14B-cfg Video Generation Model: 80GB+ VRAM"},"content":{"rendered":"<p><!-- lmw:facts --><\/p>\n<h2>At a Glance<\/h2>\n<div class=\"lmw-table-scroll\" tabindex=\"0\" style=\"overflow-x:auto;-webkit-overflow-scrolling:touch;max-width:100%;\">\n<table style=\"width:max-content;min-width:100%;border-collapse:collapse;\">\n<thead>\n<tr>\n<th>Item<\/th>\n<th>Value<\/th>\n<\/tr>\n<\/thead>\n<tbody>\n<tr>\n<td>Repository<\/td>\n<td><a href=\"https:\/\/huggingface.co\/Efficient-Large-Model\/LongLive-Plug-Wan2.1-T2V-14B-cfg\">Efficient-Large-Model\/LongLive-Plug-Wan2.1-T2V-14B-cfg<\/a><\/td>\n<\/tr>\n<tr>\n<td>Publisher guide<\/td>\n<td><a href=\"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/publisher-efficient-large-model-en\/\">Efficient Large Model (NVIDIA and MIT): models and licenses<\/a><\/td>\n<\/tr>\n<tr>\n<td>Published<\/td>\n<td>2026-09-29<\/td>\n<\/tr>\n<tr>\n<td>License<\/td>\n<td>apache-2.0<\/td>\n<\/tr>\n<tr>\n<td>Formats<\/td>\n<td><a href=\"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/format-safetensors-en\/\">safetensors<\/a><\/td>\n<\/tr>\n<tr>\n<td>Source type<\/td>\n<td>Primary source (the publisher itself)<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<\/div>\n<p><em>Values determined by this site&#8217;s code at collection time. Dates are JST.<\/em><\/p>\n<p><!-- \/lmw:facts --><\/p>\n<h2>Overview<\/h2>\n<p>Efficient-Large-Model has released <code>LongLive-Plug-Wan2.1-T2V-14B-cfg<\/code>, a LoRA adapter for <code>Wan2.1-T2V-14B<\/code>. This adapter aims to distill Classifier-Free Guidance (CFG) into conditional-only video generation. This reduces the need to run a separate unconditional inference branch during the inference process. Note that this adapter does not provide a speedup in a few steps on its own; as it functions strictly as an adapter, using it requires the corresponding base model.<\/p>\n<h2>Specifications<\/h2>\n<ul>\n<li>Base model: <code>Wan-AI\/Wan2.1-T2V-14B<\/code><\/li>\n<li>Architecture: Video Diffusion DiT (Flow Matching framework)<\/li>\n<li>Base model specifications (14B):<\/li>\n<\/ul>\n<div class=\"lmw-table-scroll\" tabindex=\"0\" style=\"overflow-x:auto;-webkit-overflow-scrolling:touch;max-width:100%;\">\n<p class=\"lmw-table-hint\" style=\"margin:0 0 4px;font-size:0.85em;opacity:0.7;\">\u2192 Scroll horizontally to see all columns<\/p>\n<table style=\"width:max-content;min-width:100%;border-collapse:collapse;\">\n<thead>\n<tr>\n<th>Model<\/th>\n<th>Dimension<\/th>\n<th>Input Dimension<\/th>\n<th>Output Dimension<\/th>\n<th>Feedforward Dimension<\/th>\n<th>Frequency Dimension<\/th>\n<th>Number of Heads<\/th>\n<th>Number of Layers<\/th>\n<\/tr>\n<\/thead>\n<tbody>\n<tr>\n<td>14B<\/td>\n<td>5120<\/td>\n<td>16<\/td>\n<td>16<\/td>\n<td>13824<\/td>\n<td>256<\/td>\n<td>40<\/td>\n<td>40<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<\/div>\n<ul>\n<li>Output resolution: Supports 480P and 720P<\/li>\n<li>Distribution format: <code>safetensors<\/code> (PEFT\/LoRA)<\/li>\n<li>License: <code>apache-2.0<\/code><\/li>\n<li>Recommended settings: It is recommended to use this in combination with the corresponding <code>few-step LoRA<\/code>. The recommended weight ratio is <code>few-step : CFG = 1 : 0.5<\/code>. However, this refers to the adapter weight ratio, not the CFG scale itself during inference.<\/li>\n<\/ul>\n<h2>Performance and Quality<\/h2>\n<p>The base model targeted by this adapter, <code>Wan2.1-T2V-14B<\/code>, is reported to outperform open-source and commercial state-of-the-art (SOTA) models across multiple benchmarks. According to the model card, evaluations were conducted using 1,035 internal prompts across 14 major dimensions and 26 sub-dimensions, confirming superior performance compared to existing models.<\/p>\n<p>Furthermore, manual evaluations with Prompt Extension applied show that it yields better results than existing open-source and closed-source models. The base model excels in its ability to generate high-quality visuals and notable motion dynamics, and is also characterized as being the only video model capable of generating text in both Chinese and English.<\/p>\n<h2>Strengths and Use Cases<\/h2>\n<p>This adapter excels at distilling Classifier-Free Guidance (CFG) into conditional-only generation during the video generation process of the base model, <code>Wan2.1-T2V-14B<\/code>. This makes it possible to eliminate the overhead of running a separate unconditional branch during inference.<\/p>\n<p>The base model, <code>Wan2.1-T2V-14B<\/code>, excels in <code>text-to-video<\/code> tasks at generating high-quality visuals and dynamic motion. It also has the capability to generate both Chinese and English text within videos.<\/p>\n<p><!-- lmw:hardware --><\/p>\n<h2>Hardware Requirements<\/h2>\n<p><strong>Estimated requirements (calculated by Local Model Watch)<\/strong> \u2014 14.3B parameters (taken from the base model Wan-AI\/Wan2.1-T2V-14B)<\/p>\n<div class=\"lmw-table-scroll\" tabindex=\"0\" style=\"overflow-x:auto;-webkit-overflow-scrolling:touch;max-width:100%;\">\n<table style=\"width:max-content;min-width:100%;border-collapse:collapse;\">\n<thead>\n<tr>\n<th>Your VRAM<\/th>\n<th>Quantization<\/th>\n<th>File size<\/th>\n<th>Est. memory needed<\/th>\n<\/tr>\n<\/thead>\n<tbody>\n<tr>\n<td>80GB class (A100 \/ H100)<\/td>\n<td>F32<\/td>\n<td>53.2GB<\/td>\n<td>63.9GB<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<\/div>\n<p><em>Memory estimates add a 20% runtime overhead (KV cache, etc.) to the actual size of the distributed files. Actual usage varies with context length, batch size and inference engine. These figures are computed by this site from file sizes, not published by the model&#8217;s authors. This release is an adapter (LoRA etc.); the table shows what the base model <a href=\"https:\/\/huggingface.co\/Wan-AI\/Wan2.1-T2V-14B\">Wan-AI\/Wan2.1-T2V-14B<\/a> needs. Compare with other models in our <a href=\"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/vram-guide-en\/\">VRAM quick reference<\/a>. What the quantization names mean: <a href=\"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/glossary-quantization-en\/\">glossary<\/a>.<\/em><\/p>\n<p><!-- \/lmw:hardware --><\/p>\n<p><!-- lmw:runnability --><\/p>\n<h2>Can You Run It Locally?<\/h2>\n<p>The publisher distributes this model as safetensors.<\/p>\n<p><strong>License \u2014 <code>apache-2.0<\/code> (Commercial use allowed):<\/strong> Permits commercial use, modification and redistribution. Redistribution requires including the license and stating changes; includes a patent grant.<\/p>\n<p><em>Compiled by this site&#8217;s code from the published formats and the license field. License summaries are not legal advice \u2014 check the publisher&#8217;s original terms before relying on them.<\/em><\/p>\n<p><!-- \/lmw:runnability --><\/p>\n<p><!-- lmw:files --><\/p>\n<h2>Distributed Files<\/h2>\n<p><em>Weight files published in <a href=\"https:\/\/huggingface.co\/Efficient-Large-Model\/LongLive-Plug-Wan2.1-T2V-14B-cfg\/tree\/main\">Efficient-Large-Model\/LongLive-Plug-Wan2.1-T2V-14B-cfg<\/a>, listed by this site from the Hugging Face API. Sizes are the actual file sizes.<\/em><\/p>\n<div class=\"lmw-table-scroll\" tabindex=\"0\" style=\"overflow-x:auto;-webkit-overflow-scrolling:touch;max-width:100%;\">\n<table style=\"width:max-content;min-width:100%;border-collapse:collapse;\">\n<thead>\n<tr>\n<th>File<\/th>\n<th>Size<\/th>\n<\/tr>\n<\/thead>\n<tbody>\n<tr>\n<td><code>adapter_model.safetensors<\/code><\/td>\n<td>2.45GB<\/td>\n<\/tr>\n<tr>\n<td><code>generator_lora.pt<\/code><\/td>\n<td>2.45GB<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<\/div>\n<p><!-- \/lmw:files --><\/p>\n<h2>How to Get It<\/h2>\n<p>This adapter is available from the following repository:<\/p>\n<ul>\n<li>Repository: <code>Efficient-Large-Model\/LongLive-Plug-Wan2.1-T2V-14B-cfg<\/code><\/li>\n<li>Distribution format: <code>safetensors<\/code><\/li>\n<li>Supported library: <code>peft<\/code><\/li>\n<\/ul>\n<p>Note that using this adapter requires the corresponding base model, <code>Wan-AI\/Wan2.1-T2V-14B<\/code>. Additionally, the publisher has released a <code>few-step LoRA<\/code> intended to be used in combination with it, and using them together is recommended.<\/p>\n<p><!-- lmw:related --><\/p>\n<h2>Related Articles<\/h2>\n<ul>\n<li><a href=\"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/2026\/09\/30\/longlive-plug-wan21-t2v-14b-few-step-2\/\">LongLive-Plug-Wan2.1-T2V-14B-few-step: 80GB+ VRAM, File List<\/a><\/li>\n<li><a href=\"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/2026\/09\/29\/longlive-plug-wan21-ti2v-5b-cfg-lora\/\">LongLive-Plug-Wan2.2-TI2V-5B-cfg Video Generation Model: 24GB+ VRAM<\/a><\/li>\n<li><a href=\"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/2026\/09\/29\/longlive-plug-wan2-2-ti2v-5b-few-step-2\/\">LongLive-Plug-Wan2.2-TI2V-5B-few-step: 24GB+ VRAM, File List<\/a><\/li>\n<li><a href=\"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/2026\/09\/10\/h3-to-ltx-latent-adapter-released\/\">H3-to-LTX-Latent-Adapter: 4GB+ VRAM, File List<\/a><\/li>\n<\/ul>\n<p><!-- \/lmw:related --><\/p>\n<p><!-- lmw:next-steps --><\/p>\n<h2>What to Read Next<\/h2>\n<ul>\n<li><strong>Formats this model is available in<\/strong> \u2192 <a href=\"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/format-safetensors-en\/\">Safetensors format guide and models<\/a><\/li>\n<li><strong>Learn about the publisher<\/strong> \u2192 <a href=\"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/publisher-efficient-large-model-en\/\">Efficient Large Model (NVIDIA and MIT): models, licenses and articles<\/a><\/li>\n<li><strong>Other models for the same task<\/strong> \u2192 <a href=\"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/models-by-task-en\/#task-video\">Other video generation models<\/a><\/li>\n<\/ul>\n<p><!-- \/lmw:next-steps --><\/p>\n<h2>Sources<\/h2>\n<ul>\n<li><a href=\"https:\/\/huggingface.co\/Efficient-Large-Model\/LongLive-Plug-Wan2.1-T2V-14B-cfg\">https:\/\/huggingface.co\/Efficient-Large-Model\/LongLive-Plug-Wan2.1-T2V-14B-cfg<\/a><\/li>\n<\/ul>\n","protected":false},"excerpt":{"rendered":"<p>Learn about LongLive-Plug-Wan2.1-T2V-14B-cfg, a LoRA adapter for Wan2.1-T2V-14B designed to distill Classifier-Free Guidance.<\/p>\n","protected":false},"author":1,"featured_media":7403,"comment_status":"closed","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[203],"tags":[1833,724,111,1547,2675,2677,2679,959],"class_list":["post-7404","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-image-video-and-audio","tag-dit-en","tag-efficient-large-model-en","tag-lora-en","tag-verified","tag-video-diffusion-en","tag-wan2-1-en","tag-wan2-1-t2v-en","tag--en"],"lang":"en","translations":{"en":7404,"ja":7402},"_links":{"self":[{"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/posts\/7404","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/comments?post=7404"}],"version-history":[{"count":1,"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/posts\/7404\/revisions"}],"predecessor-version":[{"id":7497,"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/posts\/7404\/revisions\/7497"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/media\/7403"}],"wp:attachment":[{"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/media?parent=7404"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/categories?post=7404"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/tags?post=7404"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}