{"id":6182,"date":"2026-09-28T12:05:15","date_gmt":"2026-09-28T03:05:15","guid":{"rendered":"https:\/\/localmodelwatch.tsuchitsuchi.com\/model-xingchen-agi-xing4-0-en\/"},"modified":"2026-09-28T12:10:39","modified_gmt":"2026-09-28T03:10:39","slug":"model-xingchen-agi-xing4-0-en","status":"publish","type":"page","link":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/model-xingchen-agi-xing4-0-en\/","title":{"rendered":"Xing4.0 Guide: VRAM Requirements, GGUF Builds"},"content":{"rendered":"<p>Everything Local Model Watch has published about the <strong>Xing4.0<\/strong> family: 1 article(s) covering the base model and its fine-tunes, plus converted builds we tracked after publication. Memory requirements below are computed by this site from file sizes, not quoted from model cards. Part of our <a href=\"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/models-en\/\">model family index<\/a>.<\/p>\n<h2>At a Glance<\/h2>\n<div class=\"lmw-table-scroll\" tabindex=\"0\" style=\"overflow-x:auto;-webkit-overflow-scrolling:touch;max-width:100%;\">\n<table style=\"width:max-content;min-width:100%;border-collapse:collapse;\">\n<thead>\n<tr>\n<th>Item<\/th>\n<th>Value<\/th>\n<\/tr>\n<\/thead>\n<tbody>\n<tr>\n<td>Base model(s)<\/td>\n<td><a href=\"https:\/\/huggingface.co\/XingChen-AGI\/Xing4.0-29B-A4B\">XingChen-AGI\/Xing4.0-29B-A4B<\/a><\/td>\n<\/tr>\n<tr>\n<td>Publisher<\/td>\n<td><a href=\"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/publisher-china-telecom-en\/\">China Telecom (Xing, TeleChat)<\/a><\/td>\n<\/tr>\n<tr>\n<td>Parameters<\/td>\n<td>31.2B<\/td>\n<\/tr>\n<tr>\n<td>License (model card)<\/td>\n<td>apache-2.0<\/td>\n<\/tr>\n<tr>\n<td>Smallest VRAM tier<\/td>\n<td>24GB<\/td>\n<\/tr>\n<tr>\n<td>Articles<\/td>\n<td>1<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<\/div>\n<h2>Hardware Requirements<\/h2>\n<p><strong>Estimated requirements (calculated by Local Model Watch)<\/strong> \u2014 31.2B parameters<\/p>\n<div class=\"lmw-table-scroll\" tabindex=\"0\" style=\"overflow-x:auto;-webkit-overflow-scrolling:touch;max-width:100%;\">\n<table style=\"width:max-content;min-width:100%;border-collapse:collapse;\">\n<thead>\n<tr>\n<th>Your VRAM<\/th>\n<th>Quantization<\/th>\n<th>File size<\/th>\n<th>Est. memory needed<\/th>\n<\/tr>\n<\/thead>\n<tbody>\n<tr>\n<td>24GB (RTX 4090 \/ 3090, etc.)<\/td>\n<td>IQ4_NL<\/td>\n<td>18.7GB<\/td>\n<td>22.5GB<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<\/div>\n<p><em>Memory estimates add a 20% runtime overhead (KV cache, etc.) to the actual size of the distributed files. Actual usage varies with context length, batch size and inference engine. These figures are computed by this site from file sizes, not published by the model&#8217;s authors. File sizes are measured from the converted build <a href=\"https:\/\/huggingface.co\/XingChen-AGI\/Xing4.0-29B-A4B-GGUF\">XingChen-AGI\/Xing4.0-29B-A4B-GGUF<\/a>. Compare with other models in our <a href=\"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/vram-guide-en\/\">VRAM quick reference<\/a>. What the quantization names mean: <a href=\"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/glossary-quantization-en\/\">glossary<\/a>.<\/em><\/p>\n<h2>Can You Run It Locally?<\/h2>\n<p><strong>Runs in Ollama, LM Studio and llama.cpp via a converted build.<\/strong><\/p>\n<p>The publisher ships safetensors, but <a href=\"https:\/\/huggingface.co\/XingChen-AGI\/Xing4.0-29B-A4B-GGUF\">XingChen-AGI\/Xing4.0-29B-A4B-GGUF<\/a> provides a GGUF build you can use.<\/p>\n<p><strong>License \u2014 <code>apache-2.0<\/code> (Commercial use allowed):<\/strong> Permits commercial use, modification and redistribution. Redistribution requires including the license and stating changes; includes a patent grant.<\/p>\n<p><strong>Compression:<\/strong> the IQ4_NL build measures 5.15 bits per weight \u2014 about 32% the size of the original 16-bit weights, calculated by this site from the actual file sizes.<\/p>\n<p><em>Compiled by this site&#8217;s code from the published formats, converted builds we have found, and each engine&#8217;s own model registry. &#8220;Not found&#8221; means we have not seen such a build, not that none exists. License summaries are not legal advice \u2014 check the publisher&#8217;s original terms before relying on them.<\/em><\/p>\n<h2>Quantized and Converted Variants<\/h2>\n<div class=\"lmw-table-scroll\" tabindex=\"0\" style=\"overflow-x:auto;-webkit-overflow-scrolling:touch;max-width:100%;\">\n<p class=\"lmw-table-hint\" style=\"margin:0 0 4px;font-size:0.85em;opacity:0.7;\">\u2192 Scroll horizontally to see all columns<\/p>\n<table style=\"width:max-content;min-width:100%;border-collapse:collapse;\">\n<thead>\n<tr>\n<th>Added<\/th>\n<th>Publisher<\/th>\n<th>Format<\/th>\n<th>Repository<\/th>\n<th>Smallest VRAM tier (build, est. memory)<\/th>\n<\/tr>\n<\/thead>\n<tbody>\n<tr>\n<td>2026-09-28<\/td>\n<td>XingChen-AGI<\/td>\n<td>GGUF (imatrix)<\/td>\n<td><a href=\"https:\/\/huggingface.co\/XingChen-AGI\/Xing4.0-29B-A4B-GGUF\">XingChen-AGI\/Xing4.0-29B-A4B-GGUF<\/a><\/td>\n<td>IQ4_NL 22.5GB (fits in 24GB VRAM)<\/td>\n<\/tr>\n<tr>\n<td>2026-09-28<\/td>\n<td>XingChen-AGI<\/td>\n<td>FP8<\/td>\n<td><a href=\"https:\/\/huggingface.co\/XingChen-AGI\/Xing4.0-29B-A4B-FP8\">XingChen-AGI\/Xing4.0-29B-A4B-FP8<\/a><\/td>\n<td>FP8 37.1GB (fits in 48GB VRAM)<\/td>\n<\/tr>\n<tr>\n<td>2026-09-28<\/td>\n<td>mlx-community<\/td>\n<td>MLX<\/td>\n<td><a href=\"https:\/\/huggingface.co\/mlx-community\/Xing4.0-29B-A4B-OptiQ-4bit\">mlx-community\/Xing4.0-29B-A4B-OptiQ-4bit<\/a><\/td>\n<td>MLX 4bit 23.0GB (fits in 24GB VRAM)<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<\/div>\n<p>File sizes of each build:<\/p>\n<ul>\n<li>Available builds in XingChen-AGI\/Xing4.0-29B-A4B-GGUF: IQ4_NL 18.7GB<\/li>\n<li>Available builds in XingChen-AGI\/Xing4.0-29B-A4B-FP8: FP8 30.9GB<\/li>\n<li>Available builds in mlx-community\/Xing4.0-29B-A4B-OptiQ-4bit: MLX 4bit 19.2GB<\/li>\n<\/ul>\n<p><em>This section is appended automatically by Local Model Watch when a converted build of this model appears after publication. Memory figures are estimated from the size of the distributed files.<\/em><\/p>\n<h2>Articles (the family&#8217;s own models first, then newest)<\/h2>\n<div class=\"lmw-table-scroll\" tabindex=\"0\" style=\"overflow-x:auto;-webkit-overflow-scrolling:touch;max-width:100%;\">\n<table style=\"width:max-content;min-width:100%;border-collapse:collapse;\">\n<thead>\n<tr>\n<th>Published<\/th>\n<th>Model<\/th>\n<th>Type<\/th>\n<th>Article<\/th>\n<\/tr>\n<\/thead>\n<tbody>\n<tr>\n<td>2026-09-28<\/td>\n<td>XingChen-AGI\/Xing4.0-29B-A4B<\/td>\n<td>New Models<\/td>\n<td><a href=\"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/2026\/09\/28\/xing4-0-29b-a4b-china-telecom\/\">Xing4.0-29B-A4B 29B MoE Model Strong in Coding Agents: 24GB+ VRAM<\/a><\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<\/div>\n<h2>Repositories<\/h2>\n<ul>\n<li><a href=\"https:\/\/huggingface.co\/XingChen-AGI\/Xing4.0-29B-A4B\">XingChen-AGI\/Xing4.0-29B-A4B<\/a><\/li>\n<\/ul>\n<p><em>Last updated 2026-09-28 (JST). Assembled by code from our article log; no text on this page is written by an AI model.<\/em><\/p>\n","protected":false},"excerpt":{"rendered":"<p>Everything Local Model Watch has published about the Xing4.0 family: 1 article(s) covering the base model and its fine-tunes, plus converted builds we tracked after publication. Memory requirements below are computed by this site from file sizes, not quoted from model cards. Part of our model family index. At a Glance Item Value Base model(s) [&hellip;]<\/p>\n","protected":false},"author":1,"featured_media":0,"parent":0,"menu_order":0,"comment_status":"closed","ping_status":"closed","template":"","meta":{"footnotes":""},"class_list":["post-6182","page","type-page","status-publish","hentry"],"lang":"en","translations":{"en":6182,"ja":6181},"_links":{"self":[{"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/pages\/6182","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/pages"}],"about":[{"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/types\/page"}],"author":[{"embeddable":true,"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/comments?post=6182"}],"version-history":[{"count":1,"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/pages\/6182\/revisions"}],"predecessor-version":[{"id":6208,"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/pages\/6182\/revisions\/6208"}],"wp:attachment":[{"href":"https:\/\/localmodelwatch.tsuchitsuchi.com\/en\/wp-json\/wp\/v2\/media?parent=6182"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}