← Founder Notes
Archive

The coding model just moved to a single gpu. china telecom's xing4.0-29b-a4b, out september 20, is…

Yethikrishna ROriginal on Threads

the coding model just moved to a single gpu. china telecom's xing4.0-29b-a4b, out september 20, is a full-stack domestic open-weight coding agent model that fits on one rtx 3090, built for enterprises whose data cannot leave the building.

local-first ai is no longer a compromise.

Context

The Xing4.0-29B-A4B model card on Hugging Face describes 29B total and 4B active parameters, 256K context extensible to 512K, an agent-oriented architecture and training entirely on the Ascend NPU platform with MindSpore. Its GitHub README news line reads 2026-09-17 open-sourced, with an FP8 variant the same day, and lists llama.cpp and SGLang support as pending. A third-party hardware blog of 18 September says the official 4-bit GGUF is 18.72 GiB, that the vendor's llama.cpp guide was tested on an RTX 3090 at 64K context, that it cannot run in stock llama.cpp, Ollama or LM Studio yet and that it found no public tokens-per-second figure on a 3090. The card lists vendor-reported SWE-bench Verified 75.00 and Terminal-Bench 2.1 57.50.

How it compares

The first-party date is 17 September and not 20 September. Specs and Ascend training are first-party. Fits on one RTX 3090 rests on the third-party blog and was not confirmed in first-party text, and it needs a fork or the vendor's package today. The Apache-2.0 license appears on the third-party blog only and is unconfirmed. In the vendor's own table Qwen3.6-35B-A3B scores higher on SWE-bench Verified at 76.00 while Xing is higher on Terminal-Bench 2.1, and these are vendor-run figures. No first-party data-residency statement was found, so built for enterprises whose data cannot leave the building is the author's framing. That local-first is no longer a compromise is the author's opinion.

Related work

Watch next

  • The llama.cpp support merge, a first-party hardware guide and independent coding evaluations.

Sources

  1. Xing4.0-29B-A4B (Hugging Face)huggingface.co
  2. Xing4.0-29B-A4B README (GitHub)raw.githubusercontent.com
  3. Can I run Xing 4.0 29B A4B on an RTX 3090 (OpenClaw DC, 18 Sep 2026)openclawdc.com

Provenance

The note above is reproduced unedited from the original post, first published on Threads on 20 September 2026 at 18:02 IST. Sources are the papers and datasets the note draws on.

View the original post
Embed this note
<iframe src="https://founder.myndlabs.tech/notes/embed/the-coding-model-just-moved-to-a-single-Ddgl9yZCMKg" width="480" height="420" style="border:0;max-width:100%" loading="lazy" title="The coding model just moved to a single gpu. china telecom's xing4.0-29b-a4b, out september 20, is…"></iframe>

More notes