XingChen-AGI/Xing4.0-29B-A4B — new model trending #30 on Hugging Face
China Telecom's XingChen-AGI released Xing4.0-29B-A4B, an open-weights 29B-parameter MoE model (4B active) with 256K context, trained entirely on Ascend NPUs.
XingChen-AGI, the AI unit of China Telecom and successor to the TeleChat series, published Xing4.0-29B-A4B weights in Transformers format. The MoE model has 29B total parameters with 4B activated per token, 40 layers, MLA attention with 64 routed experts, and supports 256K context extensible to 512K. It is the first model at this scale trained entirely on Ascend NPU hardware with MindSpore, reporting roughly 96% training throughput gains from co-optimization. Benchmarks include 75.0 on SWE-bench Verified, 57.5 on Terminal-Bench 2.1, and 90.0 on AIME2026, and it deploys via vLLM, SGLang, and KTransformers.