hermes-ai.net

Read docs →
HermesHermes Agent Docs
Back to News

China Telecom Releases Xing4.0, an Open 29B Coding Agent Built on Huawei Chips

China Telecom released Xing4.0-29B-A4B , a 29B MoE with 4B active parameters, Apache 2.0. Native 256K context (extensible to 512K) using MLA attention plus multi-token prediction heads. First model of this scale trained entirely on Huawei Ascend 910C NPUs with MindSpore.

China Telecom Releases Xing4.0, an Open 29B Coding Agent Built on Huawei Chips
Source
AlphaSignal
Published
Author
AlphaSignal
Read
2 min read

Takeaways

  • China Telecom released Xing4.0-29B-A4B , a 29B MoE with 4B active parameters, Apache 2.0.
  • Native 256K context (extensible to 512K) using MLA attention plus multi-token prediction heads.
  • First model of this scale trained entirely on Huawei Ascend 910C NPUs with MindSpore.
  • Hits 75 on SWE-bench Verified and 57.5 on Terminal-Bench 2.1, beating Gemma4-26B-A4B.
  • Training throughput improved ~96% via fused operators and MoE communication tuning on Ascend.
  • Ships with vLLM, SGLang, KTransformers support and adapters for Claude Code and OpenCode.
DeveloperChina Telecom Artificial Intelligence Technology Co., Ltd.
Parameters29B total, about 4B active per token
Architecture40-layer MoE with MLA, MTP, and an mHC block
Experts64 routed experts, 4 selected per token, plus 1 shared expert
Context256K native; the model card describes extension to 512K
Training stackHuawei Ascend 910C, MindSpore, and MindFormers
LicenseApache 2.0
Primary workloadsCoding agents, terminal tasks, tool use, and long-context reasoning

Reporting is indexed from AlphaSignal. Rights remain with the original publisher and cited sources.

Read original report

Sources

  1. 01huggingface.co
  2. 02arxiv.org
  3. 03arxiv.org