hermes-ai.netis an unofficial, independent community guide to Hermes Agent, with localized docs, release notes, desktop notes, and practical setup paths.
China Telecom Releases Xing4.0, an Open 29B Coding Agent Built on Huawei Chips
China Telecom released Xing4.0-29B-A4B , a 29B MoE with 4B active parameters, Apache 2.0. Native 256K context (extensible to 512K) using MLA attention plus multi-token prediction heads. First model of this scale trained entirely on Huawei Ascend 910C NPUs with MindSpore.
Source
AlphaSignal
Published
Author
AlphaSignal
Read
2 min read
Takeaways
China Telecom released Xing4.0-29B-A4B , a 29B MoE with 4B active parameters, Apache 2.0.
Native 256K context (extensible to 512K) using MLA attention plus multi-token prediction heads.
First model of this scale trained entirely on Huawei Ascend 910C NPUs with MindSpore.
Hits 75 on SWE-bench Verified and 57.5 on Terminal-Bench 2.1, beating Gemma4-26B-A4B.
Training throughput improved ~96% via fused operators and MoE communication tuning on Ascend.
Ships with vLLM, SGLang, KTransformers support and adapters for Claude Code and OpenCode.
Developer
China Telecom Artificial Intelligence Technology Co., Ltd.
Parameters
29B total, about 4B active per token
Architecture
40-layer MoE with MLA, MTP, and an mHC block
Experts
64 routed experts, 4 selected per token, plus 1 shared expert
Context
256K native; the model card describes extension to 512K
Training stack
Huawei Ascend 910C, MindSpore, and MindFormers
License
Apache 2.0
Primary workloads
Coding agents, terminal tasks, tool use, and long-context reasoning
Reporting is indexed from AlphaSignal. Rights remain with the original publisher and cited sources.