Moortech S5000 GPU Breakthrough Powers China Mobile's Jiutian AI Model
At the upcoming 9th Digital China Summit, China Mobile's self-developed "Jiutian" 35B general-purpose large language model will make its official public debut. As a significant advancement for the domestic computing ecosystem, Moore Threads recently announced that its flagship, full-featured GPU, the MTT S5000, has completed full-process adaptation and inference verification for this model.
The core of this adaptation lies in deep integration. Leveraging its proprietary MUSA software stack and the SGLang-MUSA high-performance inference engine, Moore Threads successfully implemented the entire inference pipeline for the "Jiutian" 35B model. Through collaborative optimization of the MUSA C development framework, the muDNN computing library, and the open-source MATE operator library, the MTT S5000 has been finely tuned for the specific attention mechanisms and long-sequence inference requirements of large models. This ensures efficient and stable performance when processing lengthy texts and handling high-concurrency requests.

The MTT S5000 computing card, serving as the technical foundation for this adaptation, has demonstrated exceptional capabilities. Built on the fourth-generation MUSA "Pinghu" architecture, this GPU delivers a maximum AI dense computing power of up to 1000 TFLOPS per card. Its hardware configuration features 80GB of high-capacity VRAM with a memory bandwidth of 1.6 TB/s, supporting full-precision computing from FP8 to FP64. Furthermore, a high inter-card interconnect bandwidth of 784 GB/s ensures excellent scalability in complex intelligent computing scenarios.
This collaboration not only validates the reliability of domestic GPUs in supporting core large models from central state-owned enterprises but also highlights Moore Threads' maturity in high-performance operator optimization and software ecosystem development. With the official launch of the "Jiutian" 35B model, this "domestic large model + domestic computing power" combination provides a highly relevant practical case for achieving independent and controllable computing infrastructure.
Related article
U.S. Stocks Hit Historic Milestone as AI and Aerospace Giants Prepare for Trillion-Dollar Debut
Elon Musk, Sam Altman, and Dario Amodei, three titans of the technology sector, are advancing toward initial public offerings for their respective ventures. With SpaceX, OpenAI, and Anthropic—three industry behemoths nearing trillion-dollar valuation
Swedish AI Startup Lovable Eyes $13.2 Billion Valuation After Major Funding Round
As AI-driven coding tools gain traction, Swedish startup Lovable has secured a major funding round. The company aims to raise $3 billion, potentially boosting its valuation to $13.2 billion—double the $6.6 billion recorded last December. Menlo Ventur
Google Tests Remy AI Agent for Gemini as Focus Shifts to User Control
According to Business Insider, Google is testing Remy, a new AI personal agent for Gemini. This tool aims to execute tasks on behalf of users, streamlining both professional workflows and daily routines.Currently, Remy is undergoing testing in an int
Related Special Topic Recommendations
Comments (0)
0/500
At the upcoming 9th Digital China Summit, China Mobile's self-developed "Jiutian" 35B general-purpose large language model will make its official public debut. As a significant advancement for the domestic computing ecosystem, Moore Threads recently announced that its flagship, full-featured GPU, the MTT S5000, has completed full-process adaptation and inference verification for this model.
The core of this adaptation lies in deep integration. Leveraging its proprietary MUSA software stack and the SGLang-MUSA high-performance inference engine, Moore Threads successfully implemented the entire inference pipeline for the "Jiutian" 35B model. Through collaborative optimization of the MUSA C development framework, the muDNN computing library, and the open-source MATE operator library, the MTT S5000 has been finely tuned for the specific attention mechanisms and long-sequence inference requirements of large models. This ensures efficient and stable performance when processing lengthy texts and handling high-concurrency requests.

The MTT S5000 computing card, serving as the technical foundation for this adaptation, has demonstrated exceptional capabilities. Built on the fourth-generation MUSA "Pinghu" architecture, this GPU delivers a maximum AI dense computing power of up to 1000 TFLOPS per card. Its hardware configuration features 80GB of high-capacity VRAM with a memory bandwidth of 1.6 TB/s, supporting full-precision computing from FP8 to FP64. Furthermore, a high inter-card interconnect bandwidth of 784 GB/s ensures excellent scalability in complex intelligent computing scenarios.
This collaboration not only validates the reliability of domestic GPUs in supporting core large models from central state-owned enterprises but also highlights Moore Threads' maturity in high-performance operator optimization and software ecosystem development. With the official launch of the "Jiutian" 35B model, this "domestic large model + domestic computing power" combination provides a highly relevant practical case for achieving independent and controllable computing infrastructure.
U.S. Stocks Hit Historic Milestone as AI and Aerospace Giants Prepare for Trillion-Dollar Debut
Elon Musk, Sam Altman, and Dario Amodei, three titans of the technology sector, are advancing toward initial public offerings for their respective ventures. With SpaceX, OpenAI, and Anthropic—three industry behemoths nearing trillion-dollar valuation
Swedish AI Startup Lovable Eyes $13.2 Billion Valuation After Major Funding Round
As AI-driven coding tools gain traction, Swedish startup Lovable has secured a major funding round. The company aims to raise $3 billion, potentially boosting its valuation to $13.2 billion—double the $6.6 billion recorded last December. Menlo Ventur





Home






