The Chinese GPU maker announced at the World AI Conference that it pre-trained the Mixture-of-Experts model from scratch using its own 10,000-GPU-class cluster. This achievement marks a significant step for domestic AI training at scale, previously dominated by the Huawei ecosystem.
Moore Threads' large-scale model training on domestic GPUs signals China's progress in reducing reliance on foreign AI compute for frontier models.