Skip to main content
Customer case studies

Domestic accelerator company thousand-GPU production cluster

A domestic thousand-GPU cluster has been live and stable for months, continuously carrying real traffic and serving real users.

Domestic accelerator companyDomestic thousand-GPU clusterProduction delivery

Project information

Project overview

Environment, workload, and validation approach.

Customer type
Domestic accelerator company
Project stage
Production delivery
Hardware ecosystem
Domestic thousand-GPU cluster
Workload
Real traffic and user services
Delivery path
A production loop from cluster delivery to real services
Case type
Production validation case

Background and challenges

Project background

The project targeted a thousand-GPU production cluster at a domestic accelerator company. The goal was not only to complete test validation, but also to keep the network carrying real workloads long term and serving users continuously.

Validation or delivery approach

  • Deploy a domestic thousand-GPU cluster and put it into production
  • Carry real traffic continuously and run stably for months
  • Serve real users and close the loop with customer feedback

Results

Production operation

This case presents project value with production facts and does not package operating results that lack a unified comparison baseline as performance percentages.

  1. 01

    Thousand-GPU-scale deployment

    The domestic compute cluster went live.

  2. 02

    Production operation

    Continuously carried workloads and ran stably for months.

  3. 03

    Real-user service

    Continuously served real users.

  4. 04

    Positive customer feedback

    Formed a production loop from cluster delivery to user feedback.

Value summary

Case value

ZCube moved from test validation to production validation, forming a loop of sustainable operation, real service, and customer feedback on a domestic thousand-GPU cluster.

  • Shows ZCube availability in real production workloads
  • Forms replicable delivery and operations experience for domestic thousand-GPU clusters
  • Supports long-term partnerships with continuous service capability

Customer case studies

Other customer case studies

See more ZCube validation and production practice across compute environments, workloads, and project stages.

Domestic large-model companyThousand-GPU NVIDIA inference clusterProduction delivery

Large-model company NVIDIA thousand-GPU inference production cluster

Using a coding inference service and comparing against traditional ROFT, this case validates that ZCube unlocks more effective inference compute through a flat fabric and more balanced network communication.

15%

Average GPU inference throughput increase

40.6%

Time to first token (TTFT) P99 reduction

33%

Network hardware cost reduction

View case study
Training groundDomestic GPU inference clusterIndependent validation

Training-ground domestic GPU inference cluster

In a domestic AI computing ecosystem, compare ZCube with traditional Clos across collective communication, training, and inference.

14.5%

All-to-All throughput increase

22%–30%

TTFT P99 reduction

7%–10%

Inference throughput increase

View case study
A telecom operatorDomestic GPU clusterIndependent validation

Operator domestic GPU cluster

Compare ZCube with Clos and validate performance gains in collective communication, model training, and multi-tenant end-to-end inference.

119%

Collective communication performance increase

20.1%

Model-training throughput increase

11.5%

Inference latency reduction

View case study
Start with the network

Make the next AI data center more efficient, starting with the network

Whether you are building a new cluster, upgrading an existing network, tuning performance, or deploying intelligent operations, we can assess how the network can unlock more effective compute.

Contact us