
China Merchants Bank
Building a unified heterogeneous AI compute scheduling platform on Kubernetes with HAMi for topology-aware scheduling and fine-grained accelerator sharing.
100% hardware pool utilization with topology-aware scheduling
- Hardware pool utilization reached 100% via supernode module-aware paired allocation.
- Cross-machine scheduling for distributed training dropped by 30%.
- Fine-grained partitioning down to 1 GB memory / 1% compute per card.

