Alibaba Cloud runs 2.4T-parameter model on homegrown AI supernode
The deployment marks the first time a domestically developed AI supernode has successfully run inference for a model with more than 2 trillion parameters.
The deployment marks the first time a domestically developed AI supernode has successfully run inference for a model with more than 2 trillion parameters.

Unlike conventional chips based on the von Neumann architecture, computing-in-memory integrates computation directly into memory arrays.

The company’s approach seeks to address one of the biggest constraints in AI computing: the movement of data between processors and memory.