Back to News
Project Maya Runs Z.ai's 321B GLM-5.3-Flash on Two Consumer GPUs and an SSD
Project Maya runs the 321B GLM-5.3-Flash mixture-of-experts model on one or two consumer NVIDIA GPUs by tiering experts across VRAM, RAM, and SSD.

Source
AlphaSignal
Published
Author
AlphaSignal Newsroom
Read
1 min read
Project Maya runs the 321B GLM-5.3-Flash mixture-of-experts model on one or two consumer NVIDIA GPUs by tiering experts across VRAM, RAM, and SSD.
Reporting is indexed from AlphaSignal. Rights remain with the original publisher and cited sources.
Read original report

