APIUp to 25% cheaper than official pricesTry the API →
HermesHermes Agent Docs
Back to News

Project Maya Runs Z.ai's 321B GLM-5.3-Flash on Two Consumer GPUs and an SSD

Project Maya runs the 321B GLM-5.3-Flash mixture-of-experts model on one or two consumer NVIDIA GPUs by tiering experts across VRAM, RAM, and SSD.

Project Maya Runs Z.ai's 321B GLM-5.3-Flash on Two Consumer GPUs and an SSD
Source
AlphaSignal
Published
Author
AlphaSignal Newsroom
Read
1 min read

Project Maya runs the 321B GLM-5.3-Flash mixture-of-experts model on one or two consumer NVIDIA GPUs by tiering experts across VRAM, RAM, and SSD.

Reporting is indexed from AlphaSignal. Rights remain with the original publisher and cited sources.

Read original report