hermes-ai.net

Read docs →
HermesHermes Agent Docs
Back to News

Perplexity's Lily Beats MLX-LM by 1.35x Running Qwen3.6 on Apple Silicon

Perplexity open-sources Lily, a Metal-based inference engine tuned for Qwen3.6-35B-A3B that beats MLX-LM by 1.23x prefill and 1.35x decode on M5 Max.

Perplexity's Lily Beats MLX-LM by 1.35x Running Qwen3.6 on Apple Silicon
Source
Perplexity
Published
Author
AlphaSignal Newsroom
Read
1 min read

Perplexity open-sources Lily, a Metal-based inference engine tuned for Qwen3.6-35B-A3B that beats MLX-LM by 1.23x prefill and 1.35x decode on M5 Max.

Reporting is indexed from AlphaSignal. Rights remain with the original publisher and cited sources.

Read original report