APIUp to 25% cheaper than official pricesTry the API →
HermesHermes Agent Docs
Back to News

Inco AI's Splash 1.3.0 Cuts Mac Prefill Wait by 1.49x for Coding Agents

Splash 1.3.0 cuts time to first token on local Macs by offloading KV cache to SSD and splitting prefill between the GPU and Neural Engine.

Inco AI's Splash 1.3.0 Cuts Mac Prefill Wait by 1.49x for Coding Agents
Source
Inco AI
Published
Author
AlphaSignal Newsroom
Read
1 min read

Splash 1.3.0 cuts time to first token on local Macs by offloading KV cache to SSD and splitting prefill between the GPU and Neural Engine.

Reporting is indexed from AlphaSignal. Rights remain with the original publisher and cited sources.

Read original report