←Back to NewsAI News/RetrievalnewsRetrievalInfraPerplexity's ROSE Serving Stack Beats vLLM on Speed and LatencyPerplexity open-sources details of Ivy, Tulip, and ROSE, its Rust plus Python embedding stack that beats vLLM on latency and throughput.SourcePerplexityPublishedSep 4, 2026, 9:17 PMAuthorAlphaSignal NewsroomRead1 min readPerplexity open-sources details of Ivy, Tulip, and ROSE, its Rust plus Python embedding stack that beats vLLM on latency and throughput.Reporting is indexed from AlphaSignal. Rights remain with the original publisher and cited sources.Read original report ↗Next readsGoogle Research · newsGoogle's Retrieve-for-Train Slashes AI Search Latency by 20xAlphaSignal · repoTamara Tran's fast-jev-compaction Stops Claude Code From Forgetting Critical Tool OutputExa · newsExa Snapshot Lets AI Search the Web as It Existed Years Ago