← Back to Model Beat
Hardware·Sep 6·all news from September 6, 2026

Perplexity Details Its GPU Embedding Stack: How Ivy, Tulip and ROSE Serve pplx-embed

Perplexity has published technical details regarding the infrastructure behind its pplx-embed model, highlighting the use of custom components named Ivy, Tulip, and ROSE. These systems are designed to optimize the speed and cost of running embedding models across large indexes, which is a primary constraint for the retrieval quality of AI search products. By refining this serving stack, the company aims to improve the efficiency of mapping search queries to relevant information.

Covered by 1 source

Related stories

HardwareSparks Fly: NVIDIA Accelerates Local AI at IFA 2026Sep 3 · 8 sourcesHardwareQualcomm Signs Deal to Provide Amazon With Custom AI ChipsSep 8 · 6 sourcesHardwareASML, TSMC, Samsung, Intel Back 12-Inch Masks for AI ChipsSep 8 · 2 sourcesHardwareMalaysia Eyes Huawei Chips for AI Project Despite US WarningSep 7