Perplexity Open-Sources Lily Local Inference Engine
The engine powers hybrid compute tasks in the company's Mac app.
TLDR
Perplexity announced it is open-sourcing Lily, a local inference engine built for Apple silicon. The release targets the Qwen3.6-35B-A3B model and supports hybrid compute in the Mac app by keeping on-device work from slowing cloud tasks. Company statements describe custom Metal kernels and a compact Rust runtime that manages session state and generation. Posts from Perplexity and its executives note the engine was purpose-built rather than adapted from general frameworks. The code appears in the pplx-garden repository, and a company blog post covers the on-device optimizations.
Combined views
625.1K
9 Sources, first seen 28d ago