We're open-sourcing Lily, Perplexity's local inference engine for serving models locally on Apple Silicon. This powers Perplexity's newly introduced hybrid compute feature for the Mac app.
Today we’re open-sourcing Lily, the local inference engine we built for hybrid compute in Perplexity Computer. Lily is specialized for Qwen3.6-35B-A3B on Apple ...