LaunchDevelopersSeptember 3, 2026

Perplexity open-sources Lily inference engine for Apple Silicon

Lily is a Rust + Metal inference engine specialized for Qwen3.6-35B-A3B on Apple silicon, powering hybrid compute in Perplexity Computer. It runs as a single-process runtime with an OpenAI-compatible chat-completions API.

How this story unfolded

same day · 1 report · 3 community posts · from Sep 2

  1. Sep 2
  2. Sep 3

More stories today

Open the live feed