Perplexity Open Sources Lily Inference Engine for Apple Silicon
Lily is a local inference engine built for Apple silicon and the Qwen3.6-35B-A3B model. Perplexity reports faster prompt processing and response generation than MLX-LM in tests on an M5 Max MacBook Pro.