Perplexity Will Open Source Its Faster Lily AI Engine For Apple Silicon
BrianFagioli writes: Perplexity has built a local artificial intelligence engine designed specifically for Apple silicon and the Qwen3.6-35B-A3B model. Called Lily, the engine uses a Rust runtime and custom Metal kernels, with neither PyTorch nor MLX in its execution path. Perplexity says Lily averaged 23 percent faster prompt processing and 35 percent faster token generation than MLX-LM on an M5…
We haven't written up this one. Slashdot has the full story — the link below goes straight to it.