Perplexity Lily: 1.35x Faster Local AI Than MLX on Mac
Perplexity open-sources Lily, a Rust and Metal inference engine that beats MLX by 1.35x on decode for Qwen3.6-35B-A3B on Apple Silicon. Code is ...
AI coding tools, LLMs, agents, and AI-assisted development