Perplexity Open Sources Lily: A Rust + Metal Inference Engine for Qwen3.6-35B-A3B on Apple Silicon

AITopTools Editorial TeamSeptember 3, 2026

What changed

Perplexity has released Lily as open-source software for running Qwen3.6-35B-A3B locally on Apple Silicon computers. Built with Rust and Apple’s Metal framework, it averaged faster prompt processing and response generation than MLX-LM on a 40-core, 128 GB M5 Max.

What this means for you

Developers with compatible Apple hardware can now access the code behind a local component of Perplexity Computer and potentially run this model more efficiently. Lily is narrowly optimized for one model and Apple’s chip family, so its usefulness may not extend to other hardware or models.

Related AI tools

Explore directory listings connected to the products, companies, and workflows in this story.

Related AI news

Read Google Releases Gemini 3.8 Live and 3.8 Live Extended Thinking for Production Grade Voice Agents
New AI featuresSep 15, 2026

Google Releases Gemini 3.8 Live and 3.8 Live Extended Thinking for Production Grade Voice Agents

The models are available now in Google’s Gemini API and AI Studio, allowing developers to build voice applications at $0.005 per minute for audio input. Generated audio includes Google DeepMind’s SynthID watermark, which identifies it as AI-created.

MarkTechPostSee why it matters
Read Meta now lets AI agents handle the boring parts of WhatsApp Business setup
AI toolsSep 15, 2026

Meta now lets AI agents handle the boring parts of WhatsApp Business setup

Developers can use tools such as Claude, Cursor, Codex, and ChatGPT to reduce the manual work involved in launching WhatsApp Business messaging. The feature is aimed at developers, and the feed does not specify pricing or broader access details.

TechCrunch AISee why it matters