← Back to Model Beat
Models·Sep 3·all news from September 3, 2026

Perplexity Open Sources Lily: A Rust + Metal Inference Engine for Qwen3.6-35B-A3B on Apple Silicon

Perplexity has released Lily, an open-source inference engine developed in Rust with custom Metal kernels designed specifically to run Qwen3.6-35B-A3B on Apple Silicon. By optimizing performance for this specific model and chip architecture, the engine achieves faster prefill and decode throughput compared to the existing MLX-LM framework.

ModelsQwen 3.6 35B-A3B

Covered by 1 source

Related stories

ModelsA Chinese AI called 'Qwen3.6-35B-A3B,' which is more powerful than Gemma4, has been released as an open model. - GIGAZINEApr 17ModelsQwen Team Open-Sources Qwen3.6-35B-A3B: A Sparse MoE Vision-Language Model with 3B Active Parameters and Agentic Coding Capabilities - MarkTechPostApr 16 · 3 sourcesModelsAlibaba (09988) Qwen 3.6-35B-A3B Model Open Source Release - 富途牛牛Apr 17 · 2 sources