Perplexity Open Sources Lily: A Rust + Metal Inference Engine for Qwen3.6-35B-A3B on Apple Silicon
MarkTechPosten
MarkTechPost
AI Global WirePerplexity has open sourced Lily, the local inference engine behind Hybrid Compute in Perplexity Computer. Built in Rust with custom Metal kernels for one model on one chip family, it averages 1.23x MLX-LM's prefill throughput and 1.35x its decode throughput on a 40-core, 128 GB M5 Max. The post Perplexity Open Sources Lily: A Rust + Metal Inference Engine for Qwen3.6-35B-A3B on Apple Silicon appeared first on MarkTechPost .
This is a short summary published by AI Global Wire. The full article is owned and hosted by MarkTechPost — open it there to read it in full.
Read the full story at MarkTechPost- Sök-AI
- Verktyg
Related AI news
- Cloudflare taps OpenAI’s cyber models to find and block code flawsSiliconANGLE · September 3, 2026
- OpenAI starts rolling out its next-generation GPT-6 Astra modelSiliconANGLE · September 3, 2026
- Prediction Market Betting Is Getting People Banned and ArrestedWIRED AI · September 3, 2026
- OpenAI Releases GPT-6 Astra: A 1.05M-Context Computer-Use Model Gated Behind a ‘Critical’ Cyber ThresholdMarkTechPost · September 3, 2026
- Broadcom and Supermicro unify AI factory managementSiliconANGLE · September 3, 2026
- Nvidia confirms $12.9B acquisition of AI hosting platform Hugging FaceSiliconANGLE · September 3, 2026