AINewsnow

Perplexity Open Sources Lily: A Rust + Metal Inference Engine for Qwen3.6-35B-A3B on Apple Silicon

This story is from 2026-09-03. It is preserved in the archive; the latest stories are on the live feed.

Most "runs locally on your Mac" stacks are a general-purpose runtime pointed at whatever model you downloaded. Perplexity just argued that the generality itself is the bottleneck. They open sourced Lily — the local inference engine behind Hybrid Compute in Perplexity Computer. A Rust runtime with h…

Read the full story at r/machinelearningnews ↗

Timeline · 1 report

  1. 2026-09-03 07:08 · r/machinelearningnews
    Perplexity Open Sources Lily: A Rust + Metal Inference Engine for Qwen3.6-35B-A3B on Apple Silicon

More stories

  1. Meta's personal AI agent Muse climbs to No. 1 among free apps on Apple's US App Store, ahead of ChatGPT; Muse launched on September 8 (Georgia Hennessy/Business Insider) — Techmeme
  2. I spent hours going through 100+ page PDFs, so I built a tool that highlights exactly where the answer came from. It's now completely open-source. — r/ChatGPTCoding
  3. Week in review: OpenAI ships managed Agents API, Apple's new Siri reportedly runs on Gemini, and three vendors add agent spend controls — r/artificial
  4. Apple reportedly building server packed with M-series Ultra chips for AI — Ars Technica AI
  5. Apple Reference Images Explained: The iPhone 18 Pro’s Hardware Solution to AI Slop — CNET AI
  6. AEO (Answer Engine Optimization): The Complete Guide to Getting Your Content Surfaced by AI — DEV Community — AI
  7. Apple M5 Ultra Scores Big GPU Gains in Leaked Geekbench Benchmark — r/LocalLLM
  8. Considering Purchasing a Studio M5 — r/LocalLLM

Get the daily brief of stories like this at 6:30 every morning →