AINewsnow

[Release] Nirvana Code: A single-binary Rust LLM engine built from the metal up for Apple Silicon (Metal 3, Persistent Prefix Cache, Speculative Decoding, Dual GGUF + MLX)

TL;DR: I built Nirvana Code, a native Apple Silicon coding assistant and local inference runner written in pure Rust. It bundles an interactive Terminal UI (Ratatui), a Cyberpunk Web UI, and an OpenAI-compatible API server into a single lean binary ( · GitHub: https://github.com/niravlekinwala/nirv…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-09-19 17:23 · r/LocalLLM
    [Release] Nirvana Code: A single-binary Rust LLM engine built from the metal up for Apple Silicon (Metal 3, Persistent Prefix Cache, Speculative Decoding, Dual GGUF + MLX)

More stories

  1. Week in review: OpenAI ships managed Agents API, Apple's new Siri reportedly runs on Gemini, and three vendors add agent spend controls — r/artificial
  2. Local LLM on iPhone 18 is impressive — r/LocalLLM
  3. Has anyone had this happen? — r/OpenAI
  4. Why Everyone Is Getting Excited About Personal AI Agents — The AI Daily Brief
  5. Anthropic, OpenAI, SpaceXAI, Google sued over call to ‘pace’ AI development — Politico Technology
  6. Gemini Hacked Three Companies in First Known Breakout by Google’s AI — Wall Street Journal Technology
  7. Meet the Data Agent in ChatGPT Work — OpenAI YouTube
  8. Introducing the Australian Youth Safety Blueprint — OpenAI News

Get the daily brief of stories like this at 6:30 every morning →