I run a persistent local agent on a 16GB Air. Qwen3-8B on the GPU, a second brain on the Neural Engine, no cloud.
Her name is Bad Apple. Rust core, Swift platform layer, everything on the machine. She sees, remembers, audits everything she does, and comes back after restarts. Repo: https://github.com/savageAZfck/Bad_Apple ## The brains - **GPU:** Qwen3-8B-4bit on MLX. Chat, tools, verification. ~10 tok/s decod…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-10-09 23:20 · r/LocalLLM
I run a persistent local agent on a 16GB Air. Qwen3-8B on the GPU, a second brain on the Neural Engine, no cloud.