AINewsnow

How comparable is a MacBook Pro M5 Pro 64GB 18/20 vs RTX4090 | 128 GB DDR5

Hello guys, I want to buy a MBP for coding and developing (mostly mobile apps and for apple vision pro & Unreal Engine). At work I have access to a setup with RTX4090 | 128 GB DDR5 where I can run my llama.cpp and access it with tailscale (qwen3.8-27b Q4-K-M at ~80/s and ~127k context window if i r…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-10-07 10:56 · r/LocalLLM
    How comparable is a MacBook Pro M5 Pro 64GB 18/20 vs RTX4090 | 128 GB DDR5

More stories

  1. Qwen3.8-Flash-Next-Q8_0 running on a V100 @ 130Watts 32GB Vram and 128GB System Ram — r/LocalLLM
  2. Muse launches on the iPad — The Verge AI
  3. LLM Inference Dashboard — r/LocalLLaMA
  4. RPC: add `-sm tensor` by am17an · Pull Request #26610 · ggml-org/llama.cpp — r/LocalLLaMA
  5. Looking for developer-friendly inference providers who give you enough API credits to experiment [D] — r/MachineLearning
  6. Ternary Bonsai 2 27B on a 12 GB Intel Arc B580: 128K context, ~80-90 t/s code, 250+ t/s edits, 44 t/s at 115K — r/LocalLLM
  7. Local AI ecosystem overview — r/LocalLLaMA
  8. Gemma 4 26B-A4B and a 37 GB Qwen3.6 MoE running in a browser tab on a 24 GB Mac — experts streamed from disk, output matches llama.cpp — r/LocalLLaMA

Get the daily brief of stories like this at 6:30 every morning →