TIL about llama.cpp's RPC (Remote procedure call), might be better than Vulkan? YMMV
My system is "unique" to say the least. AMD R9700 5070 TI 16gb 4070 running on a Asus WS Pro X570 Ace (96 GB DDR4) I wanted to test running the highest fidelity Qwen3.8:27b leveraging Vulkan due to the completely mismatched GPUS. Whipped together a config did some testing. Didn't think the numbers…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-09-30 17:00 · r/LocalLLM
TIL about llama.cpp's RPC (Remote procedure call), might be better than Vulkan? YMMV