WHIRL v0.1.3 — native Windows LLM engine for the Radeon AI PRO R9700: up to 2.8× llama.cpp on the same GGUF, same answers bit-for-bit
WHIRL is an open-source (Apache-2.0) inference engine for the AMD Radeon AI PRO R9700 (RDNA 4, 32 GB) on Windows : pure C++/HIP, every kernel included, no WSL or Docker. Just the AMD driver. vs llama.cpp b11214 — same GGUF, same prompts, same R9700, Swift-1.5 27B MXFP4: Decode on coding prompts: 11…
Read the full story at r/LocalLLM ↗