zram for model weight offloading — how much of a speed hit in practice?
Setup: i5-14400 + RTX 4070 + 32GB RAM Running MiniMax H3 in ComfyUI with CPU offload, hitting a 25s ceiling per generation at 0.4MP. CPU usage during offload sits around 7%, so there's clearly headroom there. Considering zram to effectively extend usable RAM for offloaded layers. Rough math with zs…
Read the full story at r/comfyui ↗
Timeline · 1 report
- 2026-09-26 17:28 · r/comfyui
zram for model weight offloading — how much of a speed hit in practice?