Benchmarking ComfyUI on Docker with CUDA 12.4 vs Bare-Metal: 0% compute penalty and how to fix the /dev/shm OOM crash
This story is from 2026-09-12. It is preserved in the archive; the latest stories are on the live feed.
I ran extensive benchmarks comparing ComfyUI in Docker (Nvidia Container Toolkit / CUDA 12.4) against a bare-metal Linux setup (Ubuntu 24.04, PyTorch 2.4, CUDA 12.4) across FLUX.1-dev, SDXL, and SD 1.5 workloads. Key Results Compute / Generation Speed: 0.0% overhead. Docker achieved identical it/s…
Read the full story at r/comfyui ↗