Best current open-source option for text rendering and reference-image conditioning in one model? (24GB)
Need both: a product/person reference image driving the composition, and legible rendered text in the output (headline, CTA). Qwen-Image-Edit handles reference well and text okay up to ~4-5 words, then it degrades into gibberish. Proprietary models are noticeably ahead here. Is anything open-source…
Read the full story at r/StableDiffusion ↗
Timeline · 1 report
- 2026-10-08 03:54 · r/StableDiffusion
Best current open-source option for text rendering and reference-image conditioning in one model? (24GB)