Qwen team also released two fine-tuned Qwen3.5-VL 9B models (LLM+Vision) for image edit and text to image prompt enhancement + system prompts
Their original Gihub repo suggests prompt enhancer model should guess aspect ratio and resolution from user input and pass it inside JSON for generation along with enhanced prompt Original prompts Text 2 image Image 2 image edit Model Weights Two fine-tuned Qwen3.5-VL 9B checkpoints with unified co…
Read the full story at r/StableDiffusion ↗
Timeline · 1 report
- 2026-09-20 19:49 · r/StableDiffusion
Qwen team also released two fine-tuned Qwen3.5-VL 9B models (LLM+Vision) for image edit and text to image prompt enhancement + system prompts