I need help with MAX OUTPUT rate limit
This story is from 2026-09-06. It is preserved in the archive; the latest stories are on the live feed.
Hey guys i tried to run locally qwen 3.6 35b a3b, i was using Cline in VSCode i’ve set the max context to 48-64k(some tests) but after long work qwen started looping, and if it was not looped, it just stops because of max tokens output, what can i do, is there any ways to fix that?
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-09-06 11:14 · r/LocalLLM
I need help with MAX OUTPUT rate limit