Qwen 3.8 vs 3.6 27b low reasoning loops way less now

Have seen some people say Qwen 3.8 still overthinks even when reasoning is set to low. Which on my case has been way better compared to 3.6, eveb on a 3 bit quant. I think it's worth mentioning that the default is actually xhigh, so first make sure to specify it if not already. Also, Qwen 3.8 has an additional parameter preserve_thinking . It allows to keep/discard the reasoning after every turn. So make sure its activated, otherwise the model may end up reasoning through the same stuff again. My personal experience is low loops way less than 3.6 Still not perfect but a significant improvement. TLDR: Qwen 3.8 on "low" loops less than 3.6. Check "preserve_thinking" and make sure you're not still on the default "xhigh".

添加评论
点赞收藏
点踩分享查看原文
评论
?
参与讨论