Qwen3.8 Flash Next Reasoning Modes: Off vs Low vs Medium vs Xhigh kaitchup.substack.com
Benjamin Marie compares Qwen3.8 Flash Next with reasoning off, low, medium and xhigh across 14 benchmarks and 23,331 problems per setting, measuring accuracy and generated tokens. He reports that low is surprisingly competitive with medium, while xhigh uses more than three times as many output tokens as medium and its accuracy gains vary by workload.