Qwen 3.8 27B speed test highlights the promise and limits of multi-token prediction
Qwen 3.8 27B is drawing fresh attention from local-AI users after a new test showed a large jump in text-generation speed when Multi-Token Prediction, or MTP, was enabled. Geeky Gadgets reported a rise from 7.9 to 17.1 tokens per second…