Running Qwen3.8-Flash-Next at Full 262K Context on a 128GB MacBook
Benchmarks for Qwen's new hybrid-attention MoE at its native 262K context on a 128GB M5 Max: the architecture that makes it fit, the day-0 recipe, and a depth sweep from 0 to 262K tokens.
7 min read