Qwen3.8-Flash-Next memory tiers: what fits in 16GB / 64·96GB / planned 128GB
Reading the 2026-08-26 Qwen3.8-Flash-Next (125B-A6B + 51B n-gram) from Unsloth GGUF, community MLX file sizes, and the official card only. Mapped onto the AI Tech Engine lab’s RTX 4080 16GB / M1 Max 64GB / M2 Max 96GB / planned 128GB. No unmeasured tok/s.