Open-weight & owned stack · Quantization efficiency
Qwen3.8 Flash Next Shows High Local Throughput
A user reported serving quantized Qwen3.8-Flash-Next at over 150 tokens per second singly and roughly 100 tokens per second across concurrent streams.
Read the original at reddit.comOpens the publisher's site in a new tab