Open-weight & owned stack · Quantization efficiency
OMLX Benchmarks Qwen 3.8 On Apple M5 Ultra
OMLX reported Qwen 3.8 27B generating 50 tokens per second and processing prompts at 1,800 tokens per second on Apple’s M5 Ultra.
Read the original at reddit.comOpens the publisher's site in a new tab