Open-weight & owned stack · Quantization efficiency
Inco AI Releases Splash Apple-Silicon Inference Engine
Inco AI announced Splash, an Apple-silicon inference engine claiming up to 144 tokens per second for Qwen3.8 27B.
Read the original at reddit.comOpens the publisher's site in a new tab