Running three Qwen3.5 122B sessions on one Mac Studio with an on-disk KV cache
A Reddit user reportedly ran three concurrent Qwen3.5 122B sessions on a single Mac Studio, serving over 93% of prompt tokens from an on-disk KV cache instead of recomputing on GPU.