P40Qwen
P40 Tested on Large Models: The Pitfalls of Local AI's Second-Hand GPU Route
A user runs Qwen 35B on a 2016 Nvidia P40, asking if IQ quantization slows inference. Behind it: the second-hand market for local AI hardware is being
Sep 28·2 min read