After minutes :D Btw this is just to get the inference graph dialed in. Q2 across two Mac Studios with 512GB can work at acceptable speed for chat at least. K3 was trained in MXFP4, I have the feeling it will quantize fine. I'll wait to have access to two Mac Studio 512GB.