Btw, are you using GPUs/TUs at all for local inference? I've been exploring what hardware options exist outside of Nvidia.