Hacker News new | ask | show | jobs
by vikmals 22 days ago
fastllm targets the GPU, while colibri uses CPU inference only
1 comments

I'd be curious about an.option that would allow glm use with a low end GPU like a 2080 ti...