Y
Hacker News
new
|
ask
|
show
|
jobs
by
vikmals
22 days ago
fastllm targets the GPU, while colibri uses CPU inference only
1 comments
aliljet
22 days ago
I'd be curious about an.option that would allow glm use with a low end GPU like a 2080 ti...
link