|
|
|
|
|
by subarctic
17 days ago
|
|
> Most useful LLM work is done in parallel I guess what I'm doing is not considered that useful then? I usually only have zero, one, or occasionally two things actively doing inference at a time, be it claude code sessions or one of the chatgpt/claude web interfaces, and i bet that's true for like 95% of people using llms. And anyway i bet even the hardcore people using a bunch of parallel agents would appreciate having access to local, private inference for some things. You're obviously right though that cloud inference isn't going away anytime soon |
|