Hacker News new | ask | show | jobs
by woadwarrior01 12 days ago
> In any case I am not sure pivoting from running local models to "cloud offering" (as in providing llm inference at their severs) is a sensible choice granted there is already competition in that space and they have no leverage there.

I agree. Incidentally, this is exactly what ollama are doing too.

1 comments

I can definitely see a world where you run stuff local first, for all the reasons we know. Sometimes, you are going to want more powerful models, faster, you’re travelling, etc. You might only use these 5 to 10% of the time, but I’m guessing that’s the market they want.
But one can already switch models in a harness between llm providers. I doubt such a pivot can get some much traction.