|
|
|
|
|
by acchow
10 days ago
|
|
I don’t think you need a 10-30b model for most smartphone use cases. But I meant to counter gp’s claim that “I can run a 27b model on an iPhone” is kind of pointless and disingenuous. Yes I’m sure someone will come up with a way to run a “27b model” at 0.1 bit quantization on an Apple Watch pretty soon misses the whole point of saying a model is “27b” in capability. Achieving a parameter count is not the point. And is almost meaningless |
|