|
|
|
|
|
by ykonstant
17 days ago
|
|
I only have 16GB of RAM in my laptop, but I would love to run a powerful model like that even at 0.5 tokens/s. For the questions I am asking, such a speed is more than enough. If anyone has any suggestions, please let me know! I am not an expert on LLMs, just a humble mathematician. |
|
https://github.com/ErikTromp/colibri-hy3