|
|
|
|
|
by i5heu
12 days ago
|
|
Looking at their career page it looks like they do not care that much about PR at the moment... considering that they have a live chat with 14000t/s via Llama 3.1 8B[0] i don't really think they need to do PR either. So i guess maybe they currently try to solve a very hard problem with a small focused group before scaling or they are dysfunctional. Also Llama 3.1 8B is a dense model AFAIK and they are fast by nature.
As there are not a lot of dense models these days i could imagine that they try to optimise for MOE models. [0] https://chatjimmy.ai/ |
|