Hacker News new | ask | show | jobs
by i5heu 12 days ago
Looking at their career page it looks like they do not care that much about PR at the moment... considering that they have a live chat with 14000t/s via Llama 3.1 8B[0] i don't really think they need to do PR either.

So i guess maybe they currently try to solve a very hard problem with a small focused group before scaling or they are dysfunctional.

Also Llama 3.1 8B is a dense model AFAIK and they are fast by nature. As there are not a lot of dense models these days i could imagine that they try to optimise for MOE models.

[0] https://chatjimmy.ai/