Hacker News new | ask | show | jobs
by smcnally 12 days ago
groq did an ASIC for llama and now for nvidia. Their cloud service is fast.

> NVIDIA Groq 3 LPU Inference Accelerator > The NVIDIA Groq 3 LPU is the next generation of Groq’s innovative language processing unit. Each LPX rack features 256 interconnected LPU accelerators that, together with the NVIDIA Vera Rubin platform, supercharge inference. Each LPU accelerator delivers 500 megabytes (MB) of SRAM, 150 terabytes per second (TB/s) of SRAM bandwidth, and 2.5 TB/s scale-up bandwidth.

https://www.nvidia.com/en-us/data-center/lpx/

1 comments

I wonder where those now worthless ASICs are rotting
if they were rotting NV wouldn't be advertising the product (nor hiring for it - which they are)