Hacker News new | ask | show | jobs
by 5555watch 2 days ago
I'm curious, how hard/expensive it is to burn a really large model into silicon, and why aren't we doing this already?

Or, when we will start doing this, who's going to be able to do that in scale?

I'm seeing the TAALAS example, but it's only an 8B model, suggesting some real limitations parameter wise. And for 2.5kW?

1 comments

I can't speak for all cases, but the AI space is seeing improvements month by month, so it is beneficial to wait until it settles (a model becomes the standard in intelligence/price) before designing and mass producing an "LLM ASIC" of said model.

The big AI labs won't do that unless they are forced to, as they want you to spend more money on the big, expensive, frontier models (so they can live up to their valuation), so it's more likely that you will see this on smaller open weights models.