Hacker News new | ask | show | jobs
by throwa356262 22 days ago
According to deepseek themselves, their current rates are NOT subsidised.

They have published tons of articles dedicated to performance and efficiency engineering. Feel free to have a look...

1 comments

Why is no other inference provider offering similar prices then?
How long did it take vLLM to implement deepseeks sparse attention from the r1 paper?

Does ananyone outside deepseek have a working code for the v4 compressed attention mechanism?

Has any other provider managed to bypass CUDA and program the compute engines in their native assembly language to get 10% more performance out of them?

There is your answer.