|
|
|
|
|
by searealist
5 hours ago
|
|
No one will ever derive any utility from running models at this speed. Please prove me wrong. Give me the number of tokens input and output (and dont forget about reasoning) and acceptable time to wait for it and the use case. |
|
This is like 98%.