Hacker News new | ask | show | jobs
by amarshall 47 days ago
Thinking doesn’t change output speed. Anthropic’s models are ~ 40–60 t/s median output speed.
1 comments

Do you have access to Anthropic model weights to run them locally?
No, and having that is not required to know output speed nor the effect of thinking, so I don’t see the point in such a superfluous, indirect question.

As for the question you’re likely asking: benchmarks that include speed across many models and providers available at various places e.g. https://artificialanalysis.ai/leaderboards/models