Hacker News new | ask | show | jobs
by foxhill 25 days ago
that's a nice example. on my M4, i measured 3.4s vs. 0.42s. honestly surprised there's ~10x improvement to be found.

as you've pointed out, you've literally micro-optimised this - isn't this what you'd expect? :)

1 comments

But what is the resulting assembly? I would assume completely different!