Y
Hacker News
new
|
ask
|
show
|
jobs
by
foxhill
25 days ago
that's a nice example. on my M4, i measured 3.4s vs. 0.42s. honestly surprised there's ~10x improvement to be found.
as you've pointed out, you've literally micro-optimised this - isn't this what you'd expect? :)
1 comments
ks6g10
24 days ago
But what is the resulting assembly? I would assume completely different!
link