So you speed up 1% of the pipeline by some factor, and the end result is unobservable for a human.
Plus, inference is not the only place tokenization happens. This can make a big difference during development of ML models.