Yes, it is a 10x markup on the API prices. Depending on whether you factor in cooling costs, data center staff, etc. Or GPU costs and the electricity the GPUs are using only.
Either way, inference is very much where the money is made, training is where the money is lost.
Either way, inference is very much where the money is made, training is where the money is lost.