Hacker News new | ask | show | jobs
by ggcr 37 days ago
With Reinforcement Learning, inference is very present in post-training stages now too