Hacker News new | ask | show | jobs
by versteegen 18 days ago
More accurate to say RLHF aligns models to human preferences, most significantly to be helpful.