|
|
|
|
|
by dopamine_daddy
2 days ago
|
|
I would suspect during pre-training. Even before RLHF was a thing models exhibited this bias. My bet is that it has to do with the training data corpora being composed in large part of left leaning content, maybe from social media platforms like reddit. |
|
It's no surprise that the models would reflect most people's feelings. However what FT showed was that models are shifting the center of gravity towards the right. Their chart showed people are on a bi-modal distribution where most are on the left and some are on the very right, less in the middle.
Then they compared AI and showed how it moved the center of mass towards the center which is moving it towards the right (since most of the center of mass is on the left).
In other words, models are providing answers to the right of where most people would be.