|
|
|
|
|
by AnthonyMouse
12 days ago
|
|
> But commercial chat models are specifically tuned in a way that maximizes user engagement. It's that specific tuning that is very easy to spot when reading AI slop, and that's not surprising that it's easy to spot automatically either. There are two problems with this. The first is that it would still misclassify human-authored text written under the same incentive, and most people have various incentives to "maximize engagement". And the second is that then people would just make other models that are tuned for defeating that sort of classifier, which would be used whenever the classifier is being used. |
|
That may not last if AI companies start trying to build models that fool it, but for the time being at least, modern models do have strong tells.