Hacker News new | ask | show | jobs
by nadam 14 days ago
Does anyone know why these Claudisms exist? Are these kind of expressions a result of some kind of 'evolution', so they are more effective in chain-of-thought reasoning that other expressions? Or the preference of people doing some kind of human-feedback in the post training? Clearly probably not the 'average' expressions of their training data, and clearly if there would be no upside and it is easy to remove them they should be removed, as they are annoying, so I guess there is an upside or hard to remove them...
1 comments

I think LLMs have stopped trying to be accurate language models for a while now. There's probably some kind of feedback loop that encourages expressions that correlate with 'good responses', if there's nothing preventing it from using the same expressions again and again then it won't stop.