Hacker News new | ask | show | jobs
by abdullahkhalids 14 days ago
An interesting solution would be for these AI companies to train a few different versions of these models, all with different speech characteristics. Then, when you start a conversation, you get a random version.
3 comments

They can't, because they use RL with synthetic data and LLMs as judges. So the system naturally convergences towards certain load bearing, genuine, not just annoying but ridiculous verbal tics.

It's probably the reason most LLMs share the same tics across labs, because they cross train and distil each other's models on an industrial scale. You also can't escape it in generated text that's already online. So if, say ChatGPT first had some random idiosyncrasies, it then contaminated the entire AI ecosystem.

Or tech companies could stop staring at their own belly-buttons and realize there's a whole big world outside of Silicon Valley, and training on the writing styles and pattens of their bubble and its hangers-on is perhaps not all that useful outside of 415.

Apple used to be guilty of this back when you'd ask Siri what the temperature was, and any number above 79°F was followed by the word "Hot!"

People outside of office workers aren't using Claude/Codex etc. though. It's the only real audience. What's the use case outside of an office? Grocery lists?
Not true, lots of people who don't have office jobs are using AI and coding agents.
Yes that would probably help!