Hacker News new | ask | show | jobs
by 0x00cl 697 days ago
> growing industry of people whose job it is to write content to train LMs against

Do you have an example of this?

How do they differentiate content written by a person v/s written by LLM, I'd expect there is going to be people trying to "cheat" by using LLMs to generate content.

1 comments

> How do they differentiate content written by a person v/s written by LLM

Honestly, not sure how to test it, but this is B2B contracts, so hopefully there's some quality control. It's part of the broad "training data labeling" business, so presumably the industry has some terms in contracts.

ScaleAI, Appen are big providers that have worked with OpenAI, Google, etc.

https://openai.com/index/openai-partners-with-scale-to-provi...