|
|
|
|
|
by Philpax
981 days ago
|
|
The idea is that you can standardise the quality of the training data by taking source articles and synthesizing new data with the same "voice" and structure, as well as being able to collate insights from multiple sources. This is the line of thinking behind the Phi lineup of models [0], as well as efforts to generate synthetic textbooks for training [1]. [0]: https://arxiv.org/abs/2309.05463 [1]: https://twitter.com/ocolegro/status/1712327588255809667 |
|