Hacker News new | ask | show | jobs
by anon373839 2 hours ago
Well, my position is that the software that defines/runs the model is the open-source part.

The weights are the asset that makes it useful, but I don’t think 40T tokens of pre-training data should be required to call the model itself open source when you can inspect every line of code in the model without that, as well as instantiate and run the model with randomly initialized weights.