|
|
|
|
|
by echelon
191 days ago
|
|
How much would it cost the community to pretrain something with a more modern architecture? Assuming it was carefully done in stages (more compute) to make sure no mistakes are made? I suppose we won't need to with the Chinese gifting so much open source recently? |
|
Quite a lot. Search for "Chroma" (which was a partial-ish retraining of Flux Schnell) or Pony (which was a partial-ish retraining of SDXL). You're probably looking at a cost of at least tens of thousands or even hundred of thousands of dollars. Even bigger SDXL community finetunes like bigASP cost thousands.
And it's not only the compute that's the issue. You also need a ton of data. You need a big dataset, with millions of images, and you need it cleaned, filtered, and labeled.
And of course you need someone who knows what they're doing. Training these state-of-art models takes quite a bit of skill, especially since a lot of it is pretty much a black art.