I’m assuming a lot of its training data is synthetic and distilled from Chinese models that were themselves trained from pirated data and distilled from American models trained on pirated data. It would be quite remarkable if the training data involved no piracy whatsoever. Then again, it’s open source, so I suppose it would be essentially reversing a reverse Robin Hood.
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments
replies: