scraping the web to create a dataset isn't plagiarism, same with training a model on said scraped data, and calculating which words should come in what order isn't plagiarism too. I agree that datasets should be ethically sourced, but scraping the web is something that allowed such things as the search engine to be created, which made the web a lot more useful. Was creating google irresponsible?
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments
replies: