1. The plagiarism aspect, and that most of the training data was used without permission
2. I haven't found it terribly useful in my personal work, as the data it was trained on was heavily polluted by incorrect information (at least in the field I'm using it)
For me there's two separate problems:
1. The plagiarism aspect, and that most of the training data was used without permission
2. I haven't found it terribly useful in my personal work, as the data it was trained on was heavily polluted by incorrect information (at least in the field I'm using it)