Thinking about it some more, say you're a PhD student. You've done a bunch of research, and have bits and pieces of your thesis drafted, and now you have ChatGPT generate the thesis document for you. So you feed it what you have, including all your data along with explanations of what it is, how it relates to your work, and a page or two about the subject of your thesis. So far so good. But to output a thesis it needs to have looked at a huge number of quality papers and prior theses, to know what it should look like. So you feed it a dataset for this. Now, the first problem you're going to have is that all the data, research, and conclusions of these will pollute the correctness of yours. You can't trust the output to be correct, because it will fill in gaps with nonsense. The second problem is you will be accused of plagiarism, and rightfully so, since it effectively takes someone else's document and reasoning and drops in your data...