Comment by jujube3

15 days ago

I've explained this many times. LLMs don't "copy original work." ChatGPT doesn't contain copies of books inside it, any more than your brain is a copy of the various books you have read. The LLM model isn't physically big enough for that, it's like saying you fit 1000 gallons of water in a 1 gallon milk bottle. Can't be done. The model may be able to quote small snippets of works, just like you might remember various quotes from Shakespeare or someone. (That's not infringing either, by the way)

Wikipedia, and LLMs, can refuse to cite sources and still not infringe copyright. Citation simply isn't relevant to copyright. Not citing a work that you read previously is not a copyright infringement. Wikipedia or OpenAI being non-profit, or for-profit businesses, has nothing to do with copyright. Consent has nothing to do with copyright. Copyrighting something doesn't mean that you can require everyone who reads it to get your consent. You can require everyone who distributes it to get your consent, but once it's been distributed to someone, they can read it freely.

Hope this helps!