← Back to context

Comment by ThrowawayTestr

2 years ago

Why don't these companies spend a bunch of money digitizing old texts? There's bound to be tons of good training data that isn't online.

That's what Google Books was and then it all got mothballed for copyright reasons. The future of AI belongs to societies that DGAF about IP.

  • In that case, big corps will eat all the free data and than buy a regulatory lock in. I hate copyright, but in case if ai it is not a bad thing.