Comment by Nicook

7 days ago

I remember when google was scanning a bunch of rare books, I mean they might still be doing that? Either way, that was cool.

I have a few "rare books" and have read many, you'd be suprised at what is publicly available on google books since like ~2010ish.

Years ago I saw an article on this. For Google Books, they had two processes.

The first was destructive. This was for mainstream books currently being published so they had no value. It's (I believe) where you cut off the spine and scan the pages.

For rarer books, there was a non-destructive process. Basically the book was opened to each page and scanned. This was slower but didn't destroy the book.

I don't understand why these companies haven't just licensed the scans Google has already done. Why is each company doing this rather than just scanning the books once and sharing the scans?

  • > I don't understand why these companies haven't just licensed the scans Google has already done.

    Because for the vast majority of the books, Google doesn't have the legal right to license those scans. It would be legal if the books were out of copyright, but despite the connotation of "rare books" those generally aren't the books we're talking about. Further, in the cases where Google didn't destroy the original of the book they scanned, their scanned copy may be considered infringing under the new standard, so Google doesn't want call undue attention to what they have.

    • yeah, I think most of this really stems from people streatching the use of "rare books". The cool rare books google scanned one page at a time were borrowed from a library because they are rare and or valuable and usually old (out of copyright by a long time).

  • Here’s a clip of The Internet Archive’s nondestructive process [39 seconds]:

    https://nitter.net/internetarchive/status/135809098218971955...

      6 Feb 2021
    
      At the Internet Archive, this is how we digitize a book. 
    
      We never destroy a book by cutting off its binding. Instead, we digitize it the hard way--one page at a time.
    

    Wonder the additional cost to add automated page turning. More than Bezos can afford, impoverished chap he is.

  • If I remember correctly, it wasn't a distinction between common and rare books.

    It was a distinction between library books and others. They partnered with libraries, and obviously, libraries didn't want their books destroyed, so they devised a non-destructive book scanner. Some of the books they scanned were indeed rare, but rare or not, you don't destroy books you borrowed from a library!