Internet Archive Seeks Material for Library of Amateur Radio and Communications

4 years ago (blog.archive.org)

Internet Archive has begun gathering content for the Digital Library of Amateur Radio and Communications (DLARC), which will be a massive online library of materials and collections related to amateur radio and early digital communications.

The DLARC project is looking for contributors with troves of ham radio, amateur radio, and early digital communications related books, magazines, documents, catalogs, manuals, videos, software, personal archives, and other historical records collections, no matter how big or small. In addition to physical material to digitize, we are looking for podcasts, newsletters, video channels, and other digital content that can enrich the DLARC collections.

Probably totally digress, but I wish IA can organize their digital library slightly better.

One day I was checking some manga books by ISBN on IA just out of curiosity. And for some reason, it put the ISBNs for all the volumes of a manga into one single entry (https://archive.org/details/isbn_1919979003907, check "ISBN" metadata section) and unsurprisingly, the actual content is only one volume, vol.43 (not even vol.1!). I have a feeling other volumes may exist somewhere there, but there is no way to search for them.

This isn't a one-off occurrence either, it reflects my experience for trying to find specific item there well, especially for non-English books.

  • A lot of the time the metadata accuracy is up to the original uploader. IA's upload system doesn't magically fill in all the metadata details for an item.

    • Also doesn't allow other to update metadata or even submit for review.

      Wikidata has a property for Internet Archive ID, so it wouldn't be conceptually hard to construct a parallel metadata store there, but it would involve hundreds of millions of triples so it's definitely "hard" in other senses.

  • While I also wish the Archive to be more precise - e.g. in the "Author" and in the "Year of publication" fields -,

    I suggest that you check their RSS feeds to see how staggeringly high the rate of uploads is. That uploading is "frenetic" (in a good way of course) reveals where the focus is. For re-assessing and fixing the records a parallel team would probably be needed.

    I would gladly help towards that: I never checked but maybe one can volunteer.

  • I agree. I had wondered how successful and easy it would be to create a "front end" site that does a better job of searching, organizing archive.org.

There are millions of "reflector" messages that contain a tremendous amount of knowledge. I hope the project manages to archive those as well.

I have a shelf full of books (already got rid of all my QSTs I can access digitally) I will be able to get rid of soon!

Ahh, this is great. I was already seriously impressed by the amount of amateur radio content on the Internet Archive. I'm happily surprised to see some solicitation for even more content! Passing this along to my relevant communities/clubs/etc. (also just emailed one possible place to archive :))

It would be awesome if one day we go back and hear all these small town and college radio stations and the types of shows in them.

Woukd be a cool time capsule.

The IA has really lost its reputation as an archive by choosing to remove content down for political purposes. They see themselves as publishers rather than an archive.

  • I attempted to google to guess what you meant.

    I found: (1) they recently removed KiwiFarms; (2) in 2020, they began labeling certain pages with "fact checks"; (3) they remove content by request of the site owner or by copyright complaint.

    Of those, (2) seems the most political, but it's not removing content. Was there something else you had in mind?

  • What did they remove?

    • Evidence of unsavoury and most likely illegal behavior of a man with powerful friends in the tech industry.

      If you know the right people with the right political views, even the IA will capitulate.

As far as I'm concerned the IA can stick this up their arse. They REFUSE to acknowlege my request to have a personal website that they've managed to archive removed. It's a person blog that I want online so I can give the URL to family and friends, but I keep it out of Google etc with robots.txt. But once during an upgrade of the backend software I stuffed up the robots.txt and they crawled it until I fixed it up.

Will they reply to my emails? They will not. I'm so frustrated, they just IGNORE emails.

Don't support these clowns.

  • You picked the dumbest possible method to share something privately with family and friends but it's the Internet Archive's issue to fix? A robots.txt provides zero protection for a public website.

    • I didn't say it was private. It's obviously public, I give you the URL and you can pull it up. I'm not stupid enough to think that robots.txt means no one can read it! There's some RBAC control so that you need to be logged in etc if I publish anything I wish to remain private. Most _reputable_ organisiations will respect robots.txt and not crawl your site, nor cache/archive your data if you request it. What I want is the decision that someone's going to archive/store my website somewhere else to be MY decision, not theirs.

  • > they just ignore

    I have hints that they may be understaffed. If you had a little spare time, you could lend a hand to their legendary effort, and maybe contribute in fixing a few things such as your issue.

    • people who are politically connected/aligned to the organization seem to have no problems getting things memory holed

  • I suggest you publishing something againts their political agenda, and they will remove and censor it asap