# Can all information be found online?

**URL:** <https://boards.straightdope.com/t/can-all-information-be-found-online/802884>\
**Category:** Factual Questions\
**Created:** [November 27, 2017, 6:50pm UTC](https://boards.straightdope.com/t/can-all-information-be-found-online/802884 "2017-11-27T18:50:31Z")\
**Posts on this page:** 20\
**Page:** 3

<div class="post-metadata">

**Author:** ![Chronos](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/chronos/32/134_2.png) [@Chronos](https://boards.straightdope.com/u/Chronos)\
**Post date:** [November 28, 2017, 6:33pm UTC](https://boards.straightdope.com/t/can-all-information-be-found-online/802884/41 "2017-11-28T18:33:11Z")

</div>

A hash still doesn’t add any meaningful security for a four-digit PIN, because it’s trivial to just try all 10,000 PINs until you find the one that matches the hash. Hashes are only relevant when you’re hashing passwords, which (due to greater length and larger alphabet) can be much harder to search exhaustively. Well, at least if the password was chosen well, which most of them aren’t, but security is at least possible.

And even if data is digitized, it might still never make it online. When I was an undergrad, the astronomy department had huge racks of paper tape of old data, that had never been transferred over to a new medium, and the machines built to read it were all scrapped. Every so often, a student would start some effort to get those read by improvising a scanner to do the job, or the like, but (being undergrads) never got the chance to finish the job. Eventually, all those racks of data just got thrown out. This sort of thing doesn’t happen any more, now that people realize that keeping media updated is something that needs to be done, but it was pretty common back in the day.

---

<div class="post-metadata">

**Author:** ![Yllaria](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/yllaria/32/3452_2.png) [@Yllaria](https://boards.straightdope.com/u/Yllaria)\
**Post date:** [November 28, 2017, 7:03pm UTC](https://boards.straightdope.com/t/can-all-information-be-found-online/802884/42 "2017-11-28T19:03:56Z")

</div>

> [@beowulff](#):
>
> There’s vast amounts of information that isn’t online, but even more information that is on-line but nearly impossible to retrieve, because it is so poorly indexed.

For several years my grandfather worked as an ice man. He worked at an ice manufacturing plant in San Pedro CA, which delivered blocks of ice to residents. When my mother died I inherited a couple of regional industry newsletters. I enjoyed reading the articles talking earnest trash about the perils and inefficiencies of refrigeration. I posted scans on a blog. I’ve never been able to google a link to them directly. I can google to my blog, because I know its name, and then search there, but googling directly to the newsletters is effectively impossible.

It doesn’t help that the name of the newsletter is “Ice Picks”.

---

<div class="post-metadata">

**Author:** ![Bullitt](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/bullitt/32/5725_2.png) [@Bullitt](https://boards.straightdope.com/u/Bullitt)\
**Post date:** [November 28, 2017, 7:18pm UTC](https://boards.straightdope.com/t/can-all-information-be-found-online/802884/43 "2017-11-28T19:18:57Z")

</div>

> [@Ludovic](#):
>
> _Everyone’s_ PIN is online assuming they use a 4-digit one. There’s just no telling which one is yours 🙂

I have all of them! 🙂

---

<div class="post-metadata">

**Author:** ![erysichthon](https://avatars.discourse-cdn.com/v4/letter/e/d07c76/32.png) [@erysichthon](https://boards.straightdope.com/u/erysichthon)\
**Post date:** [November 28, 2017, 7:21pm UTC](https://boards.straightdope.com/t/can-all-information-be-found-online/802884/44 "2017-11-28T19:21:33Z")

</div>

I was a librarian for 30 years (retired now). From the mid-2000s on, I had lots of conversations that went like this:

Person: “What’s it like being a librarian now that Google\* has digitized everything?”  
Me: “Google hasn’t digitized everything.”  
Person: “Yes, they have. I read it somewhere. Have you started looking for a new career yet?”

\*Sometimes it was the Library of Congress instead of Google.

Oh, and don’t get me started on the people who claim to have a USB drive containing “every song ever written.”

---

<div class="post-metadata">

**Author:** ![Lucas\_Jackson](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/lucas_jackson/32/303_2.png) [@Lucas\_Jackson](https://boards.straightdope.com/u/Lucas_Jackson)\
**Post date:** [November 28, 2017, 7:44pm UTC](https://boards.straightdope.com/t/can-all-information-be-found-online/802884/45 "2017-11-28T19:44:06Z")

</div>

I was talking with an old friend of mine from Texas last night about Texas music. He asked me if I remembered the song, “Blah Blah Blah”. I’d heard it before so this morning I tried googling it. I found this thread: [http://boards.straightdope.com/sdmb/showthread.php?t=588560](http://boards.straightdope.com/sdmb/showthread.php?t=588560) from 2010. I tried find a way to buy it online but didn’t have any more luck that the OP. So the internet doesn’t have that…

ETA: It’s this one: [Blah, Blah, Blah Chords - Sam Oaks - Cowboy Lyrics](https://www.cowboylyrics.com/tabs/sam-oaks/blah-blah-blah-27335.html)

Other searcher: [Yahoo | Mail, Weather, Search, Politics, News, Finance, Sports & Videos](https://answers.yahoo.com/question/index?qid=20060620231445AAYmOK6)

---

<div class="post-metadata">

**Author:** ![Chronos](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/chronos/32/134_2.png) [@Chronos](https://boards.straightdope.com/u/Chronos)\
**Post date:** [November 28, 2017, 7:54pm UTC](https://boards.straightdope.com/t/can-all-information-be-found-online/802884/46 "2017-11-28T19:54:08Z")

</div>

Google has certainly made it a goal of theirs to digitize everything and get it all online. And they’ve made quite impressive progress. But they’re still not remotely close to done, and won’t be for a very long time, even if you only count public and semi-public information like books and newsletters. And there will always be some material, like most love letters, which will never be digitized.

---

<div class="post-metadata">

**Author:** ![CalMeacham](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/calmeacham/32/35_2.png) [@CalMeacham](https://boards.straightdope.com/u/CalMeacham)\
**Post date:** [November 28, 2017, 8:05pm UTC](https://boards.straightdope.com/t/can-all-information-be-found-online/802884/47 "2017-11-28T20:05:40Z")

</div>

> [@Chronos](#):
>
> Google has certainly made it a goal of theirs to digitize everything and get it all online. And they’ve made quite impressive progress. But they’re still not remotely close to done, and won’t be for a very long time, even if you only count public and semi-public information like books and newsletters. And there will always be some material, like most love letters, which will never be digitized.

They may have made it their goal to digitize everything, but they certainly have not made it their goal to make it all available online. I frequently run into situations where they certainly have the book digitized, but only offer it in a “snippet” view, or when, after going through a few pages, I find a block that says that the rest isn’t available.

I’m not just talking about recent things still under copyright. I frequently am frustrated by old, long out-of-copyright books that GoogleBooks witholds from me, although my search engine (usually Google itself) tantalizingly tells me that they have something with one of my keywords. It makes me want to strangle them.

---

<div class="post-metadata">

**Author:** ![Darren\_Garrison](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/darren_garrison/32/92_2.png) [@Darren\_Garrison](https://boards.straightdope.com/u/Darren_Garrison)\
**Post date:** [November 28, 2017, 8:41pm UTC](https://boards.straightdope.com/t/can-all-information-be-found-online/802884/48 "2017-11-28T20:41:02Z")

</div>

> [@Yllaria](#):
>
> For several years my grandfather worked as an ice man. He worked at an ice manufacturing plant in San Pedro CA, which delivered blocks of ice to residents. When my mother died I inherited a couple of regional industry newsletters. I enjoyed reading the articles talking earnest trash about the perils and inefficiencies of refrigeration. I posted scans on a blog. I’ve never been able to google a link to them directly. I can google to my blog, because I know its name, and then search there, but googling directly to the newsletters is effectively impossible.
> 
> It doesn’t help that the name of the newsletter is “Ice Picks”.

You should put them on [archive.org](http://archive.org).

---

<div class="post-metadata">

**Author:** ![Two\_Many\_Cats](https://avatars.discourse-cdn.com/v4/letter/t/8dc957/32.png) [@Two\_Many\_Cats](https://boards.straightdope.com/u/Two_Many_Cats)\
**Post date:** [November 28, 2017, 11:15pm UTC](https://boards.straightdope.com/t/can-all-information-be-found-online/802884/49 "2017-11-28T23:15:38Z")

</div>

Is the OP kidding? All information is not even close to being posted online. Puh-leese. Put the laptop down and get with the real world.

---

<div class="post-metadata">

**Author:** ![DPRK](https://avatars.discourse-cdn.com/v4/letter/d/4491bb/32.png) [@DPRK](https://boards.straightdope.com/u/DPRK)\
**Post date:** [November 28, 2017, 11:43pm UTC](https://boards.straightdope.com/t/can-all-information-be-found-online/802884/50 "2017-11-28T23:43:46Z")

</div>

> [@Darren\_Garrison](#):
>
> You should put them on [archive.org](http://archive.org).

Seconded.

And if you are a librarian, curator, or otherwise in charge of a valuable collection, and a corporation like Google comes to you with a generous offer to come in, ransack your stacks and digitise everything for free, please make triple-sure, in writing, that the license terms require them to make all derived materials openly available to you, your patrons, and the general public in perpetuity.

---

<div class="post-metadata">

**Author:** ![sunstone](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/sunstone/32/3372_2.png) [@sunstone](https://boards.straightdope.com/u/sunstone)\
**Post date:** [November 29, 2017, 1:10am UTC](https://boards.straightdope.com/t/can-all-information-be-found-online/802884/51 "2017-11-29T01:10:52Z")

</div>

Not everything is online. Some of the research papers that I’ve published are online, but others are not. This seems to be more a factor of where and when they were published rather than the subject material.

Some professional literature in the sciences have been digitized and others have not.

---

<div class="post-metadata">

**Author:** ![Paul\_in\_Qatar](https://avatars.discourse-cdn.com/v4/letter/p/ccd318/32.png) [@Paul\_in\_Qatar](https://boards.straightdope.com/u/Paul_in_Qatar)\
**Post date:** [November 29, 2017, 2:27am UTC](https://boards.straightdope.com/t/can-all-information-be-found-online/802884/52 "2017-11-29T02:27:46Z")

</div>

Lots of stuff is kept offline because it is secret or because nobody has ever bothered to digitize it. But even in this day and age, there are things that are not hopelessly obscure that seem to have no internet presence.

Ike was the president of Columbia University after the war. I have long maintained there are no images of the interior of the president’s residence on the internet. (I am really looking for the interior during Ike’s tenure.) No reason it is not online. It just isn’t.

---

<div class="post-metadata">

**Author:** ![Projammer](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/projammer/32/559_2.png) [@Projammer](https://boards.straightdope.com/u/Projammer)\
**Post date:** [November 29, 2017, 8:12am UTC](https://boards.straightdope.com/t/can-all-information-be-found-online/802884/53 "2017-11-29T08:12:16Z")

</div>

There is no central database to search firearm serial numbers or ownership. Many of the NICS forms are available online to the BATF in PDF form, but they are not digitally searchable.

---

<div class="post-metadata">

**Author:** ![Francis\_Vaughan](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/francis_vaughan/32/3093_2.png) [@Francis\_Vaughan](https://boards.straightdope.com/u/Francis_Vaughan)\
**Post date:** [November 29, 2017, 11:24am UTC](https://boards.straightdope.com/t/can-all-information-be-found-online/802884/54 "2017-11-29T11:24:06Z")

</div>

Just to add a pet peeve. I’m something of an Apollo missions tragic. I love to read the old reports and papers surrounding the programme. NASA published a huge volume of work, detailing pretty much everything you want to know. Most of it only resides on paper. There are scans of some documents available. NASA itself has a web page that provides links to many. And many of these links are to scans made by private individuals that took the time to go to the libraries where the now aged documents reside. Many of these scans are pretty mediocre, often with essentially impossible to read figures and barely discernable pictures.  
Not just these documents, but the blueprints to most of the systems reside on microfiche. Again essentially impossible to access. With the 50th anniversary looming there is going to be another lift in interest in the moon landings. It would be brilliant to see a lot more of the core information made available, but there is also no doubt, it isn’t a small undertaking to take a small library’s worth of publications and to curate a digital on-line version.

Rather than Google, maybe someone could convince Elon or Jeff to fund it.

But this is just my little obsession. The world is filled with such repositories. What is important is that much of the information in these repos is high quality. It isn’t the masses of barely curated detritus of modern life.

---

<div class="post-metadata">

**Author:** ![DPRK](https://avatars.discourse-cdn.com/v4/letter/d/4491bb/32.png) [@DPRK](https://boards.straightdope.com/u/DPRK)\
**Post date:** [November 29, 2017, 11:35am UTC](https://boards.straightdope.com/t/can-all-information-be-found-online/802884/55 "2017-11-29T11:35:43Z")

</div>

[archive.org](http://archive.org) will digitise microfilm/microfiche as well as books. I am sure they would be happy if some billionaire dropped extra money on them.

---

<div class="post-metadata">

**Author:** ![Darren\_Garrison](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/darren_garrison/32/92_2.png) [@Darren\_Garrison](https://boards.straightdope.com/u/Darren_Garrison)\
**Post date:** [November 29, 2017, 11:44am UTC](https://boards.straightdope.com/t/can-all-information-be-found-online/802884/56 "2017-11-29T11:44:44Z")

</div>

> [@DPRK](#):
>
> And if you are a librarian, curator, or otherwise in charge of a valuable collection, and a corporation like Google comes to you with a generous offer to come in, ransack your stacks and digitise everything for free, please make triple-sure, in writing, that the license terms require them to make all derived materials openly available to you, your patrons, and the general public in perpetuity.

Don’t blame Google for all that information not being available to the public–blame the Writer’s Guild of America for throwing a hissy fit to stop them.

---

<div class="post-metadata">

**Author:** ![DPRK](https://avatars.discourse-cdn.com/v4/letter/d/4491bb/32.png) [@DPRK](https://boards.straightdope.com/u/DPRK)\
**Post date:** [November 29, 2017, 11:51am UTC](https://boards.straightdope.com/t/can-all-information-be-found-online/802884/57 "2017-11-29T11:51:29Z")

</div>

My apologies to Google if they tried to do a good thing but got their arm twisted through no fault of their own. Is there anything they could have done in advance to avoid such problems, or anything they can do now to resolve them?

---

<div class="post-metadata">

**Author:** ![Yllaria](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/yllaria/32/3452_2.png) [@Yllaria](https://boards.straightdope.com/u/Yllaria)\
**Post date:** [November 30, 2017, 9:35pm UTC](https://boards.straightdope.com/t/can-all-information-be-found-online/802884/58 "2017-11-30T21:35:12Z")

</div>

> [@Darren\_Garrison](#):
>
> You should put them on [archive.org](http://archive.org).

Thanks. I’ll check them out. If they’re interested in PDFs of just two newsletters, that would be cool.

---

<div class="post-metadata">

**Author:** ![Darren\_Garrison](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/darren_garrison/32/92_2.png) [@Darren\_Garrison](https://boards.straightdope.com/u/Darren_Garrison)\
**Post date:** [November 30, 2017, 9:42pm UTC](https://boards.straightdope.com/t/can-all-information-be-found-online/802884/59 "2017-11-30T21:42:45Z")

</div>

> [@Yllaria](#):
>
> Thanks. I’ll check them out. If they’re interested in PDFs of just two newsletters, that would be cool.

No interest or permission needed, you can just post it.

---

<div class="post-metadata">

**Author:** ![SamuelA](https://avatars.discourse-cdn.com/v4/letter/s/c77e96/32.png) [@SamuelA](https://boards.straightdope.com/u/SamuelA)\
**Post date:** [November 30, 2017, 10:32pm UTC](https://boards.straightdope.com/t/can-all-information-be-found-online/802884/60 "2017-11-30T22:32:04Z")

</div>

> [@DPRK](#):
>
> My apologies to Google if they tried to do a good thing but got their arm twisted through no fault of their own. Is there anything they could have done in advance to avoid such problems, or anything they can do now to resolve them?

If you read the articles on wired’s website about this topic, it comes down to a Federal judge blocking the agreement between the writer’s guild and google. I do hope this problem can be solved, because it would enable much better access to the information.

I wonder if google can use all their scanned books (a significant fraction of every book ever published in English) to train an artificial intelligence.

It would not be illegal for the AI to answer questions based on those books, in the same way it is not illegal for me to read a book and then tell you any facts I know from it and make short quotes from the book.

Of course, google was going to take it a step further. You would have been able to purchase, for a minimal cost, access to the full text of any book ever published from your computer.

[Previous page](https://boards.straightdope.com/t/can-all-information-be-found-online/802884.md?page=2)

[Next page](https://boards.straightdope.com/t/can-all-information-be-found-online/802884.md?page=4)
