# Old newspaper archives online?

**URL:** <https://boards.straightdope.com/t/old-newspaper-archives-online/358806>\
**Category:** Factual Questions\
**Created:** [May 30, 2006, 11:56pm UTC](https://boards.straightdope.com/t/old-newspaper-archives-online/358806 "2006-05-30T23:56:53Z")\
**Posts on this page:** 17\
**Page:** 1

<div class="post-metadata">

**Author:** ![II\_Gyan\_II](https://avatars.discourse-cdn.com/v4/letter/i/bbe5ce/32.png) [@II\_Gyan\_II](https://boards.straightdope.com/u/II_Gyan_II)\
**Post date:** [May 30, 2006, 11:56pm UTC](https://boards.straightdope.com/t/old-newspaper-archives-online/358806/1 "2006-05-30T23:56:53Z")

</div>

Are there searchable (or atleast indexed) online databases of the content of old newspapers (starting around mid or late 1800s; mainly prominent American & British ones)?

---

<div class="post-metadata">

**Author:** ![MsRobyn](https://avatars.discourse-cdn.com/v4/letter/m/90ced4/32.png) [@MsRobyn](https://boards.straightdope.com/u/MsRobyn)\
**Post date:** [May 31, 2006, 12:50am UTC](https://boards.straightdope.com/t/old-newspaper-archives-online/358806/2 "2006-05-31T00:50:12Z")

</div>

There’s always Proquest. It’s a subscription service, and IIRC, it costs boucoup bux. Many libraries have subscriptions, though, so you can always try your local public or university library.

Robin

---

<div class="post-metadata">

**Author:** ![samclem](https://avatars.discourse-cdn.com/v4/letter/s/a9a28c/32.png) [@samclem](https://boards.straightdope.com/u/samclem)\
**Post date:** [May 31, 2006, 2:20am UTC](https://boards.straightdope.com/t/old-newspaper-archives-online/358806/3 "2006-05-31T02:20:02Z")

</div>

If you have just a simple search or two, I’d be glad to do it. I subscribe.

---

<div class="post-metadata">

**Author:** ![II\_Gyan\_II](https://avatars.discourse-cdn.com/v4/letter/i/bbe5ce/32.png) [@II\_Gyan\_II](https://boards.straightdope.com/u/II_Gyan_II)\
**Post date:** [May 31, 2006, 2:42am UTC](https://boards.straightdope.com/t/old-newspaper-archives-online/358806/4 "2006-05-31T02:42:25Z")

</div>

**MsRobyn** , was Proquest Newsstand the service you were referring to? If so, it seems the coverage starts [from](http://www.proquest.com/products_pq/descriptions/newsstand.shtml) 1980. Is that right?

Thanks for the offer, **samclem** , but I’ll want to dig in.

---

<div class="post-metadata">

**Author:** ![jiggs](https://avatars.discourse-cdn.com/v4/letter/j/ecc23a/32.png) [@jiggs](https://boards.straightdope.com/u/jiggs)\
**Post date:** [May 31, 2006, 3:24am UTC](https://boards.straightdope.com/t/old-newspaper-archives-online/358806/5 "2006-05-31T03:24:12Z")

</div>

Try

[http://www.newspaperarchive.com](http://www.newspaperarchive.com)

Their subscriptions are much more reasonable, and they have a huge quantity of searchable newspapers online. If you have to use Proquest (they have the biggies like the NY Times, Wall Street Journal) you’ll find that most libraries subscribe so you can use the service for free. It’s pretty pricey to get a home subscription to Proquest.

---

<div class="post-metadata">

**Author:** ![jiggs](https://avatars.discourse-cdn.com/v4/letter/j/ecc23a/32.png) [@jiggs](https://boards.straightdope.com/u/jiggs)\
**Post date:** [May 31, 2006, 3:28am UTC](https://boards.straightdope.com/t/old-newspaper-archives-online/358806/6 "2006-05-31T03:28:19Z")

</div>

Oh, and Proquest does not just have 1980+ papers archived. The amount archived varies by paper, and they have archives going back into the 1800s on some papers (as does [newspaperarchive.com](http://newspaperarchive.com)). I think some Proquest subscriptions only let you search 1980+, and the more expensive version gives you the whole range of their archive.

---

<div class="post-metadata">

**Author:** ![samclem](https://avatars.discourse-cdn.com/v4/letter/s/a9a28c/32.png) [@samclem](https://boards.straightdope.com/u/samclem)\
**Post date:** [May 31, 2006, 3:45am UTC](https://boards.straightdope.com/t/old-newspaper-archives-online/358806/7 "2006-05-31T03:45:04Z")

</div>

> [@II Gyan II](#):
>
> **MsRobyn** , was Proquest Newsstand the service you were referring to? If so, it seems the coverage starts [from](http://www.proquest.com/products_pq/descriptions/newsstand.shtml) 1980. Is that right?
> 
> Thanks for the offer, **samclem** , but I’ll want to dig in.

In a previous thread about online databases, I offered

> [@](#):
>
> The other database I pay to access is ProQuest Historical Newspapers, the version available through SABR. This costs $60. US/yr. and is the bargain of the century. The list of papers available by joining them is  
> Quote:  
> American Periodical Series Online (1749-1900)  
> The New York Times (1851-2001)  
> The Boston Globe, (1872-1922; not all years are yet available)  
> The Washington Post (1877-1988)  
> The Los Angeles Times (1881-1964)  
> The Chicago Tribune (1872-1964; not all years are yet available)  
> The Atlanta Constitution (1868-1925; not all years are yet available)  
> from [http://www.sabr.org/sabr.cfm?a=cms,c,861,35,0](http://www.sabr.org/sabr.cfm?a=cms,c,861,35,0)

from [http://boards.straightdope.com/sdmb/showthread.php?t=357993&highlight=proquest+newspaperarchive](http://boards.straightdope.com/sdmb/showthread.php?t=357993&highlight=proquest+newspaperarchive)

If you want to be able to access all the papers in the comfort of your home, doing it on YOUR schedule, note that it only costs $60/year. This is an incredible bargain. Incredible.

---

<div class="post-metadata">

**Author:** ![MsRobyn](https://avatars.discourse-cdn.com/v4/letter/m/90ced4/32.png) [@MsRobyn](https://boards.straightdope.com/u/MsRobyn)\
**Post date:** [May 31, 2006, 10:34am UTC](https://boards.straightdope.com/t/old-newspaper-archives-online/358806/8 "2006-05-31T10:34:47Z")

</div>

I’m still in school, so I have access to a bunch of databases, Proquest included. I was not aware of the [sabr.org](http://sabr.org) site.

Robin

---

<div class="post-metadata">

**Author:** ![anson2995](https://avatars.discourse-cdn.com/v4/letter/a/c77e96/32.png) [@anson2995](https://boards.straightdope.com/u/anson2995)\
**Post date:** [May 31, 2006, 2:02pm UTC](https://boards.straightdope.com/t/old-newspaper-archives-online/358806/9 "2006-05-31T14:02:21Z")

</div>

In the previous thread that \*\*samclem \*\* mentioned, I suggested [paperofrecord.com](http://paperofrecord.com), which is the only electronic source for the Sporting News. The rest of the papers they have are obscure canadian weeklies, but TSN is invaluable for sports researchers.

Proquest is awesome, perhaps the coolest innovation in human history.

---

<div class="post-metadata">

**Author:** ![Zsofia](https://avatars.discourse-cdn.com/v4/letter/z/7bcc69/32.png) [@Zsofia](https://boards.straightdope.com/u/Zsofia)\
**Post date:** [May 31, 2006, 2:22pm UTC](https://boards.straightdope.com/t/old-newspaper-archives-online/358806/10 "2006-05-31T14:22:40Z")

</div>

Most public libraries pay for a lot of database subscriptions that their patrons have no idea they can access. It’s really frustrating sometimes. Actually, a lot of _state_ libraries pay for _all state residents_ to have access to certain databases - here in South Carolina we have DISCUS, for example, and Georgia has GALILEO. Don’t pay for a database yourself - call your library first! Most people are really shocked to see what we have to offer that they can access from home with just a library card!

---

<div class="post-metadata">

**Author:** ![anson2995](https://avatars.discourse-cdn.com/v4/letter/a/c77e96/32.png) [@anson2995](https://boards.straightdope.com/u/anson2995)\
**Post date:** [June 9, 2006, 5:35pm UTC](https://boards.straightdope.com/t/old-newspaper-archives-online/358806/11 "2006-06-09T17:35:39Z")

</div>

> [@samclem](#):
>
> The other database I pay to access is ProQuest Historical Newspapers, the version available through SABR. This costs $60. US/yr. and is the bargain of the century.

SABR members received an email today stating that their access to ProQuest historical newspaper collection would be discontinued at the end of the year. It said, in part:

> [@](#):
>
> Recently ProQuest sent SABR a letter explaining a new business decision to no longer offer remote access to genealogical or historical societies, beginning at the end of the current subscription year. For SABR, that means beginning on January 1, 2007, SABR members will no longer have remote access to the various historical newspapers to which SABR subscribes, nor will we have access to the HeritageQuest databases.

---

<div class="post-metadata">

**Author:** ![samclem](https://avatars.discourse-cdn.com/v4/letter/s/a9a28c/32.png) [@samclem](https://boards.straightdope.com/u/samclem)\
**Post date:** [June 9, 2006, 8:09pm UTC](https://boards.straightdope.com/t/old-newspaper-archives-online/358806/12 "2006-06-09T20:09:45Z")

</div>

> [@anson2995](#):
>
> SABR members received an email today stating that their access to ProQuest historical newspaper collection would be discontinued at the end of the year. It said, in part:

Yeah, I’m bummed. It is due to ProQuest’s financial problems, self-inflicted. They may not survive.  
[http://www.infotoday.com/it/jun06/Britt.shtml](http://www.infotoday.com/it/jun06/Britt.shtml)

I may have to actually get out in the sunshine starting next year. 🙂

---

<div class="post-metadata">

**Author:** ![II\_Gyan\_II](https://avatars.discourse-cdn.com/v4/letter/i/bbe5ce/32.png) [@II\_Gyan\_II](https://boards.straightdope.com/u/II_Gyan_II)\
**Post date:** [June 9, 2006, 9:17pm UTC](https://boards.straightdope.com/t/old-newspaper-archives-online/358806/13 "2006-06-09T21:17:55Z")

</div>

I tried out the Proquest service at one of my uni libraries. I am disappointed that I can’t locate the keyword within articles, i.e. if I search for ‘chicago’, it throws up a lot of results, but they are all PDFed images. Often these articles are long, and I have to read the whole thing to figure out where the word occurs. Is there some tool/trick I am missing?

---

<div class="post-metadata">

**Author:** ![Exapno\_Mapcase](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/exapno_mapcase/32/1051_2.png) [@Exapno\_Mapcase](https://boards.straightdope.com/u/Exapno_Mapcase)\
**Post date:** [June 10, 2006, 2:45pm UTC](https://boards.straightdope.com/t/old-newspaper-archives-online/358806/14 "2006-06-10T14:45:55Z")

</div>

> [@II Gyan II](#):
>
> I tried out the Proquest service at one of my uni libraries. I am disappointed that I can’t locate the keyword within articles, i.e. if I search for ‘chicago’, it throws up a lot of results, but they are all PDFed images. Often these articles are long, and I have to read the whole thing to figure out where the word occurs. Is there some tool/trick I am missing?

You can search inside a .pdf document. It works like “find” inside a word document or web file. Adobe uses a binocular icon to launch the search.

I haven’t yet used ProSearch, so it may have some weird non-standard version of Abode that doesn’t show this, but it should be in the file menus somewhere.

---

<div class="post-metadata">

**Author:** ![Dewey\_Finn](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/dewey_finn/32/4222_2.png) [@Dewey\_Finn](https://boards.straightdope.com/u/Dewey_Finn)\
**Post date:** [June 10, 2006, 3:14pm UTC](https://boards.straightdope.com/t/old-newspaper-archives-online/358806/15 "2006-06-10T15:14:11Z")

</div>

The PDF files that **II Gyan II** aren’t searchable because, as he said, they are images, not text files. That’s what results from scanning old newspapers. (It’s possible to scan the pages and then run OCR software on them to get searchable text documents, but the character recognition software is imperfect at best, and on old newspapers will be really inaccurate.)

---

<div class="post-metadata">

**Author:** ![Exapno\_Mapcase](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/exapno_mapcase/32/1051_2.png) [@Exapno\_Mapcase](https://boards.straightdope.com/u/Exapno_Mapcase)\
**Post date:** [June 10, 2006, 3:18pm UTC](https://boards.straightdope.com/t/old-newspaper-archives-online/358806/16 "2006-06-10T15:18:19Z")

</div>

> [@Dewey Finn](#):
>
> The PDF files that **II Gyan II** aren’t searchable because, as he said, they are images, not text files. That’s what results from scanning old newspapers. (It’s possible to scan the pages and then run OCR software on them to get searchable text documents, but the character recognition software is imperfect at best, and on old newspapers will be really inaccurate.)

[NewspaperArchives.com](http://NewspaperArchives.com) does this and any search gives back your term in 99% gibberish. But you can re-search the .pdf’s and hone in on your term. I hadn’t realized that ProSearch didn’t include OCR recognition.

---

<div class="post-metadata">

**Author:** ![anson2995](https://avatars.discourse-cdn.com/v4/letter/a/c77e96/32.png) [@anson2995](https://boards.straightdope.com/u/anson2995)\
**Post date:** [June 10, 2006, 4:07pm UTC](https://boards.straightdope.com/t/old-newspaper-archives-online/358806/17 "2006-06-10T16:07:36Z")

</div>

> [@Exapno Mapcase](#):
>
> I hadn’t realized that ProSearch didn’t include OCR recognition.

ProQuest does, use OCR technology to index the pages, at least for the major papers that I use most often. That’s the best thing about it, in my mind. I can search on a book title or an author’s name to find a book review. Without it, I have to know when and where the review was published. In doing sports research, I can search on a player’s name to find out when he did something in a game worthy of being mentioned, rather than scanning stories every day for an entire season.
