# PDF / .DOC =\> CSV parsing conversion - cheaply, efficiently, accurately

**URL:** <https://boards.straightdope.com/t/pdf-doc-csv-parsing-conversion-cheaply-efficiently-accurately/584258>\
**Category:** Marketplace\
**Created:** [June 3, 2011, 11:12pm UTC](https://boards.straightdope.com/t/pdf-doc-csv-parsing-conversion-cheaply-efficiently-accurately/584258 "2011-06-03T23:12:16Z")\
**Posts on this page:** 2\
**Page:** 1

<div class="post-metadata">

**Author:** ![code\_grey](https://avatars.discourse-cdn.com/v4/letter/c/f4b2a3/32.png) [@code\_grey](https://boards.straightdope.com/u/code_grey)\
**Post date:** [June 3, 2011, 11:12pm UTC](https://boards.straightdope.com/t/pdf-doc-csv-parsing-conversion-cheaply-efficiently-accurately/584258/1 "2011-06-03T23:12:16Z")

</div>

Parse your PDF and DOC format documents into fields and tables for cheap. You know you want to 😉

I am offering automated outsourced extraction of text fields and table contents from text based documents, e.g. the filing forms from financial industry, healthcare related paperwork and similar. Thanks to my nifty proprietary parsing framework I can make parsers for new document formats fairly quickly, and ambiguity problems associated with missing fields in tables are minimized.

I will only work with not-very-confidential documents because I don’t currently distribute my framework. So if your stuff is sufficiently confidential to require you to run the parser on your machine instead of transmitting it to me for parsing, unfortunately I cannot be of service.

---

<div class="post-metadata">

**Author:** ![code\_grey](https://avatars.discourse-cdn.com/v4/letter/c/f4b2a3/32.png) [@code\_grey](https://boards.straightdope.com/u/code_grey)\
**Post date:** [June 3, 2011, 11:38pm UTC](https://boards.straightdope.com/t/pdf-doc-csv-parsing-conversion-cheaply-efficiently-accurately/584258/2 "2011-06-03T23:38:42Z")

</div>

I can be reached at parseyourpdfs (gmail)

**We also herd cats and find clever solutions to other complex problems**
