See unedited process of how a programmer would go about parsing a PDF in Python 🐍 to compare one terms of service to another. I'll also be on the lookout for Ballmer's Peak
👍 Like and Subscribe
👪 Follow me on Twitter / matteohoch
📮 Join my mailing list https://www.ergosum.co/subscribe/
💬 Share Your Thoughts and Suggest New Content for me to do!
🗒️ View Blog Post https://www.ergosum.co/pdf-parsing-in...
https://github.com/rogerfitz/tutorial...
Happy to answer any questions. Thanks for watching!
Sections:
Introduction - 0:00
Finding a Library - 6:54
Using the Library and Creating the Algorithms - 13:09
Getting Upset About Copy and Paste Not Working - 30:25
Looking Over Github's Text Comparison Utility - 35:23
Linus Torvald is a Coding Hero - 36:08
Schrodinger is a weekend warrior - 38:38
Back to Coding - 41:42
EDGAR is Awesome for Researching Public Companies (and a great source of PDFs/scrapeable data) - 47:18
Back to Coding.. Again - with some thoughts on PDF nuances - 48:22
Loading the Data in a Pandas Dataframe - 55:22
Finding Differencing Libraries Like Git Does and Tying it All Back Together- 1:02:40