Is it easier to scrape websites or PDFs? What are the pros and cons of each?

18 Replies, 1576 Views

Wow, thanks for all the suggestions, everyone! I tried PyPDF2 and it worked pretty well for a few PDFs I was struggling with. Still, websites are def easier to scrape overall. I’m gonna give Selenium a shot for those pesky dynamic sites. Anyone have tips for handling CAPTCHAs? That’s my next hurdle lol.

Messages In This Thread



Users browsing this thread: 1 Guest(s)