Wow, thanks for all the suggestions, everyone! I tried PyPDF2 and it worked pretty well for a few PDFs I was struggling with. Still, websites are def easier to scrape overall. I’m gonna give Selenium a shot for those pesky dynamic sites. Anyone have tips for handling CAPTCHAs? That’s my next hurdle lol.
Is it easier to scrape websites or PDFs? What are the pros and cons of each?
18 Replies, 1576 Views| Messages In This Thread |
|
Is it easier to scrape websites or PDFs? What are the pros and cons of each? - by - 11-02-2025, 11:57 PM
|
Users browsing this thread: 1 Guest(s)
