Wow, thanks for all the suggestions, everyone! I tried BulkGPT first, and it worked pretty well for scraping whole website robots.txt how bulkgpt does it, but I’m definitely gonna check out some of the other tools you mentioned, like Octoparse and Screaming Frog. I also tried writing a quick Python script with Requests, but I think I need to tweak the delays a bit—got blocked on one site lol. Anyway, really appreciate all the tips! If anyone’s got more advice on handling proxies, let me know!
How to Scrape a Whole Website's robots.txt File Using BulkGPT: Any Tips or Tools?
16 Replies, 1670 Views| Messages In This Thread |
|
How to Scrape a Whole Website's robots.txt File Using BulkGPT: Any Tips or Tools? - by - 25-10-2024, 08:12 PM
|
Users browsing this thread: 1 Guest(s)
