Hey everyone! đź‘‹
So, I’ve been trying to scrape twitter data for a project, and man, it’s been a rollercoaster. Twitter’s API is kinda restrictive, and scraping twitter without getting blocked feels like a game of cat and mouse.
I’ve tried a few tools like Tweepy, Snscrape, and even some custom scripts with BeautifulSoup + Selenium. Tweepy’s great if you’re okay with the API limits, but if you wanna scrape twitter at scale, Snscrape seems to be the go-to in 2023.
Anyone else have tips or tools they’re using? I’ve heard mixed things about Octoparse and Scrapy, but not sure if they’re worth the hassle. Also, how do y’all handle rate limits and IP bans? Proxies? Rotating user agents?
Would love to hear what’s working for you guys! Cheers! 🍻
So, I’ve been trying to scrape twitter data for a project, and man, it’s been a rollercoaster. Twitter’s API is kinda restrictive, and scraping twitter without getting blocked feels like a game of cat and mouse.
I’ve tried a few tools like Tweepy, Snscrape, and even some custom scripts with BeautifulSoup + Selenium. Tweepy’s great if you’re okay with the API limits, but if you wanna scrape twitter at scale, Snscrape seems to be the go-to in 2023.
Anyone else have tips or tools they’re using? I’ve heard mixed things about Octoparse and Scrapy, but not sure if they’re worth the hassle. Also, how do y’all handle rate limits and IP bans? Proxies? Rotating user agents?
Would love to hear what’s working for you guys! Cheers! 🍻
