Yo! If you’re looking into how to scrape user accounts on instagram and tiktok aws, I’d recommend Selenium over Puppeteer. It’s more flexible and works better with dynamic content.
For rate limits, use a combo of proxies and user-agent rotation. I’ve had success with Scrapy + Scrapy-rotating-proxies.
Storage-wise, S3 is a no-brainer for bulk data. If you need structured data, maybe RDS or even Redshift if you’re dealing with a ton of it.
Just be careful with IG’s TOS, they’re pretty strict. 😅
For rate limits, use a combo of proxies and user-agent rotation. I’ve had success with Scrapy + Scrapy-rotating-proxies.
Storage-wise, S3 is a no-brainer for bulk data. If you need structured data, maybe RDS or even Redshift if you’re dealing with a ton of it.
Just be careful with IG’s TOS, they’re pretty strict. 😅
