Source description
About the role
Architect, develop, and maintain scalable and distributed web scraping systems using Node.js . Design and implement data extraction pipelines to process large volumes of structured and unstructured data. Develop solutions to bypass anti-bot mechanisms , including CAPTCHA handling, session management, fingerprinting, and IP rotation . Optimize scraping processes for performance, reliability, and efficiency while managing proxy services (residential, datacenter, rotating).Oversee data storage and processing strategies , ensuring high availability and consistency. Collaborate with Product, DevOps, and Data Science teams to integrate extracted data into analytics and business applications. Implement best practices for microservices, API integrations, and real-time data streaming .
More at PortPro
