11 days ago
Remote, IndiaMid Level
Responsibilities
- Refactor existing scraping scripts to improve reliability, maintainability, and efficiency.
- Design, build, and maintain robust web scrapers for complex and dynamic websites.
- Use fingerprinting, cookies, headers, user-agent rotation, proxies, and request-signature techniques to avoid detection and blocking.
- Manage dynamic content, complex DOM structures, browser rendering, and session or cookie lifecycles.
- Collaborate with analysts and stakeholders to gather requirements, align targets, and ensure data quality.
- Develop monitoring and alerting solutions to identify scraper failures and availability issues.
- Diagnose bottlenecks and scale scraping systems for efficiency, reliability, and performance.
- Propose new tooling, methodologies, and technologies for web data extraction.
- Provide documentation, support, and best practices to internal stakeholders.
Requirements
- At least 3 years of experience with web scraping frameworks such as Selenium, Playwright, or Puppeteer.
- Strong understanding of HTTP, RESTful APIs, HTML parsing, browser rendering, and TLS/SSL mechanics.
- Expertise in browser fingerprint spoofing, fingerprinting, evasion strategies, and request-signature manipulation.
- Deep experience managing cookies, headers, session states, and residential and data-center proxy rotations.
- Experience with logging, metrics, and alerting for high availability.
- Ability to troubleshoot and optimize scraper performance for efficiency, reliability, and scalability.
- Effective English communication with technical and non-technical stakeholders.
Benefits
- Fully remote opportunity based in India.
- Standard working hours are 11am–8pm IST, with flexibility.
- Vacation time, parental leave, team events, and learning reimbursement.
- Growth and development based on impact, ownership, and skill mastery.
Tech Stack
HTMLSelenium
