02 Sep
|
Scrapingdog
|
Jaipur
02 Sep
Scrapingdog
Jaipur
Company Description
Scrapingdog is a specialized web scraping API that enables developers and data teams to extract data from search engines, e-commerce platforms, and websites without managing proxies, browsers, or CAPTCHAs. The platform delivers clean HTML or structured data in a single request, automatically handling IP rotation, browser rendering, retries, and anti-bot protections. Organizations use Scrapingdog for SEO, SERP tracking, price monitoring, market research, LLM dataset collection, and large-scale data gathering.
Built for teams that need reliable scraping infrastructure, Scrapingdog emphasizes predictable performance, straightforward integration, and scalability across diverse data use cases. The team is focused on delivering robust tools and responsive support for technical users and partners.
Role Description The Senior Full Stack Engineer role at Scrapingdog is a full-time, on-site position based in Jaipur. In this role, the engineer designs, develops, and maintains end-to-end web applications and services that power Scrapingdog’s scraping infrastructure and user interfaces. Day-to-day responsibilities include implementing backend APIs, optimizing data processing pipelines, building responsive frontend components, and ensuring system performance and reliability at scale.
The engineer collaborates closely with product, data, and support teams to refine requirements, improve developer experience, and deliver new features. The role also involves code reviews, mentoring team members, and contributing to architectural decisions and best practices.
What you'll work on
- Building and maintaining browser automation infrastructure (Playwright/Puppeteer) against a distributed remote browser pool
- Anti-bot evasion and fingerprint consistency — TLS/JA3 fingerprints, canvas/WebGL spoofing, sec-ch-ua and other header consistency, viewport/UA matching
- Reverse-engineering internal/undocumented endpoints (XHR/RPC calls,
token generation, signed request blobs) on major platforms (Google, Amazon, Walmart, TikTok, etc.)
- Designing and scaling cookie/session harvesting pipelines (Redis/MongoDB-backed pooling, session tiering)
- Working against enterprise-grade anti-bot systems — Cloudflare, Akamai, DataDome, PerimeterX, Google BotGuard — and staying ahead as they evolve
- Proxy infrastructure strategy — residential/ISP/datacenter rotation, IP reputation, geo-targeting
- Diagnosing and fixing infra-level issues affecting scrape reliability (rendering engine quirks, headless detection, server migrations)
- Owning scraper reliability at scale — this directly powers a production API serving 1,000+ paying customers
What we're looking for
- Deep, hands-on experience with browser automation frameworks (Playwright, Puppeteer, or Selenium) — not just using them, but understanding their internals well enough to patch around detection
- Strong understanding of how modern anti-bot systems fingerprint traffic — TLS, HTTP/2 fingerprinting, canvas/WebGL, font/audio fingerprinting, behavioral signals
- Experience reverse-engineering web/mobile app traffic — reading obfuscated JS, replicating signed requests, understanding token generation logic
- Solid grasp of proxy ecosystems and how to use them well (not just "rotate IPs")
- Backend chops: Node.js/Express, MongoDB, Redis, and comfort operating infrastructure (PM2, Nginx, Linux servers) — this isn't a role where someone else handles ops
- A track record of scraping at scale — ideally you've broken and rebuilt scrapers against major platforms before, more than once
- Comfortable working with ambiguity — platforms change their defenses constantly; you'll need to diagnose "why did this suddenly break" fast
Nice to have
- Experience with CAPTCHA-solving integration and detection avoidance
- Familiarity with BotGuard-style challenge-response systems
- Contributions to open-source scraping/anti-detect tooling
📌 Full Stack Engineer (Jaipur)
🏢 Scrapingdog
📍 Jaipur