AI & Web Scraping-Python and Javascript (Remote) (Bengaluru)

AI & Web Scraping-Python and Javascript (Remote) (Bengaluru)

09 Sep
|
AIMLEAP
|
Bengaluru

09 Sep

AIMLEAP

Bengaluru

Python & Java Script Developer – AI & Web Scraping

Experience: 3-5 Years
Location: Remote
Mode of Engagement: Full-time
No of Positions: 3
Educational Qualifications: Bachelor's degree in computer science, Information Technology
Industry: IT / Software Development

Notice Period: Immediate Joiners Preferred

About the Role

We are looking for a hands-on Web Scraping / Crawling Engineer with 3–5 years of experience in web scraping, browser automation, and scalable data extraction.

The ideal candidate should have robust experience working with dynamic and Java Script-heavy websites, building reliable crawling workflows, handling crawling failures, and working with distributed processing systems.

This is a highly technical, hands-on role. You will be expected to design, develop, debug, optimize, and maintain web crawling and data extraction systems.

Responsibilities

- Design, develop, and maintain scalable web crawling and scraping systems for dynamic and Java Script-heavy websites.
- Develop browser automation workflows using Playwright, Selenium, Puppeteer, or similar frameworks.
- Investigate and resolve crawling issues such as 403/429 responses, redirects, timeouts, rendering failures, session issues, and anti-bot challenges.
- Develop robust retry, fallback, and failure-handling mechanisms to improve crawler reliability.
- Work with cookies, sessions, browser contexts, headers, proxies, and related crawling mechanisms to maintain state and improve crawling reliability.
- Build and optimize concurrent and distributed scraping workflows using asynchronous processing, queues, and worker-based architectures.
- Design data pipelines covering URL processing, crawling, extraction, validation, transformation, and storage.




- Optimize crawler performance, including concurrency, browser lifecycle, resource utilization, timeouts, and request handling.
- Implement monitoring and observability for crawl success rates, failure types, latency, retries, and worker performance.
- Debug complex crawling problems and identify root causes rather than relying only on one-off fixes.
- Develop reusable crawling components and frameworks that can support multiple websites and use cases.
- Collaborate with data engineering, AI, backend, and product teams to deliver reliable and structured web data.
- Evaluate and adopt new web crawling, browser automation, and data extraction technologies where appropriate.

Required Skills

- 3–5 years of hands-on experience in web scraping, web crawling, browser automation, or web data engineering.
- Strong programming experience in Python. Java Script/Node.js is a plus.
- Strong hands-on experience with Scrapy, Playwright, Selenium, Puppeteer, or similar browser automation frameworks.
- Strong understanding of HTTP, HTML, DOM, Java Script rendering, redirects, cookies, sessions, headers, and browser contexts.
- Experience with browser fingerprinting, WAFs, and modern anti-bot mechanisms.
- Experience working with dynamic, Java Script-heavy, and asynchronous websites.
- Good understanding of asynchronous programming, concurrency, and parallel processing.




- Experience with job queues, distributed workers, or message-based processing systems such as Redis, RabbitMQ, Kafka, Celery, or similar technologies.
- Understanding of retry mechanisms, error handling, idempotency, rate limiting, and backpressure in distributed systems.
- Hands-on experience with proxy management, session handling, and anti-bot challenges.
- Strong debugging and analytical skills, with the ability to investigate and resolve complex crawling failures.
- Experience working with SQL databases such as PostgreSQL or MySQL.
- Familiarity with Docker and cloud platforms such as AWS, GCP, or Azure is a plus.
- Strong hands-on engineering mindset rather than purely project or delivery management experience.
- Ability to independently debug, investigate, and solve difficult crawling problems.
- Strong understanding of how browsers, HTTP requests, sessions, and websites interact.
- Ability to design systems that are reliable, scalable, and fault tolerant.
- Good understanding of how to distribute and coordinate large-scale scraping workloads.
- Ability to identify the root cause of crawling failures and develop generalized, reusable solutions.
- Willingness to work across scraping, backend services, distributed processing, data pipelines, and infrastructure when required.
- Strong problem-solving skills and curiosity to understand how websites behave rather than relying solely on existing scraping tools.

Education

- Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field is preferred.
- Equivalent practical experience in software engineering or web scraping will also be considered.

📌 AI & Web Scraping-Python and Javascript (Remote) (Bengaluru)
🏢 AIMLEAP
📍 Bengaluru

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: ai & web scraping-python and javascript (remote) (bengaluru) / bengaluru

Subscribe to this job alert:

Get the latest job offers by email for: ai & web scraping-python and javascript (remote) (bengaluru) / bengaluru