29 Aug
|
Valasys Media
|
Pune
29 Aug
Valasys Media
Pune
Web Scraper
Location: Pune | Work From Office
Experience: 3–5 Years Preferred | Minimum 2 Years
Working Days: Monday to Friday
Shift: EST – 6:00 PM to 3:00 AM IST
Employment Type: Full-Time
About the Role
We are looking for a skilled and hands-on Web Scraper to join our technology team.
The ideal candidate will have strong experience in Python development, web scraping, data engineering, automation, API integrations, and B2B data intelligence. You will be responsible for building scalable data collection and processing systems that power research, lead intelligence, enrichment, validation, and automation workflows.
This is an ideal opportunity for a developer who enjoys solving complex scraping and data challenges and working on production-level systems.
What You'll Do
- Develop and maintain scalable Python-based web scraping solutions.
- Build scraping pipelines using Scrapy, BeautifulSoup, Requests, Playwright, and/or Selenium.
- Handle modern, JavaScript/AJAX-driven websites, pagination, infinite scroll, sessions, cookies, DOM, XPath, and CSS selectors.
- Develop high-performance scripts using asynchronous and concurrent programming.
- Build ETL pipelines for data extraction, cleaning, transformation, normalization, and processing.
- Work with large volumes of B2B data for company research, contact discovery, persona identification, ICP filtering, and enrichment.
- Develop email validation and data-quality workflows, including deduplication, domain/MX validation, and bounce/suppression handling.
- Integrate REST APIs, webhooks, JSON, OAuth, and third-party services.
- Build automated workflows using Cron, Celery, task queues, schedulers, and retry mechanisms.
- Integrate approved outreach/email platforms and CRM/marketing systems through APIs.
- Work with AI/LLM APIs for research, classification,
enrichment, entity extraction, and automation.
- Implement data-quality controls, validation, confidence scoring, monitoring, and auditability.
- Work with databases such as PostgreSQL/MySQL, with MongoDB/Redis exposure being an advantage.
- Use Git, Linux, Docker, AWS/Azure, and related development/deployment tools.
- Troubleshoot, optimize, and continuously improve scraping and data-processing pipelines.
What We're Looking ForMust-Have Skills
- Strong hands-on experience with Python
- Strong understanding of Object-Oriented Programming (OOP)
- Practical experience in Web Scraping
- Experience with Scrapy / BeautifulSoup / Requests
- Experience with Playwright and/or Selenium
- Understanding of asynchronous/concurrent programming
- Experience with Pandas and ETL workflows
- Knowledge of PostgreSQL/MySQL
- Experience with REST APIs and third-party integrations
- Robust understanding of HTML/DOM, XPath, and CSS selectors
- Experience handling dynamic and JavaScript-driven websites
- Strong debugging and problem-solving skills
Good to Have
- MongoDB / Redis
- Celery / task queues
- Docker and Linux
- AWS/Azure
- CRM and marketing-platform integrations
- Email/outreach platform integrations
- AI/LLM API experience
- B2B data intelligence and lead enrichment experience
- Data validation and quality frameworks
Experience & Compensation
- Ideal Experience: 3–5 years
- Minimum Experience: 2 years,
provided the candidate demonstrates exceptional hands-on capability
- Preferred Mid-Level Range: 2.5–4 years
- Flexibility may be considered for exceptionally strong candidates with proven production-level experience.
- Freshers will not be considered.
Compliance & Responsible Data Collection Responsible data collection is an important part of this role.
Candidates should have an understanding of and experience working within applicable website terms, privacy/data-protection requirements, access controls, and email/anti-spam requirements while designing and operating data-collection systems.
Why This Role?
- Work on real-world, production-level scraping and data engineering challenges
- Opportunity to work across Python, automation, APIs, databases, AI/LLMs, and data intelligence
- Build scalable systems used for B2B research, lead intelligence, enrichment, and automation
- Work in a technically challenging environment with opportunities to expand your engineering skill set
Interested? If you are a hands-on Python developer who enjoys web scraping, automation, data engineering, and solving complex technical problems, we'd love to hear from you.
Apply now and be part of our technology team!
Pay: ₹59,000.00 - ₹70,000.00 per month
Benefits
- Food provided
- Health insurance
- Leave encashment
- Paid sick time
- Paid time off
- Provident Fund
Application Question(s):
- What is your current CTC?
- What is your expected CTC?
- What is your Notice Period?
- Are you comfortable with WFO- 6:00PM till 3:00AM (5 days working)?
- Do you have experience with working in B2B Industry?
- Do you have experience in Python?
- Do you have experience in Web Scrapping?
Work Location: In person
📌 Senior Data Engineer (Pune)
🏢 Valasys Media
📍 Pune