02 Sep
|
Infoweave
|
Coimbatore
02 Sep
Infoweave
Coimbatore
DataWeave is a cutting-edge AI-powered digital commerce analytics platform that empowers retailers with competitive intelligence and equips consumer brands with digital shelf analytics on a global scale. By harnessing the power of DataWeave, retailers gain the ability to make smarter pricing and merchandising decisions, while consumer brands can optimize their digital shelf performance across key performance indicators such as share of search, content quality, price competitiveness, and stock availability.
At the heart of DataWeave's capabilities lies its state-of-the-art AI-powered proprietary technology, which aggregates and analyzes 500 billion data points, covering over 400,000 brands, 4,000 websites, and spanning more than 20 industry verticals.
We are a globally distributed team, composed of over 200 talented engineers, product managers, and eCommerce experts located across San Francisco, Seattle, Austin, and Toronto in North America, complemented by our technology-focused offices in Bangalore and Coimbatore.
Data Engineering and Delivery @DataWeave
We the Delivery / Data engineering team at DataWeave, deliver the Intelligence with actionable data to the customer. One part of the work is to write effective crawler bots to collect data over the web, which calls for reverse engineering and writing scalable python code. Other part of the job is to crunch data with our big data stack / pipeline. Underpinnings are Tooling, domain awareness, fast paced delivery, and pushing the envelope.
How we work?
It's hard to tell what we love more, problems or solutions! Every day, we choose to address some of the hardest data problems that there are. We are in the business of making sense of messy public data on the web. At serious scale! Read more on Become a DataWeaver
Role Overview
We are seeking a proactive and sharp Data Engineering Intern to join our Delivery Team. In this role, youll get hands-on experience with next-gen technologies, AI-powered coding assistants, and modern data engineering practices. You’ll work on real client-facing projects, help build and maintain data pipelines, and respond swiftly to evolving needs.
You'll also gain exposure to startup dynamics—solving high-impact problems under tight timelines, while maintaining a strong focus on code quality and performance.
Key Responsibilities
- Build and maintain scalable data pipelines and ETL workflows using modern data stacks.
- Leverage tools like GitHub Copilot, Cursor, and other GenAI-assisted platforms to rapidly prototype and develop solutions.
- Collaborate closely with cross-functional teams to integrate data engineering components into AI-powered applications.
- Respond quickly to delivery timelines and shifting project requirements in a fast-paced environment.
- Ensure all code is well-documented, testable, and aligned with best practices.
- Showcase prior experience through Kaggle competitions, GitHub repos, or personal projects.
Required Skills & Qualifications
- Currently pursuing or recently completed a degree in Computer Science, Data Science, Engineering, or a related discipline.
- Proficiency in Python and SQL for data processing tasks.
- Experience with data engineering tools and frameworks (e.g., Airflow, dbt, Pandas, Spark, etc.).
- Exposure to cloud platforms (AWS, GCP, Azure) and version control systems (Git).
- Familiarity with GenAI and LLM tools such as OpenAI APIs, LangChain, and vector databases (e.g., Pinecone, Weaviate).
- Experience using AI-assisted development tools like Cursor and GitHub Copilot for faster delivery and prototyping.
Bonus / Preferred Skills
- Experience with web scraping frameworks (e.g., Scrapy, BeautifulSoup, Selenium) to gather external data.
- Understanding of big data technologies (e.g., Apache Spark, Hadoop, Kafka) and their application in large-scale data processing.
- Knowledge of data warehousing concepts and tools (e.g., Snowflake, BigQuery, Redshift).
- Participation in Kaggle competitions, hackathons, or public data science/data engineering projects.
What We’re Looking For
- A self-driven individual with solid problem-solving and analytical skills.
- Excellent communication skills to work effectively across tech and non-tech teams.
- Ability to thrive in a high-pressure, fast-moving startup environment.
- A strong commitment to quality and attention to detail in every aspect of your work.
- Hunger to learn, experiment, and grow with emerging technologies and real-world challenges.
What You’ll Gain
- Real-world experience with a cutting-edge GenAI-driven tech stack.
- Mentorship from experienced engineers and project leads.
- Involvement in full-cycle delivery for live client projects.
- Exposure to startup operations and decision-making.
- A potential pathway to a full-time opportunity based on performance and business needs.
📌 Data Engineer Intern (Coimbatore)
🏢 Infoweave
📍 Coimbatore