Data Engineer –
Experience: 5+ Years
Location: Hyderabad / Chennai
Work Mode: On-site
Employment Type: Full-time
About the Role
We are looking for a passionate and experienced Data Engineer with 5+ years of experience in designing, developing, and maintaining scalable data platforms and ETL/ELT pipelines.
The ideal candidate should have strong expertise in Python, SQL, PySpark, Apache Spark, cloud data services, and modern data engineering frameworks, along with hands-on experience in Generative AI and Large Language Models (LLMs).
The candidate will play a key role in building reliable, scalable, and high-performance data solutions that support analytics, reporting, and AI/ML initiatives.
This is an on-site opportunity in Hyderabad or Chennai. Preference will be given to candidates currently based in these locations or willing to relocate.
Key Responsibilities
- Design, develop, and maintain scalable data pipelines and ETL/ELT workflows.
- Build and optimize data ingestion processes from multiple structured and unstructured data sources.
- Develop robust data models and data warehouses for analytics and reporting.
- Design and optimize SQL queries for high-performance data processing.
- Build and maintain data lakes using cloud storage solutions.
- Implement data validation, cleansing, transformation, and quality checks.
- Integrate data solutions with cloud platforms such as AWS, Azure, or Google Cloud Platform (GCP).
- Develop batch and real-time data processing pipelines using modern data processing frameworks.
- Collaborate with data scientists, analysts,
and application teams to deliver reliable data solutions.
- Follow software engineering best practices, including Git, CI/CD, testing, monitoring, and technical documentation.
Required Skills
- 5+ years of experience in Data Engineering.
- Strong programming experience in Python.
- Hands-on experience with PySpark and Apache Spark.
- Advanced SQL skills, including query optimization and performance tuning.
- Experience with AWS, Azure, or GCP.
- Strong experience in building scalable ETL/ELT pipelines.
- Strong experience in Generative AI and Large Language Models (LLMs).
- Valuable understanding of data warehousing concepts and dimensional modelling.
- Experience working with large-scale distributed datasets.
- Familiarity with Git and CI/CD pipelines.
- Understanding of Agile/Scrum methodologies.
- Preferred Qualifications
- Experience with Apache Airflow or similar workflow orchestration tools.
- Knowledge of Kafka and real-time data streaming.
- Familiarity with Docker, Kubernetes, and Terraform (IaC).
- Exposure to Databricks and modern data lake technologies such as Delta Lake, Apache Iceberg, or Apache Hudi.
- Understanding of data governance and metadata management.
- Strong exposure to Generative AI and LLM-based solutions.
Strong analytical, problem-solving, and communication skills
Educational Qualification
Bachelor's degree in Computer Science, Information Technology, Engineering, or a related technical discipline.
Location Requirement
? Hyderabad / Chennai
? On-site only
? 5+ Years of Experience
📌 Data Engineer (Chennai)
🏢 SMILO
📍 Chennai