Data Engineer (Machine Learning) | 3–7 Years | Bangalore
We are looking for a skilled Data Engineer with 3–7 years of experience in Data Engineering and Machine Learning to build scalable data pipelines and support AI/ML solutions.
Key Responsibilities:
Design, develop, and maintain scalable ETL/ELT data pipelines.
Build and optimize data ingestion from multiple data sources.
Process large datasets using PySpark/Apache Spark.
Clean, transform, and prepare data for machine learning models.
Collaborate with Data Scientists and ML Engineers to develop and deploy ML solutions.
Perform feature engineering and optimize datasets for model training.
Monitor, troubleshoot, and improve data pipeline performance.
Ensure data quality, integrity, security, and governance.
Optimize SQL queries and database performance.
Deploy and maintain data workflows on cloud platforms.
Automate data pipelines using Airflow or similar orchestration tools.
Participate in code reviews, documentation, and Agile development practices.
Required Skills:
Python, SQL
PySpark / Apache Spark
ETL/ELT, Data Warehousing
Databricks or Snowflake
Machine Learning (Scikit-learn, TensorFlow, or PyTorch)
AWS, Azure, or GCP
Airflow, Git, Docker (preferred)