Job DescriptionCandidates ready to join immediately can share their details via email for quick processing.
NCCTC | ECTC | Notice Period | Location Preference
[email protected]
NAct fast for immediate attention!
N___________________________________________________________________________________________________
NMust-Have Skills
n
N
- 6+ years of overall experience in Data Engineering / Big Data.
N
- Strong understanding ofBig Data concepts and architecture.
N
- Strong hands-on experience with Apache Spark.
N
- Expertise in:
N
- Spark Performance Tuning
N
- Spark Optimization
n
- Query/Job Performance Improvement
N
- Troubleshooting Spark workloads
N
- Solid hands-on experience with PySpark and Spark.
N
- Strong programming experience in Python.
N
- Good experience workingwith MySQL / SQL.
N
- Strong experience in designing and developing Data Pipelines.
N
- Hands-on experience with Apache Airflow for data pipeline orchestration and scheduling.
N
- Experience working withat least one Cloud Platform.
N
- GCP experience is preferred.
N
- Good understanding of CI/CD and DevOps concepts.
N
- Experience integrating data engineering workloads with CI/CD pipelines.
N
- Strong debugging, troubleshooting, and problem-solving skills.
N
nPreferred Skills
N
n
- Hands-on exposureto relevant GCP data services.
N
- Experience handling large-scale and high-volume datasets.
N
- Understanding of distributed data processing and data architecture.
N
- Experience improving the scalability, reliability, and performance of data pipelines.
N
- Exposure to Agile development and DevOps practices.
N
nKey Responsibilities
N
n
- Design, develop, and maintain scalable Big Data and Data Engineering solutions.
N
- Develop data processingapplications using Python, PySpark, and Apache Spark.
N
- Perform Spark performance tuning and optimization for large-scale workloads.
N
- Build, maintain, and monitor robust ETL/ELT data pipelines.
N
- Develop and manage workflow orchestration using Apache Airflow.
N
- Work with MySQL/SQL fordata extraction, transformation, and validation.
N
- Deploy and support dataengineering solutions in cloud environments, preferably GCP.
N
- Work with DevOps teams to implement and maintain CI/CD pipelines.
N
- Troubleshoot productionissues and optimize data processing performance.
N
- Collaborate with engineering and business teams to deliver reliable and scalable data solutions.
N
nPrimary Skill Combination
NBig Data + Apache Spark +PySpark + Python + Airflow + SQL/MySQL + Cloud (GCP Preferred) + CI/CD
NMandatory Focus: Strong hands-on Apache Spark performance tuning and optimization experience.
📌 6 + Yoe - Data Engineer - Big Data / Pyspark - Any Ust Location - Immediate Joiner (Bengaluru)
🏢 UST
📍 Bengaluru