16 Sep
|
Anonymous
|
Bengaluru
16 Sep
Anonymous
Bengaluru
Job Type: Full-time
Candidates ready to join immediately can share their details via email for quick processing.CCTC | ECTC | Notice Period | Location PreferenceAct fast for immediate attention! ⏳___________________________________________________________________________________________________Must-Have Skills6+ years of overall experience in Data Engineering / Big Data.Strong understanding of Big Data concepts and architecture.Strong hands-on experience with Apache Spark.Expertise in:Spark Performance TuningSpark OptimizationQuery/Job Performance ImprovementTroubleshooting Spark workloadsStrong hands-on experience with Py Spark and Spark.Strong programming experience in Python.Good experience working with My SQL / SQL.Strong experience in designing and developing Data Pipelines.Hands-on experience with Apache Airflow for data pipeline orchestration and scheduling.Experience working with at least one Cloud Platform.GCP experience is preferred.Good understanding of CI/CD and Dev Ops concepts.Experience integrating data engineering workloads with CI/CD pipelines.Strong debugging, troubleshooting,
and problem-solving skills.Preferred SkillsHands-on exposure to relevant GCP data services.Experience handling large-scale and high-volume datasets.Understanding of distributed data processing and data architecture.Experience improving the scalability, reliability, and performance of data pipelines.Exposure to Agile development and Dev Ops practices.Key ResponsibilitiesDesign, develop, and maintain scalable Big Data and Data Engineering solutions.Develop data processing applications using Python, Py Spark, and Apache Spark.Perform Spark performance tuning and optimization for large-scale workloads.Build, maintain, and monitor robust ETL/ELT data pipelines.Develop and manage workflow orchestration using Apache Airflow.Work with My SQL/SQL for data extraction, transformation, and validation.Deploy and support data engineering solutions in cloud environments, preferably GCP.Work with Dev Ops teams to implement and maintain CI/CD pipelines.Troubleshoot production issues and optimize data processing performance.Collaborate with engineering and business teams to deliver reliable and scalable data solutions.Primary Skill CombinationBig Data + Apache Spark + Py Spark + Python + Airflow + SQL/My SQL + Cloud (GCP Preferred) + CI/CDMandatory Focus: Robust hands-on Apache Spark performance tuning and optimization experience.
📌 6 + yoe - data engineer – big data / pyspark - any ust location - immediate joiner (Bengaluru)
🏢 Anonymous
📍 Bengaluru