Job Title: Data Engineer
Experience: 69 Years Location: Pune (Hybrid – Client Office)
Job Summary:
Seeking a skilled Senior Data DevOps Engineer having experience in Cloudera platforms, data engineering, and DevOps automation. The ideal candidate will manage and optimize Cloudera environments, build CI/CD pipelines, and support enterprise-scale data processing workloads.
Key Responsibilities
Administer and support Cloudera CDP/CDH platforms, including HDFS, Hive, Spark, YARN, Hue, and CDE. Develop, deploy, and optimize PySpark and Python-based data processing solutions. Build and maintain CI/CD pipelines using Jenkins and GitHub/Bitbucket. Integrate security and code quality tools such as Checkmarx into delivery pipelines. Write and optimize SQL queries for Hive and Impala environments.
Monitor platform health, perform upgrades, troubleshoot issues, and ensure high availability. Implement security controls, access management, and governance best practices.
Required Skills Hands on experience with Cloudera Hadoop/CDP platforms. Solid expertise in HDFS, Spark/PySpark, Python, Hive, YARN, Hue, and CDE. Hands-on experience with Jenkins, GitHub/Bitbucket, and CI/CD practices. Positive knowledge of SQL and performance tuning. Experience with Checkmarx or similar SAST tools. Robust Linux administration and scripting skills. Excellent troubleshooting and problem-solving abilities
📌 Data Engineer Pune
🏢 hackajob
📍 Pune