Hadoop Administrator & Data Engineer (Thiruvananthapuram)

Hadoop Administrator & Data Engineer (Thiruvananthapuram)

12 Sep
|
Qzigma Technologies
|
Thiruvananthapuram

12 Sep

Qzigma Technologies

Thiruvananthapuram

We are looking for an experienced Hadoop Administrator & Data Engineer to manage and optimize production Hadoop environments while working closely with the Data Engineering team. The role requires robust hands-on expertise in HDFS, YARN, Hive, Spark, Linux administration, Ambari, and cloud-based data pipelines.

The ideal candidate should be comfortable troubleshooting production infrastructure, managing mixed Hadoop versions, performing capacity planning, optimizing resource utilization, and supporting data pipelines running through Apache Airflow on AWS MWAA.

Key Responsibilities

Hadoop & Cluster Administration

- Administer and optimize production HDFS, YARN, Hive, and Spark clusters.
- Manage Hadoop environments with mixed software versions and troubleshoot compatibility issues.
- Monitor cluster health, performance, capacity, and resource utilization.
- Administer clusters using Apache Ambari, including service configurations, alerts, rolling restarts, and version upgrades.
- Perform capacity planning, HDFS quota management, data retention, and storage cleanup.
- Monitor and optimize NameNode, DataNode, ResourceManager, NodeManager, Hive Metastore, and HiveServer2 services.

Hive & Spark Administration

- Manage Hive Metastore and HiveServer2, including JVM sizing and memory management.
- Troubleshoot Hive failures such as partition-related errors, serialization issues, and Thrift session failures.
- Support Spark-on-YARN workloads and troubleshoot driver and executor failures.
- Tune Spark memory overhead, resource allocation, concurrency, and workload performance.
- Analyze NameNode, Hive Metastore, Spark, and other service logs to identify root causes.

Linux & Infrastructure Management

- Perform hands-on Linux administration on RHEL/CentOS-family systems.
- Troubleshoot processes, memory utilization, disk capacity, network connectivity, and system performance.
- Manage gateway hosts and ensure adequate capacity for concurrent workloads.
- Monitor worker pools,



multiprocessing workloads, Spark drivers, and gateway resource consumption.
- Establish appropriate resource limits to prevent individual workloads from impacting shared infrastructure.
- Manage SSH access, SSH key hygiene, service accounts, and permissions across gateways and cluster nodes.

Monitoring & Performance

- Develop and enhance monitoring and alerting for:
- Host load and system health
- Per-tenant resource consumption

- HDFS capacity and utilization

- Hive Metastore health

- Gateway capacity

- Cluster services and availability
- Identify early warning indicators and proactively resolve infrastructure issues.
- Perform capacity planning and recommend resource allocation based on actual workload evidence and performance data.

Data Engineering & Cloud Support

- Work closely with Data Engineering teams to support production data pipelines.
- Provide infrastructure-level troubleshooting for Apache Airflow workflows running on AWS MWAA.
- Diagnose whether pipeline failures are related to infrastructure, cluster configuration, resource constraints, or application logic.
- Support data engineering teams through SSH-based access from gateway hosts.
- Collaborate with developers and data engineers to ensure reliable and scalable data processing.

Database & Search Infrastructure

- Support operations and troubleshooting for MySQL and MongoDB environments at scale.
- Work with Elasticsearch or Sphinx/Manticore search infrastructure.
- Monitor performance, troubleshoot issues, and coordinate infrastructure improvements as required.





Required Skills & Experience

- 4+ years of hands-on experience administering production Hadoop clusters.
- Strong experience with HDP, CDH, or equivalent Hadoop distributions.

- Deep knowledge of

- HDFS

- YARN

- Hive

- Spark
- Strong experience in Linux system administration, particularly RHEL/CentOS.
- Experience with process and memory troubleshooting, SSH, systemd, disk, and network troubleshooting.
- Hands-on experience with Spark-on-YARN, including memory overhead tuning and driver-side failure troubleshooting.
- Strong scripting skills in Bash and Python.
- Hands-on experience with Apache Ambari administration and Hadoop version upgrades.
- Experience with Apache Airflow, preferably AWS MWAA.
- Experience supporting MySQL and MongoDB in production/large-scale environments.
- Experience with Elasticsearch, Sphinx, or Manticore is preferred.
- Strong analytical and troubleshooting skills with the ability to perform root-cause analysis.
- Good understanding of shared infrastructure, resource management, capacity planning, and workload optimization.

Preferred Candidate Profile

- Strong production troubleshooting and incident-management experience.
- Ability to analyze logs, JVM behavior, memory utilization, and system performance.
- Comfortable working with large-scale distributed data platforms.
- Strong judgment regarding infrastructure capacity and resource requests.
- Proactive approach toward monitoring, alerting, optimization, and preventive maintenance.
- Good communication and collaboration skills with Data Engineering and technical teams.
- Ability to work independently in a remote/hybrid environment.

Work Schedule & Location

- Location: Thiruvananthapuram, Kerala
- First 6 Months: Hybrid work model
- After 6 Months: Fully Remote
- Shift: US Shift
- Candidate must be comfortable working during 6:00 AM5:00 PM Pacific Time business hours.

📌 Hadoop Administrator & Data Engineer (Thiruvananthapuram)
🏢 Qzigma Technologies
📍 Thiruvananthapuram

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: hadoop administrator & data engineer (thiruvananthapuram) / thiruvananthapuram

Subscribe to this job alert:

Get the latest job offers by email for: hadoop administrator & data engineer (thiruvananthapuram) / thiruvananthapuram