Track Manager - High Performance Computing, Red Hat Enterprise Linux (India)

Track Manager - High Performance Computing, Red Hat Enterprise Linux (India)

19 Aug
|
HCLTech
|
India

19 Aug

HCLTech

India

Bengaluru, Karnataka
Job Summary

Key Responsibilities • Own Linux infrastructure lifecycle: provisioning, configuration, patching, upgrades, and capacity planning. • Design & implement: Server provisioning (Kickstart, PXE, Foreman, HA/DR architectures (e.g., Pacemaker; multi-AZ/multi-region patterns). • Automate everything: golden images, OS baselines, config compliance (Puppet/Ansible pipelines). • Build & run : Golden Image creations with lifecycle management and push to Infrastructure. • Security hardening at OS and network layers (SELinux, firewalld/nftables/IPSec, CIS baselines). • Performance engineering: low-latency tuning, kernel parameters, NUMA, I/O optimization. • Observability: Prometheus/Grafana dashboards knowledge. • Lead incident response and postmortems with blameless RCAs and concrete action plans. • Mentor engineers; contribute to architecture standards, design reviews, and runbooks. • Partner with Security, Networking, App, and DevOps teams on cross-functional deliverables. Mandatory Qualifications • 15+ years Linux engineering in production at scale (RHEL). • Expert in systemd, boot process, LVM, XFS/ext4, iSCSI/multipath, RAID; tcpdump, iproute2. • Strong Bash and Python; comfortable with regex and writing robust automation. • Hands-on with experts role in Puppet, Ansible, Cloud Migration (Azure and AWS), Git, and CI/CD tools (GitHub Actions). • Infrastructure as Code with Puppet and Ansible. • Security: SELinux, SSH hardening, vulnerability scanning (OpenSCAP/Qualys). • Cloud: experience with AWS/Azure, VPC/networking, IAM, On-prem to cloud migrations and managed K8s desirable. • Demonstrated leadership in complex troubleshooting, incident management, and mentoring. Success Metrics (First 6–12 Months) • MTTR reduced by 30% through runbooks/automation. • Baseline OS compliance 95% within 90 days; patch SLAs consistently met. • Automation coverage ( 70% of repetitive ops codified). • Documentation (golden images, infra patterns, puppet modules, playbooks) published and adopted. • Cloud Migration on-prem to AWS/Azure

Key Responsibilities

Key Responsibilities • Own Linux infrastructure lifecycle: provisioning, configuration, patching, upgrades, and capacity planning. • Design & implement: Server provisioning (Kickstart, PXE, Foreman, HA/DR architectures (e.g., Pacemaker; multi-AZ/multi-region patterns). • Automate everything: golden images, OS baselines, config compliance (Puppet/Ansible pipelines). • Build & run : Golden Image creations with lifecycle management and push to Infrastructure. • Security hardening at OS and network layers (SELinux, firewalld/nftables/IPSec, CIS baselines). • Performance engineering: low-latency tuning, kernel parameters, NUMA, I/O optimization. • Observability: Prometheus/Grafana dashboards knowledge. • Lead incident response and postmortems with blameless RCAs and concrete action plans. • Mentor engineers; contribute to architecture standards, design reviews, and runbooks.



• Partner with Security, Networking, App, and DevOps teams on cross-functional deliverables. Mandatory Qualifications • 15+ years Linux engineering in production at scale (RHEL). • Expert in systemd, boot process, LVM, XFS/ext4, iSCSI/multipath, RAID; tcpdump, iproute2. • Strong Bash and Python; comfortable with regex and writing robust automation. • Hands-on with experts role in Puppet, Ansible, Cloud Migration (Azure and AWS), Git, and CI/CD tools (GitHub Actions). • Infrastructure as Code with Puppet and Ansible. • Security: SELinux, SSH hardening, vulnerability scanning (OpenSCAP/Qualys). • Cloud: experience with AWS/Azure, VPC/networking, IAM, On-prem to cloud migrations and managed K8s desirable. • Demonstrated leadership in complex troubleshooting, incident management, and mentoring. Success Metrics (First 6–12 Months) • MTTR reduced by 30% through runbooks/automation. • Baseline OS compliance 95% within 90 days; patch SLAs consistently met. • Automation coverage ( 70% of repetitive ops codified). • Documentation (golden images, infra patterns, puppet modules, playbooks) published and adopted. • Cloud Migration on-prem to AWS/Azure

Skill Requirements

Key Responsibilities • Own Linux infrastructure lifecycle: provisioning, configuration, patching, upgrades, and capacity planning. • Design & implement: Server provisioning (Kickstart, PXE, Foreman, HA/DR architectures (e.g., Pacemaker; multi-AZ/multi-region patterns). • Automate everything: golden images, OS baselines, config compliance (Puppet/Ansible pipelines). • Build & run : Golden Image creations with lifecycle management and push to Infrastructure. • Security hardening at OS and network layers (SELinux, firewalld/nftables/IPSec, CIS baselines). • Performance engineering: low-latency tuning, kernel parameters, NUMA, I/O optimization. • Observability: Prometheus/Grafana dashboards knowledge. • Lead incident response and postmortems with blameless RCAs and concrete action plans. • Mentor engineers; contribute to architecture standards, design reviews, and runbooks. • Partner with Security, Networking, App, and DevOps teams on cross-functional deliverables. Mandatory Qualifications • 15+ years Linux engineering in production at scale (RHEL). • Expert in systemd, boot process, LVM, XFS/ext4, iSCSI/multipath, RAID; tcpdump, iproute2. • Strong Bash and Python; comfortable with regex and writing robust automation. • Hands-on with experts role in Puppet, Ansible, Cloud Migration (Azure and AWS), Git, and CI/CD tools (GitHub Actions). • Infrastructure as Code with Puppet and Ansible. • Security: SELinux,



SSH hardening, vulnerability scanning (OpenSCAP/Qualys). • Cloud: experience with AWS/Azure, VPC/networking, IAM, On-prem to cloud migrations and managed K8s desirable. • Demonstrated leadership in complex troubleshooting, incident management, and mentoring. Success Metrics (First 6–12 Months) • MTTR reduced by 30% through runbooks/automation. • Baseline OS compliance 95% within 90 days; patch SLAs consistently met. • Automation coverage ( 70% of repetitive ops codified). • Documentation (golden images, infra patterns, puppet modules, playbooks) published and adopted. • Cloud Migration on-prem to AWS/Azure

Other Requirements

Key Responsibilities • Own Linux infrastructure lifecycle: provisioning, configuration, patching, upgrades, and capacity planning. • Design & implement: Server provisioning (Kickstart, PXE, Foreman, HA/DR architectures (e.g., Pacemaker; multi-AZ/multi-region patterns). • Automate everything: golden images, OS baselines, config compliance (Puppet/Ansible pipelines). • Build & run : Golden Image creations with lifecycle management and push to Infrastructure. • Security hardening at OS and network layers (SELinux, firewalld/nftables/IPSec, CIS baselines). • Performance engineering: low-latency tuning, kernel parameters, NUMA, I/O optimization. • Observability: Prometheus/Grafana dashboards knowledge. • Lead incident response and postmortems with blameless RCAs and concrete action plans. • Mentor engineers; contribute to architecture standards, design reviews, and runbooks. • Partner with Security, Networking, App, and DevOps teams on cross-functional deliverables. Mandatory Qualifications • 15+ years Linux engineering in production at scale (RHEL). • Expert in systemd, boot process, LVM, XFS/ext4, iSCSI/multipath, RAID; tcpdump, iproute2. • Solid Bash and Python; comfortable with regex and writing robust automation. • Hands-on with experts role in Puppet, Ansible, Cloud Migration (Azure and AWS), Git, and CI/CD tools (GitHub Actions). • Infrastructure as Code with Puppet and Ansible. • Security: SELinux, SSH hardening, vulnerability scanning (OpenSCAP/Qualys). • Cloud: experience with AWS/Azure, VPC/networking, IAM, On-prem to cloud migrations and managed K8s desirable. • Demonstrated leadership in complex troubleshooting, incident management, and mentoring. Success Metrics (First 6–12 Months) • MTTR reduced by 30% through runbooks/automation. • Baseline OS compliance 95% within 90 days; patch SLAs consistently met. • Automation coverage ( 70% of repetitive ops codified). • Documentation (golden images, infra patterns, puppet modules, playbooks) published and adopted. • Cloud Migration on-prem to AWS/Azure

#body.unify div.unify-button-container .unify-apply-now: focus, #body.unify div.unify-button-container .unify-apply-#body.unify div.unify-button-container .unify-apply-now: focus, #body.unify div.unify-button-container .unify-apply-

📌 Track Manager - High Performance Computing, Red Hat Enterprise Linux (India)
🏢 HCLTech
📍 India

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: track manager - high performance computing, red hat enterprise linux (india) / india

Subscribe to this job alert:

Get the latest job offers by email for: track manager - high performance computing, red hat enterprise linux (india) / india