06 Aug
|
National Payments Corporation Of India (NPCI)
|
Hyderabad
06 Aug
National Payments Corporation Of India (NPCI)
Hyderabad
Job Summary
About the Role We are seeking a Incharge/Deputy Incharge Site Reliability Engineering to lead and scale our on-premises payments infrastructure. This role requires deep expertise in hybrid environments (bare metal, virtual machines, and containerized workloads) along with a strong focus on system reliability, performance, and security. You will drive platform resilience strategy, lead high-performing SRE teams, and integrate AI/ML-driven automation and observability into infrastructure operations.
Location: Hyderabad Experience: 15+ years
Key Responsibilities
- Infrastructure &
- Platform Engineering - Lead the design, operation, and scaling of on-premise infrastructure supporting mission-critical payments systems
- Manage hybrid environments: Virtual machines (VMware/KVM)
- Containerized workloads (Docker on bare metal, Kubernetes preferred)
- Middleware &
- Distributed Systems - Oversee and optimize: Nginx (reverse proxy, traffic routing, load balancing)
- Redis (caching, clustering, HA, persistence tuning)
- Kafka (streaming, partition design, replication, performance tuning)
- Drive architecture improvements for low latency and high throughput systems
- IT Ops - Build and scale CI/CD pipelines using Jenkins (or equivalent tools)
- Implement GitOps and infrastructure-as-code practices
- Lead automation initiatives using: Ansible
- Shell/Python scripting
- SRE &
- Reliability Engineering - Define and enforce SLOs, SLIs, and SLAs
- Lead incident management, RCA, and postmortem culture
- Drive proactive monitoring, alerting, and observability strategy
- Networking &
- Security - Manage and troubleshoot: TCP/IP, DNS, Load Balancing
- Firewalls, WAF, CDN (Akamai preferred)
- Ensure security, compliance (PCI-DSS preferred), and governance standards
- Lead vulnerability management and risk mitigation
- AI-Driven Operations &
- Innovation - Introduce AI/ML-based observability,
anomaly detection, and predictive maintenance
- Leverage AI tools for: Incident prediction and auto-remediation
- Intelligent log analysis
- Capacity planning and forecasting
- Evaluate and integrate AIOps platforms
- Leadership &
- Stakeholder Management - Lead and mentor a team of SREs/DevOps engineers (minimum 3+ years of people management)
- Drive hiring, performance management, and capability building
- Collaborate with Engineering, Security, Product, and Infra teams
- Own strategic roadmap for platform scalability and resilience
- Governance &
- Operations - Maintain infrastructure inventory, asset lifecycle, and documentation
- Ensure adherence to audit and compliance requirements
- Drive operational excellence and cost optimization
Required Qualifications
- 15+ years of experience in Infrastructure, SRE, or DevOps roles
- 3+ years in leadership/people management roles
- Strong experience managing on-premise or hybrid infrastructure at scale
- Deep hands-on expertise in: Linux (RHEL/CentOS/Ubuntu)
- Nginx, Redis, Kafka
- Jenkins, Git-based workflows
- Ansible and scripting
- Robust understanding of networking and security concepts
- Proven experience in high-availability systems and production support
Preferred Qualifications
- Exposure to Kubernetes and container orchestration platforms
- Experience in payments domain or financial services (highly desirable)
- Familiarity with PCI-DSS or similar compliance frameworks
- Experience with AIOps tools (e.g., Grafana, Victoria Logs, Victoria Metrics)
- Knowledge of cloud integration (AWS/Azure hybrid setups)
- Strong programming skills (Python/Go good to have)
Job Type: Full Time
Disclaimer: This job posting has been aggregated from external source. Role details, content, and availability are subject to change. Applicants are advised to confirm the latest information directly on the company website before applying.
📌 Deputy Incharge - Engineering Delivery & Reliability (Hyderabad)
🏢 National Payments Corporation Of India (NPCI)
📍 Hyderabad