01 Sep
|
Unique Identification Authority of India (UIDAI)
|
Bengaluru
01 Sep
Unique Identification Authority of India (UIDAI)
Bengaluru
Platform Engineer
Role Overview
- The Platform Engineer will be responsible for designing, automating, and operating UIDAI’s core application and infrastructure platform to ensure high availability, scalability, and reliability across multiple data centres
- This role focuses on Service Reliability Engineering (SRE) principles, observability, and infrastructure automation, ensuring that UIDAI systems consistently meet defined SLA/SLO targets
- The position requires in-depth technical expertise in container orchestration, service mesh, CI/CD, monitoring, and distributed system reliability, combined with solid ownership, operational discipline, and collaborative problem-solving skills
Key Responsibilities
Platform Engineering & Reliability
- Design and maintain scalable, secure, and fault-tolerant infrastructure for UIDAI’s core systems deployed across multiple DCs
- Apply SRE principles — define, measure, and continuously improve SLIs/SLOs/SLAs for all mission-critical services
- Build self-healing and automated recovery mechanisms for core applications and APIs
- Perform capacity planning, failure testing, and reliability analysis
Infrastructure Automation & Configuration Management
- Design and manage infrastructure provisioning and configuration using automation frameworks
- Experience with Infrastructure as Code (IaC) tools like Terraform, Ansible, or Chef is desirable
- Develop repeatable automation for environment setup, deployment, and monitoring across hybrid DC environments
Containerization, Orchestration & Service Mesh
- Manage and optimise large-scale Kubernetes (K8s)
clusters for production workloads
- Implement and maintain Service Mesh (Istio/Linkerd) for secure, observable, and resilient service-to-service communication
- Work with CI/CD pipelines using Jenkins and ArgoCD (experience with any equivalent CI/CD platform is acceptable)
- Ensure isolation, policy enforcement, and resource quota management across namespaces and tenants
Monitoring, Observability & Incident Management
- Design and maintain the observability stack (Prometheus, Grafana, Loki, and OpenTelemetry)
- Define and monitor SLOs (availability, latency, and error budgets) for all critical services
- Develop automated alerting and runbooks aligned with UIDAI’s reliability goals
- Lead Root Cause Analysis (RCA) and Post Incident Reviews (PIR) for continuous improvement
Collaboration & Governance
- Collaborate with Product Managers, Architects, and Developers to ensure reliability and performance objectives are integrated into design
- Standardise environments across development, staging, and production
- Participate in architectural reviews and propose platform improvements to enhance scalability, resilience, and efficiency
Essential
- B.E./B.Tech/B.Sc. in Computer Science, Information Technology, or an equivalent discipline from a reputed university
- Minimum 6 years of experience
Desirable
- Certifications such as CKA/CKAD, Terraform Associate, AWS/GCP DevOps Engineer, or equivalent
- Recognized SRE/DevOps certifications (Google SRE, Linux Foundation, etc.)
UIDAI Values
- Citizen Centricity
- Ownership
- Excellence
- Comfort with Ambiguity
📌 Platform Engineer (Bengaluru)
🏢 Unique Identification Authority of India (UIDAI)
📍 Bengaluru