04 Sep
|
TRIGENT SOFTWARE PRIVATE
|
India
04 Sep
TRIGENT SOFTWARE PRIVATE
India
As a Senior Dev
Ops Platform Engineer, a typical day involves designing, implementing, enhancing, and supporting cloud-native infrastructure platforms on AWS.
The role focuses on Kubernetes platform engineering, Infrastructure as Code (IaC), CI/CD architecture, reliability engineering, security governance, and operational excellence for large-scale digital platforms.
The candidate will work closely with development teams, operations teams, architects, and client stakeholders to build scalable, secure, highly available, and cost-effective platform solutions while following Dev
Ops, Site Reliability Engineering (SRE), and cloud engineering best practices.
Roles & Responsibilities
Expected to perform independently and become a Subject Matter Expert (SME) in AWS cloud infrastructure, Kubernetes platform engineering, and Dev
Ops practices.
Own AWS infrastructure and container platform architecture, including network topology, cluster design, scaling models, and multi-AZ or multi-region deployment and establish Infrastructure as Code standards using Terraform and configuration management standards using Ansible.
Define reusable Terraform module structures, state management strategies, code review processes, testing approaches, and governance standards.
Design and maintain enterprise-grade CI/CD pipeline architectures, including Jenkins shared libraries, reusable templates, security gates, quality controls, and deployment promotion deployment modernization initiatives through progressive delivery strategies to improve deployment reliability, reduce change failure rates, and accelerate delivery cycles.
Own platform reliability initiatives, including SLO/SLI definitions, error budget management, capacity planning, disaster recovery strategies, and business continuity testing.
Design, implement, and maintain highly available Kubernetes platforms in production environments.
Lead Kubernetes platform lifecycle activities including cluster upgrades, node management, tenant isolation, networking policies, and platform scalability improvements.
Manage EKS platform architecture, including IRSA implementation, VPC networking, storage integrations, CSI drivers, load balancer controllers, and cluster security.
Define and implement platform security controls including IAM least-privilege access, secrets management, security scanning, compliance monitoring, and vulnerability AWS cost optimization initiatives through resource governance, workload optimization, and Fin
Ops best practices.
Manage Helm chart lifecycle processes and Git
Ops operating models using tools such as ArgoCD or Flux.
Lead Terraform provider upgrades, Kubernetes version upgrades, and platform modernization root cause analysis (RCA) for critical incidents and lead cross-functional troubleshooting efforts across infrastructure, platform, networking, and application layers.
Represent the platform engineering team during CAB meetings, release planning discussions, and change management reviews.
Assess infrastructure enhancement requests, evaluate operational impacts, estimate capacity requirements, and provide technical with engineering, security, architecture, and product teams to build scalable platform solutions.
Mentor platform engineers through technical reviews, architecture discussions, troubleshooting sessions, and knowledge-sharing as the primary technical point of contact for client stakeholders and platform-related discussions.
Contribute to technical documentation, operational runbooks, disaster recovery procedures, and platform standards.
Support continuous improvement initiatives focused on automation, reliability, developer experience, and operational excellence.
Skilled & Technical Skills
Must Have Skills
Infrastructure as Code & Automation
Deep expertise in Terraform with a minimum of 4 years of hands-on production experience
📌 Software Engineering - DevOps Engineer (India)
🏢 TRIGENT SOFTWARE PRIVATE
📍 India