08 Oct
|
Terrabit Consulting- India
|
India
08 Oct
Terrabit Consulting- India
India
Task Description: 5+ years
of hands-on Dev Ops/SRE experience operating production system located in MS Azure, experienced in used technologies: Terraform, Kubernetes, Elastic Search, Kafka, CI/CD pipeline, Grafana.
Skilful in security, networking, resiliency, failover process areas.
Ability to work on cost optimization and cloud governance.
Work Result: Smoohly runnig system, build team supported from devops POV
Skill Area: Infrastructure
Technology: Microsoft Azure
Proficiency - Technology: Expert
Secondary Skill Area: Infrastructure
Secondary Technology: Kubernetes
Proficiency - Secondary Technology: Expert
Other Skills: Excellent communication skills and a passion for sharing knowledge with engineering teams about infrastructure and cloud-native patterns.
Experience with open-source governance, license evaluation, and community-driven tooling.
Senior Dev Ops Engineer – Job Requirements
Key Responsibilities
• Design, build, and continuously improve cloud infrastructure in Microsoft
Azure,
ensuring scalability, reliability, and cost efficiency.
• Develop and maintain Infrastructure as Code (IaC) using Terraform, enabling
consistent and automated environment provisioning.
• Implement, optimize, and support CI/CD pipelines using Git Hub Actions,
ensuring
smooth and reliable delivery workflows.
• Operate and enhance Kubernetes clusters, including deployment strategies,
scaling, and troubleshooting.
• Manage and optimize Elastic Search and Confluent Kafka deployments for
performance, reliability, and cost.
• Strengthen platform security through patching, vulnerability mitigation,
backup
strategies, and proactive risk analysis.
• Drive operational excellence by defining, executing, and improving resiliency
procedures, including disaster recovery simulations and incident response
drills.
• Work closely with development teams to ensure smooth integration of
applications
into production environments
• Ensure best practices for cloud security and compliance are followed across
all
environments
• Monitor system health using Prometheus → Grafana stack and implement
automated alerting and remediation.
• Troubleshoot complex issues across infrastructure, applications, and
integrations,
collaborating with SMEs and vendor support when needed.
• Evaluate and recommend open-source technologies, especially when licensing,
support, or ecosystem changes occur.
• Promote Dev Ops best practices across engineering teams, mentoring developers
on infrastructure patterns, automation, and cloud-native principles.
• Document new solutions and maintain existing documentation, ADRs
Core Requirements & Skills
• 5+ years of hands-on Dev Ops/SRE experience operating production systems at
scale.
• Deep practical expertise with Microsoft Azure (networking, compute, storage,
identity, security).
• Strong proficiency with Terraform and Infrastructure as Code best practices.
• Solid experience running and troubleshooting Kubernetes clusters (AKS
preferred).
• Hands-on experience with Elastic Search and Confluent Kafka in production
environments.
• Strong automation mindset with scripting skills (Bash, Python, or similar).
• Experience building and maintaining CI/CD pipelines using Git Hub Actions.
• Proficiency with Prometheus and Grafana for monitoring, alerting, and
observability.
• Strong understanding of platform security: patching, secrets management,
RBAC,
network security, identity; PKI basics (certificates, chain of trust)
• Networking basics: TCP/IP, DNS, segmentation, routing, troubleshooting
• Experience designing and executing resiliency, DR, and failover procedures.
• Ability to troubleshoot complex distributed systems across infra, networking,
and
application layers.
• Comfortable evaluating and integrating open-source technologies, including
assessing licensing and support risks.
• Experience with cost optimization and cloud governance.
• Excellent communication skills and a passion for sharing knowledge with
engineering teams about infrastructure and cloud-native patterns.
Preferred Qualifications
• Azure certifications (AZ-104, AZ-305, AZ-500, AZ-400) or equivalent cloud
credentials.
• Hashi Corp Terraform Associate certification.
• Kubernetes certification (CKA, CKS)
• Familiarity with SRE practices (SLIs/SLOs, error budgets).
• Experience with open-source governance, license evaluation, and community driven tooling.
Responsibilities
Task Description: 5+ years
of hands-on Dev Ops/SRE experience operating production system located in MS Azure, experienced in used technologies: Terraform, Kubernetes, Elastic Search, Kafka, CI/CD pipeline, Grafana.
Skilful in security, networking, resiliency, failover process areas.
Ability to work on cost optimization and cloud governance.
Work Result: Smoohly runnig system, build team supported from devops POV
Skill Area: Infrastructure
Technology: Microsoft Azure
Proficiency - Technology: Expert
Secondary Skill Area: Infrastructure
Secondary Technology: Kubernetes
Proficiency - Secondary Technology: Expert
Other Skills: Excellent communication skills and a passion for sharing knowledge with engineering teams about infrastructure and cloud-native patterns.
Experience with open-source governance, license evaluation, and community-driven tooling.
Senior Dev Ops Engineer – Job Requirements
Key Responsibilities
• Design, build, and continuously improve cloud infrastructure in Microsoft
Azure,
ensuring scalability, reliability, and cost efficiency.
• Develop and maintain Infrastructure as Code (IaC) using Terraform, enabling
consistent and automated environment provisioning.
• Implement, optimize, and support CI/CD pipelines using Git Hub Actions,
ensuring
smooth and reliable delivery workflows.
• Operate and enhance Kubernetes clusters, including deployment strategies,
scaling,
and troubleshooting.
• Manage and optimize Elastic Search and Confluent Kafka deployments for
performance, reliability, and cost.
• Strengthen platform security through patching, vulnerability mitigation,
backup
strategies, and proactive risk analysis.
• Drive operational excellence by defining, executing, and improving resiliency
procedures, including disaster recovery simulations and incident response
drills.
• Work closely with development teams to ensure smooth integration of
applications
into production environments
• Ensure best practices for cloud security and compliance are followed across
all
environments
• Monitor system health using Prometheus → Grafana stack and implement
automated alerting and remediation.
• Troubleshoot complex issues across infrastructure, applications, and
integrations,
collaborating with SMEs and vendor support when needed.
• Evaluate and recommend open-source technologies, especially when licensing,
support, or ecosystem changes occur.
• Promote Dev Ops best practices across engineering teams, mentoring developers
on infrastructure patterns, automation, and cloud-native principles.
• Document new solutions and maintain existing documentation, ADRs
Core Requirements & Skills
• 5+ years of hands-on Dev Ops/SRE experience operating production systems at
scale.
• Deep practical expertise with Microsoft Azure (networking, compute, storage,
identity, security).
• Strong proficiency with Terraform and Infrastructure as Code best practices.
• Solid experience running and troubleshooting Kubernetes clusters (AKS
preferred).
• Hands-on experience with Elastic Search and Confluent Kafka in production
environments.
• Robust automation mindset with scripting skills (Bash, Python, or similar).
• Experience building and maintaining CI/CD pipelines using Git Hub Actions.
• Proficiency with Prometheus and Grafana for monitoring, alerting, and
observability.
• Strong understanding of platform security: patching, secrets management,
RBAC,
network security, identity; PKI basics (certificates, chain of trust)
• Networking basics: TCP/IP, DNS, segmentation, routing, troubleshooting
• Experience designing and executing resiliency, DR, and failover procedures.
• Ability to troubleshoot complex distributed systems across infra, networking,
and
application layers.
• Comfortable evaluating and integrating open-source technologies, including
assessing licensing and support risks.
• Experience with cost optimization and cloud governance.
• Excellent communication skills and a passion for sharing knowledge with
engineering teams about infrastructure and cloud-native patterns.
Preferred Qualifications
• Azure certifications (AZ-104, AZ-305, AZ-500, AZ-400) or equivalent cloud
credentials.
• Hashi Corp Terraform Associate certification.
• Kubernetes certification (CKA, CKS)
• Familiarity with SRE practices (SLIs/SLOs, error budgets).
• Experience with open-source governance, license evaluation, and community driven tooling.
Qualifications
Any Degree
📌 Senior DevOps Engineer (India)
🏢 Terrabit Consulting- India
📍 India