We are looking for a highly motivated and technically strong Junior Site Reliability Engineer SRE with 2-4 years of hands-on experience to join our ServiceNow Platform team in Bellandur, Bangalore. The ideal candidate will have a solid foundation in Linux systems, cloud infrastructure, CI/CD tooling, Kubernetes-based Java application deployments, and infrastructure automation.
Roles and Responsibility
- Act as the operational bridge between ServiceNow platform teams and SRE Infrastructure teams.
- Support and maintain ServiceNow distributed infrastructure, including MID Servers, ACC Agents, and API integrations.
- Monitor platform reliability, availability, and performance to meet SLA targets.
- Participate in incident response, troubleshooting root cause analysis (RCA), and remediation activities.
- Build and maintain Infrastructure as Code IaC using Puppet and Terraform.
- Automate deployments and operational tasks using GitLab CI/CD pipelines.
- Manage on-premise deployments and configuration management using Puppet.
- Support Kubernetes-based service architecture for Java applications, building, deploying, and troubleshooting containerized workloads in Kubernetes environments.
- Create and maintain Linux deployment artifacts and deployment pipelines, ensuring platform scalability, resilience, and operational stability.
- Implement and maintain logging, monitoring, and alerting solutions for applications and infrastructure, building telemetry dashboards and proactive monitoring for platform health.
- Analyze logs and system metrics to identify operational risks and performance bottlenecks, working with tools such as Prometheus, Grafana, OTel, or similar observability platforms.
- Manage secrets and credentials using HashiCorp Vault or equivalent secret management tools, ensuring secure access patterns and compliance with enterprise security standards.
- Support least privilege access models and infrastructure hardening initiatives, applying common troubleshooting methodologies for distributed systems and cloud-native environments.
- Support production incidents involving Kubernetes, Linux systems, networking deployments, and integrations, identifying operational toil and automating repetitive tasks.
- Collaborate with developers and platform teams to improve system reliability and deployment quality.
Job Requirements
- 2-4 years of experience in an SRE, DevOps, or Platform Engineering role.
- Hands-on experience with GitLab CI/CD, Puppet, Terraform, Kubernetes, Linux administration, AWS infrastructure, and Kubernetes service architecture for Java applications.
- Experience managing secrets using Vault and a solid understanding of logging, monitoring, and alerting systems.
- Strong analytical and problem-solving skills, with familiarity with image-building pipelines and automation workflows.
- Knowledge of Infrastructure as Code IaC and configuration management principles, along with experience troubleshooting distributed systems deployment failures and infrastructure issues.
- Working knowledge of Git version control systems and exposure to ServiceNow platform infrastructure such as MID Servers, ACC Agents, or Integration Hub.
- Experience with Docker container technologies and knowledge of Python, Java, or Go for automation scripting.
About Company
We are a global e-commerce company with business operations in nearly every country and city on the planet, and we want to make it easy for everyone anywhere in the world to pay for their travel or do business with our platform whenever and however it's convenient for them.
Disclaimer: This job posting has been aggregated from external source. Role details, content, and availability are subject to change. Applicants are advised to confirm the latest information directly on the company website before applying.
📌 Specialist - Package Implementation (Karnataka)
🏢 LTM
📍 Karnataka