09 Sep
|
Infowaysolutions
|
India
09 Sep
Infowaysolutions
India
Role: Site Reliability Engineer
About this Position
A Cloud Site Reliability Engineering Engineer closely works with app developers to tide the cloud infrastructure to the application behavior or deployment like a software engineer. The close collaboration with Cloud Architects is a key duty to align on-going implementations with architectural principals.
SRE is a key success factor of frequent and secure deployments of application in cloud.
Senior
App developers with cloud native knowledge are truly capable to take over devops tasks, too.
- Manage Cloud Platform in an agile product team with an DevSecOps mindset.
- Ability to learn SRE practices across Red Hat Open Shift, Google Cloud and Azure.
- Options to regularly become part of business projects through our Cloud Adoption practice.
- Define technical solution based on requirements and existing standards by Architects.
- Verify and improve cloud blueprints in collaboration with User, Engineers and DC Operators.
- Maintain Platforms features in SRE manner and improve availability on Image Factory, Monitoring, Code Maintenance or Configuration Managements.
- Upskill yourself constantly by regular research, testing and training.
- Ability in establishing automated monitoring or System healthy checks to ensure continuous operations (Azure Monitor & insights Services).
- Mentality to automate repetitive tasks at utmost level as a software engineer.
- IT security understanding is a must (Firewalls, Identity & Access or segregation of duties).
- Good practical experiences with Infrastructure-as-Code or Config frameworks like Hashicorp Terraform or Ansible.
- Experiences in using Hybrid Cloud technologies like in GCP, Azure, or RH OpenShift.
- Knowledge about hardware, software and concepts in the area of network & security.
- Practical experience with software development and DevOps methodologies.
- Experiences with the design and implementation of cloud based IoT or Data & Analytics solutions are beneficial.
- Incorporates aspects of software engineering and applies them to infrastructure and operations problems. The main goals are to create scalable and highly reliable software platforms and systems.
- Maintain platform services in Infrastructure as Code or Configuration Management and continuously improve the operations towards enterprise scale automation incl. shared services or workload related cloud services.
- Response to incidents or troubleshoot requests as 2nd/3rd level to Operational Support Engineers and take up improvement as well as required mitigation actions into automation.
- Walk through possible housekeeping opportunities to save costs and decrease security risks on infrastructure level.
- Implement and automate detail designs defined by Cloud architects or Engineers.
- Preferred is a university degree in Information Technology or a related degree in this subject.
- Excellent English skills, Open-mindedness and intrinsic motivation to learn recent topics.
Methodological Skills: Problem Solving, Technical Writing, Object-oriented programming, Test-Driven Design, Developer Toolchain, IT Customer Orientation.
Technical skills desired: Terraform, Ansible, PowerShell, Shellscripting (Bash), KVM Virtualization, Containers [Docker, Kubernetes], Git, GitOps, Infrastructure as Code, Configuration Management, SDLC, CI/CD, Grafana, Prometheus, ELK Stack, RedHat Openshift, Google Cloud Platform, MS Azure.
📌 Cloud Site Reliability Engineer (India)
🏢 Infowaysolutions
📍 India