**Before you apply to a job, select your language preference from the options available at the top right of this page.** Explore your next opportunity at a Fortune Global 500 organization. Envision cutting-edge possibilities, experience our rewarding culture, and work with talented teams that help you become better every day. We know what it takes to lead UPS into tomorrow-people with a unique combination of skill + passion. If you have the qualities and drive to lead yourself or teams, there are roles ready to cultivate your skills and take you to the next level. **Job Description:** **Responsibilities:** **System Reliability:** Ensure the reliability and uptime of critical services and infrastructure. **Google Cloud Expertise:** Design, implement, and manage cloud infrastructure using Google Cloud services. **Automation:** Develop and maintain automation scripts and tools to improve system efficiency and reduce manual intervention. **Monitoring and Incident Response:** Implement monitoring solutions and respond to incidents to minimize downtime and ensure quick recovery. **Collaboration:** Work closely with development and operations teams to improve system reliability and performance.
**Capacity Planning:** Conduct capacity planning and performance tuning to ensure systems can handle future growth. **Documentation:** Create and maintain comprehensive documentation for system configurations, processes, and procedures. **Qualifications:** Education: Bachelor's degree in computer science, Engineering, or a related field. Experience: 4+ years of experience in site reliability engineering or a similar role. **Skills:** Proficiency in Google Cloud services (Compute Engine, Kubernetes Engine, Cloud Storage, BigQuery, Pub/Sub, etc.). Familiarity with Google BI and AI/ML tools (Looker, BigQuery ML, Vertex AI, etc.) Experience with automation tools (Terraform, Ansible, Puppet). Familiarity with CI/CD pipelines and tools (Azure pipelines Jenkins, GitLab CI, etc.). Strong scripting skills (Python, Bash, etc.). Knowledge of networking concepts and protocols. Experience with monitoring tools (Prometheus, Grafana, etc.). **Preferred Certifications:** Google Cloud Professional DevOps Engineer Google Cloud Professional Cloud Architect Red Hat Certified Engineer (RHCE) or similar Linux certification **Employee Type:** Permanent UPS is committed to providing a workplace free of discrimination, harassment, and retaliation.
📌 Site Reliability Engineers (SREs) - Robust background in Google Cloud Platform (GCP) | RedHat OpenShift administration
🏢 UPS
📍 Chennai