Senior Site Reliability Engineer (Hyderabad)

Senior Site Reliability Engineer (Hyderabad)

18 Sep
|
FTD India Private
|
Hyderabad

18 Sep

FTD India Private

Hyderabad

As a Senior Site Reliability Engineer, you will build and maintain the core platform automation, infrastructure, and observability systems that power FTD's high-scale commerce engine. In this developer-centric role, you will apply software engineering principles to operations—treating infrastructure as code, building self-service tooling for development teams, and optimizing application performance and reliability. You will work closely with cross-functional product teams to establish operational standards, define reliability metrics, and build automated systems that eliminate toil.

This position supports a hybrid work model including onsite presence in our Hyderabad, India office as needed. Occasional on-call and overtime work will be required but generally it is not expected to be significant.

KEY RESPONSIBILITIES::

· Maintain availability, performance, and scalability of critical services and production environments.

· Collaborate closely with developers to design reliable applications and improve deployment practices (you will be embedded in the Development team but reporting to infrastructure leader).

· Break down walls and build trust between developers and infrastructure teams

· Participate in end-to-end application ownership throughout the CI/CD process, including automated testing, observability, dependency management, and other operational concerns.

· Improve CI/CD pipeline reliability, traceability, and security

· Build automation for provisioning, configuration, deployments, and incident response.

· Improve observability using metrics, logs, distributed tracing, dashboards, and alerting.

· Participate in on-call rotation, lead incident response, and drive root cause analysis.

· Conduct capacity planning, chaos testing, and reliability reviews.

· Implement infrastructure-as-code using Terraform, Helm, Jenkins/GitHub Actions/etc.

· Optimize CI/CD pipelines and ensure safe, repeatable deployments (i.e. ArgoCD).

· Champion SRE principles: SLIs/SLOs, error budgets, toil reduction, problem management,



blameless postmortems.

· Embrace a culture of enablement, customer service, continuous improvement, transparency, and fiscal responsibility

· Perform other duties as directed

REQUIRED SKILLS AND ABILITIES

· 5+ years designing, developing, delivering, and operating scalable, available, high-performance applications (Java and node.js, etc.)

· Bachelor's or advanced degree in Computer Science, Information Systems, or a related field

· Familiarity with modern application languages and concepts, with hands-on e-commerce software development experience preferred

· Advanced hands-on experience with continuous integration and delivery / deployment methodologies and technologies

· Solid experience with containerization (e.g. Docker), Kubernetes, and Infrastructure as Code (Terraform preferred)

· Proficient understanding of microservices principles and orchestration. Working knowledge of Helm.

· Excellence in navigating and prioritizing multiple simultaneous responsibilities of varying scope and complexity

· Ability to effectively articulate technical concepts to audiences at all organizational levels via oral, written, and other non-verbal communications

· Demonstrated desire and ability to be self-directed, take ownership of issues, and establish a prominent level of credibility

· Ability to work well independently and within dynamic, cross-functional teams

· Experience with rapid detection and resolution of technical issues using various monitoring and application performance management tools

· Proficiency with shell scripting, Python and/or other scripting languages

· Ability to operate effectively under pressure, both independently and in collaboration with other resources

· Ability to rapidly learn new technologies via mentoring, formal training, independent research and testing

· A genuine desire and willingness to share knowledge effectively with others

Skills: Kubernetes, Cicd, Site Reliability Engineering, Development, Java, Cloud, Sla

Experience: 5.00-8.00 Years

Education: Bachelor Of Technology (B.Tech/B.E), Masters in Technology (M.Tech/M.E)

📌 Senior Site Reliability Engineer (Hyderabad)
🏢 FTD India Private
📍 Hyderabad

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: senior site reliability engineer (hyderabad) / hyderabad

Subscribe to this job alert:

Get the latest job offers by email for: senior site reliability engineer (hyderabad) / hyderabad