Senior Reliability Engineer (Hyderabad)

Senior Reliability Engineer (Hyderabad)

06 Aug
|
LSEG
|
Hyderabad

06 Aug

LSEG

Hyderabad

Cloud Reliability Engineer

? Hyderabad, India | ? 24×7 Reliability Engineering Model

Build Reliability at Scale.

Own Critical Cloud

Systems.

We’re hiring a Cloud SRE Engineer to join our Operations Reliability Engineering (ORE) team—focused on keeping mission-critical systems highly available, scalable, and resilient.

This role goes beyond traditional support. You’ll operate at the intersection of cloud engineering, reliability, and automation, owning the health of distributed systems running on AWS while driving operational excellence.

If you enjoy deep system visibility, proactive reliability engineering, and solving production-scale challenges, this role is for you.

What You’ll Own

? Cloud Reliability & Platform Operations

Own and manage reliability of AWS-based infrastructure (EC2, S3, RDS/Aurora)

Drive best practices for resource management, tagging, and optimization

Perform lifecycle operations on cloud infrastructure aligned to business needs

Manage AMI strategies, backup readiness, and system recovery posture

? Observability, Monitoring & Incident Engineering

Leverage Datadog to proactively detect and resolve issues

Build strong situational awareness across:

Infrastructure health (CPU, memory, storage)

Database performance and replication lag

Act on alerts with a focus on root cause identification, not just resolution

Improve alert quality, reduce noise, and enhance system observability

? Database Reliability Engineering

Ensure high availability and performance of Aurora/RDS clusters

Monitor and optimize

Replication and mirroring health

Query performance and database load

Partner with engineering teams on performance tuning and resilience improvements

? Data Protection, Backup & Recovery Engineering

Own the integrity and reliability of enterprise backup systems (Commvault)

Ensure backups meet strict SLA and compliance requirements

Perform and validate restore operations (critical recovery scenarios)

Continually improve backup validation, reporting, and audit readiness

? Security, Access & Compliance

Manage certificates, access provisioning, and credential governance

Support SOX compliance and audit requirements with precision

Ensure all operations meet security and regulatory standards

⚙️ Operational Excellence & Automation Opportunities

Drive consistency through runbooks, SOPs, and process improvements

Identify opportunities to automate repetitive operational tasks

Contribute to evolving the team from reactive support → proactive SRE culture

What Makes You a Strong Fit

? Core Skills

Strong hands-on experience with core AWS services, including EC2, S3,



RDS/Aurora, and CloudWatch

Deep expertise across AWS service domains:

Compute: EC2, Auto Scaling, AWS Lambda

Storage: S3, EBS, EFS

Databases: RDS, Aurora, DynamoDB

Networking: VPC, Subnets, Load Balancers (ALB/NLB), Route 53

Proven ability to design and implement highly available, scalable, and fault-tolerant architectures, including multi-AZ deployments and disaster recovery (DR) strategies

Strong experience with observability and monitoring tools (e.g., Datadog preferred, CloudWatch, Prometheus/Grafana)

Hands-on experience with backup and recovery solutions (e.g., Commvault), including DR planning and testing

Solid understanding of

Distributed systems and reliability engineering principles

Database operations, including replication, backup, restore, and performance tuning

Experience working in production environments with high availability and critical uptime requirements

? Engineering Mindset

You think beyond tasks—focused on system reliability and resilience

Strong troubleshooting skills with a bias for root cause analysis

Comfort working in mission-critical environments with real-time impact

? Ops + SRE Balance

Experience in 24×7 production environments

Ability to operate calmly under pressure while maintaining precision

Interest in moving towards automation, efficiency, and reliability engineering practices

➕ Nice to Have

Scripting experience (Python, Shell) for automation

Exposure to ITIL, incident management frameworks

Experience in regulated environments (SOX compliance)

✅ Recommended Certifications (Preferred)

AWS Certified Cloud Practitioner (CLF-C02)

AWS Certified Solutions Architect – Associate

Why This Role Stands Out

? Work on business-critical, high-scale cloud systems

⚙️ Be part of a team evolving towards true SRE practices—not just operations

? Gain deep exposure to AWS, observability, and database reliability at scale

? High ownership, real impact, and strong growth in Cloud + SRE + DevOps

? If you’re passionate about reliability, enjoy solving real production challenges, and want to operate at scale—this is your role. Apply now.

We're proud to have been recognised as a Great Place to Work® in India ‘25

Career Stage

Senior Associate





London Stock Exchange Group (LSEG) Information:

Join us and be part of a team that values innovation, quality, and continuous improvement. If you're ready to take your career to the next level and make a significant impact, we'd love to hear from you.

LSEG is a leading global financial markets infrastructure and data provider. Our purpose is driving financial stability, empowering economies and enabling customers to create sustainable growth.

Our purpose is the foundation on which our culture is built. Our values of Integrity, Partnership, Excellence and Change underpin our purpose and set the standard for everything we do, every day. They go to the heart of who we are and guide our decision making and everyday actions.

Working with us means that you will be part of a energetic organisation of 25,000 people across 65 countries. However, we will value your individuality and enable you to bring your true self to work so you can help enrich our diverse workforce.

We are proud to be an equal opportunities employer. This means that we do not discriminate on the basis of anyone’s race, religion, colour, national origin, gender, sexual orientation, gender identity, gender expression, age, marital status, veteran status, pregnancy or disability, or any other basis protected under applicable law. Conforming with applicable law, we can reasonably accommodate applicants' and employees' religious practices and beliefs, as well as mental health or physical disability needs.

You will be part of a collaborative and creative culture where we encourage new ideas. We are committed to sustainability across our global business and we are proud to partner with our customers to help them meet their sustainability objectives. Our charity, the LSEG Foundation provides charitable grants to community groups that help people access economic opportunities and build a secure future with financial independence. Colleagues can get involved through fundraising and volunteering.

LSEG offers a range of tailored benefits and support, including healthcare, retirement planning, paid volunteering days and wellbeing initiatives.

Please take a moment to read this privacy notice carefully, as it describes what personal information London Stock Exchange Group (LSEG) (we) may hold about you, what it’s used for, and how it’s obtained, your rights and how to contact us as a data subject. If you are submitting as a Recruitment Agency Partner, it is essential and your responsibility to ensure that candidates applying to LSEG are aware of this privacy notice.

📌 Senior Reliability Engineer (Hyderabad)
🏢 LSEG
📍 Hyderabad

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: senior reliability engineer (hyderabad) / hyderabad

Subscribe to this job alert:

Get the latest job offers by email for: senior reliability engineer (hyderabad) / hyderabad