Site Reliability Engineer (Bengaluru)

Site Reliability Engineer (Bengaluru)

06 Aug
|
Synechron
|
Bengaluru

06 Aug

Synechron

Bengaluru

Good day,

We have opportunity for Site Reliability Engineer (SRE)

Job Role: Site Reliability Engineer (SRE)

Job Location: Synechron ( Bengaluru BCIT )

Experience- 8 to 15 years

Notice: Immediate joiner to 15 days.

About Synechron: Synechron is a global technology consulting firm that helps leading organizations accelerate digital transformation through innovation, expertise, and agility. With more than 16,500 professionals across around 60 offices in over 20 countries, we combine deep industry knowledge with advanced capabilities in AI, cloud, cybersecurity, and data engineering.

Our regional teams, supported by strategic delivery centers, provide scalable, cost-efficient solutions tailored to local markets. Through our award-winning Synechron FinLabs accelerators and strategic partnerships with AWS, Microsoft, Databricks, Salesforce, and ServiceNow, we enable clients to innovate fast and lead with confidence. For more information on the company, please visit our website or LinkedIn community. Job Summary

Synechron is seeking an experienced Site Reliability Engineer (SRE) to enhance the stability, resilience, and operational maturity of our critical Financial Crime and Transaction Monitoring platforms. This role is vital in embedding SRE best practices across observability, automation, incident management, and production support. The successful candidate will be responsible for proactively managing service health, reducing operational risks, and supporting regulatory-critical services, thereby enabling the organization to deliver reliable, scalable, and compliant solutions aligned with business objectives.

Software Requirements

Required

Solid understanding and hands-on experience managing production-grade systems with high reliability and availability requirements

Expertise in SRE principles, monitoring, logging, alerting, and defining SLOs/SLA tuning

Proficiency with AWS services including EC2, S3, RDS, VPC, IAM, and CloudWatch (latest versions or equivalents)

Linux system administration and troubleshooting skills for enterprise environments

Experience with Oracle databases, including performance tuning, RAC, or RMAN in large data environments

Automation scripting skills using Python and Shell (Bash/sh) for operational automation

Experience with monitoring tools such as Prometheus, Grafana, ELK/EFK, and PagerDuty

Familiarity with CI/CD tools like Jenkins, GitLab CI, or AWS CodePipeline

Preferred

Knowledge of OFSAA, Oracle Rules Engine, or ML-enabled platform support (e.g., TRACE)

Infrastructure-as-Code tools such as CloudFormation or Terraform

Experience with support for high-performance Oracle environments (performance tuning, RAC, RMAN)

Exposure to cloud-native and containerized environments (Kubernetes, Docker)





Overall Responsibilities

Improve the reliability, availability, and recoverability of Financial Crime and Transaction Monitoring platforms.

Define, monitor, and manage SLIs/SLOs to proactively ensure service health and detect anomalies.

Provide Level 1 and Level 2 support for AWS and Oracle-based platforms, handling incident resolution and root cause analysis.

Build and sustain automation solutions for monitoring, logging, alerting, and operational workflows to reduce manual toil.

Lead incident response activities, conduct post-incident reviews, and implement preventative measures.

Develop, operate, and enhance CI/CD pipelines and infrastructure automation across environments.

Collaborate with engineering teams to design scalable, resilient, and secure systems; participate in capacity planning and performance tuning.

Support deployment, patching, and configuration changes, ensuring compliance with policies and standards.

Maintain comprehensive documentation of operational procedures, configurations, and incident resolutions.

Lead continuous process improvements to enhance system reliability, operational efficiency, and compliance adherence.

Technical Skills (By Category)

Systems & Support (Essential):

Enterprise-level system operation and support for AWS and Oracle environments

Linux system administration and troubleshooting

Incident management and escalation procedures

Monitoring & Automation (Essential):

Monitoring and alerting using Prometheus, Grafana, ELK/EFK, CloudWatch

Automation scripting with Python and Shell for operational tasks and event handling

Cloud & Infrastructure (Preferred):

Cloud deployment, scaling, and management (AWS, Azure, GCP)

Infrastructure-as-Code (Terraform, CloudFormation)

Databases/Data Management (Essential)

Oracle database management, performance tuning, and recovery

Data extraction and validation for high-volume transactional data

Development Tools & Methodologies (Essential):

Jenkins, GitLab CI, AWS CodePipeline for CI/CD pipelines

Version control with Git

Experience Requirements

Minimum of 8+ years supporting high-availability, mission-critical enterprise systems, particularly in financial services or comparable regulated environments.

Proven experience supporting Oracle databases, Oracle RAC, or RMAN in a high-volume context.





Strong background in enterprise support for Financial Crime and Transaction Monitoring platforms.

Demonstrated ability to lead operational support teams, manage incident escalations, and implement automation solutions.

Experience in cloud-native architecture, infrastructure automation, and observability tools.

Support experience working under regulatory and audit constraints is preferred.

Day-to-Day Activities

Monitor platform dashboards, logs, and alerts to ensure system health and performance.

Troubleshoot and resolve incidents related to operational, performance, or security issues proactively.

Conduct root cause analysis, document incident reports, and lead corrective action plans.

Automate routine operational tasks, alerts, and workflows to improve efficiency.

Collaborate with platform engineers, developers, and security teams on change management and capacity planning.

Participate in on-call rotations, incident reviews, and readiness exercises.

Continuously evaluate and recommend tools, procedures, and automation that improve reliability and reduce manual intervention.

Maintain detailed documentation of configurations, procedures, and lessons learned.

Qualifications

Bachelor’s degree in Computer Science, Engineering, or a related discipline.

8+ years supporting enterprise-scale, high-availability systems with operational excellence focus.

Experience supporting regulatory-critical platforms in financial services, especially in Fraud, Risk, or Transaction Monitoring.

Certifications in cloud platforms (AWS Certified Solutions Architect, Azure) and SRE foundations (Google SRE or equivalent) are advantageous.

Proven track record of automation, incident management, and operational improvements.

Professional Competencies

Critical thinking and analytical skills to diagnose and resolve complex operational issues.

Leadership and team management skills to guide operational teams and support team development.

Effective communication for stakeholder reporting, incident updates, and cross-team collaboration.

Ability to work under pressure, prioritize multiple tasks, and meet strict SLAs.

Adaptability to evolving technology landscapes and regulatory requirements.

Focus on continuous improvement, automation, and operational excellence. To expedite the application process, please share the following details at your earliest convenience:

Tentative date to join (if selected)

Current salary

Expected salary

Total experience

Relevant experience

Official email confirmation of notice period or last working day

Primary skills (hands-on)

Secondary skills

Reason for change

Current location

Preferred location

For More information contact – [email protected] OR DM

📌 Site Reliability Engineer (Bengaluru)
🏢 Synechron
📍 Bengaluru

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: site reliability engineer (bengaluru) / bengaluru

Subscribe to this job alert:

Get the latest job offers by email for: site reliability engineer (bengaluru) / bengaluru