AWS Engineer (Bengaluru)

AWS Engineer (Bengaluru)

09 Oct
|
Lumos Learning
|
Bengaluru

09 Oct

Lumos Learning

Bengaluru

AWS Systems Engineer (Site Reliability & Observability)
Lumos Learning · Bangalore (on-site)

Role overview

Lumos Learning is hiring an AWS Systems Engineer with 3–5+ years of hands-on AWS experience to keep our platform healthy, observable and quick for the students and educators who rely on it every day. You will own cloud health monitoring in Amazon CloudWatch, run our Amazon RDS Serverless databases, and lead the selection and rollout of a new application performance monitoring (APM) tool that lets engineering teams find and fix performance problems quickly. When something is slow or broken, you find the cause and fix it, working with developers on application-level issues.

Key responsibilities

Monitoring, observability and incident response

- Design, build and maintain CloudWatch dashboards, metrics, composite alarms, log groups, Logs Insights queries and synthetic canaries to give clear, real-time health visibility across AWS services.
- Define SLIs, SLOs and alert thresholds with engineering teams, and tune alerts to cut noise and false positives.
- Evaluate, select and implement a new application monitoring tool (distributed tracing, service maps, real-user and infrastructure monitoring), and instrument key applications so teams can troubleshoot performance issues end to end.
- Build runbooks and troubleshooting guides, and train developers and support staff to use the current monitoring tools effectively.
- Lead root-cause analysis for incidents and performance degradations, participate in on-call rotation, and drive post-incident reviews to completion.

Databases

- Administer Amazon RDS and Aurora Serverless databases: capacity (ACU) tuning, scaling behavior, backups, point-in-time recovery, patching, parameter groups and Performance Insights.
- Identify slow database queries and work with developers to improve and optimize them.

Infrastructure, automation and cost





- Automate provisioning and configuration with infrastructure as code (Terraform or CloudFormation) and scripting, with changes reviewed through version control.
- Manage core AWS infrastructure: EC2, VPC networking, load balancers, IAM, S3 and Auto Scaling, following least-privilege and security best practices.
- Monitor and optimize cloud cost, capacity and performance, and report on trends to engineering leadership.
- Support backup, disaster recovery and high-availability planning, including regular recovery testing.

Required qualifications

- 3–5+ years of hands-on experience as a systems, cloud, DevOps or site reliability engineer, with at least 3 years working primarily in AWS.
- Strong working knowledge of Amazon CloudWatch: metrics, alarms, dashboards, Logs Insights, agent configuration and EventBridge integrations.
- Hands-on experience running Amazon RDS and Aurora Serverless (v2 preferred), including scaling, performance tuning, backup and recovery.
- Experience implementing or rolling out an APM or observability platform (for example Datadog, New Relic, Dynatrace, Grafana or AWS X-Ray/OpenTelemetry) and using it to troubleshoot application performance.
- Solid understanding of AWS core services: EC2, VPC, IAM, S3, ELB and Route 53.
- Proficiency with infrastructure as code (Terraform or CloudFormation) and scripting in Python, Bash or PowerShell.
- Working knowledge of Linux administration and TCP/IP, DNS and HTTP troubleshooting.
- Experience with incident response, root-cause analysis and on-call support.
- Clear written and verbal communication, including documenting systems and explaining technical issues to non-technical colleagues.





Preferred qualifications

- AWS certification such as Solutions Architect Associate or Professional, SysOps Administrator or DevOps Engineer Professional.
- Experience with containers and orchestration (ECS, EKS, Docker) and serverless services such as Lambda.
- Familiarity with CI/CD pipelines (GitHub Actions, GitLab CI, Jenkins or AWS CodePipeline).
- Experience with OpenTelemetry instrumentation and centralized log management.
- Knowledge of AWS security services and compliance practices (GuardDuty, Security Hub, Config, CloudTrail), ideally with student data privacy requirements such as FERPA or COPPA.
- Experience with AWS cost optimization and FinOps practices.
- Background in an education technology or SaaS environment with seasonal traffic spikes, such as back-to-school peaks.
- Bachelor's degree in computer science or a related field, or equivalent practical experience.

Tools and working style

- Monitoring and observability: CloudWatch, X-Ray, Logs Insights, new APM platform (to be selected)
- Databases:Amazon RDS, Aurora Serverless, Performance Insights
- Infrastructure as code: Terraform or CloudFormation, Git Compute and network
- : EC2, VPC, ELB, Route 53, Auto Scaling
- Automation: Python, Bash, Lambda, EventBridge

The role suits someone who is curious about why systems slow down, stays calm during incidents, writes things down so others can follow, and works well with developers, QA and support teams.

Pay: ₹1,200,000.00 - ₹2,000,000.00 per year

Application Question(s):

- How many years of hands-on experience do you have as a systems, cloud, DevOps or site reliability engineer?

- How many of those years were primarily in AWS?
- Are you comfortable working on-site in Bangalore?
- What is your notice period?
- What is your current CTC?
- What is your expected CTC?
- Do you have any experince working in Edtech?

Work Location: In person

📌 AWS Engineer (Bengaluru)
🏢 Lumos Learning
📍 Bengaluru

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: aws engineer (bengaluru) / bengaluru