29 Aug
|
Angel One
|
Bengaluru
29 Aug
Angel One
Bengaluru
About Angel One
Angel One is one of India’s fastest growing fintechs, on a bold mission to make investing simple, smart, and inclusive for every Indian. With over 3 crore clients, we’re building at scale – and building for impact. Our Super App helps clients manage their investments, trade seamlessly, and access financial tools tailored to their goals. We are working to build personalized financial journeys for our clients, powered by new-age tech, AI, Machine Learning and Data Science.
We're a builder's company at heart. You’ll have the space to experiment, the freedom to move with velocity, and the mandate to make bold, user-first decisions – every single day.
The vibe? Think less hierarchy, more momentum. Everyone has a seat at the table and a shot to build something that lasts. Be part of a team that’s scaling sustainably, thinking big, and building for the next billion.
Why You'll Love Working at Angel One!
- Tech Systems that run at Scale: From AI to real-time data infra, you’ll work on tech that’s ahead of the curve and solve problems that truly matter.
- Build one of India’s Leading Fintech Platform: We’re not just disrupting finance – we’re shaping how billion Indians access wealth.
- Own It. Drive It. Scale It: You’ll have the freedom to lead, the resources to build, and the opportunity to leave your mark.
- Empowered Growth: We invest in your growth and empower you to explore your full potential.
- Exceptional Benefits: Our comprehensive benefits package includes health insurance, wellness programs, learning & development opportunities, and more.
Job Title: SRE 2 /3
Location:Bangalore
What you will do:
- Architect, design and build highly available, scalable, resilient and secure platforms for mission-critical services.
- Define and drive reliability engineering strategy, standards, SLOs/SLIs, error budgets and operational excellence across services and platforms.
- Lead the architecture and evolution of Kubernetes, cloud and platform infrastructure at scale.
- Build scalable automation, self-healing systems and developer platforms that significantly reduce operational toil.
- Drive complex technical initiatives across multiple teams, including platform modernization, cloud migration, reliability and resilience improvements.
- Provide technical leadership during critical incidents, drive effective RCA and ensure systemic corrective actions and permanent resolution.
- Design and lead disaster recovery, business continuity,
resilience and failure-testing strategies for critical systems.
- Partner with engineering, architecture, security and product teams throughout the lifecycle to ensure reliability, scalability, performance and security are built into products.
- Establish and evolve Infrastructure as Code, CI/CD and GitOps practices across the organization.
- Drive observability strategy, including metrics, logs, traces, SLOs, capacity planning and proactive detection of reliability risks.
- Identify and deliver significant cloud cost, capacity and performance optimization opportunities without compromising reliability.
- Evaluate and adopt emerging technologies, including AI/GenAI and AIOps, to improve engineering productivity, automation, incident response and operational intelligence.
- Provide technical mentorship and coaching to SRE-1/SRE-2 engineers and raise engineering standards across the organization.
- Influence architecture and technology decisions through design reviews, technical proposals and engineering best practices.
- Own and deliver multiple high-impact initiatives simultaneously while providing technical direction and removing execution blockers.
- Drive a culture of automation, continuous improvement, operational ownership and engineering excellence across SRE and development teams.
Who you are:
- Bachelor's degree or equivalent experience in Software Engineering, Computer Science or a related discipline.
- 4–8 years of experience in SRE, DevOps, Cloud, Platform Engineering or Software Engineering, with demonstrated technical leadership.
- Strong software engineering experience in Python, Go or equivalent, with ability to build production-grade platforms, automation and tooling.
- Expert understanding of distributed systems, scalability, high availability, fault tolerance, performance and reliability engineering.
- Strong hands-on expertise in Kubernetes/EKS, including architecture, cluster lifecycle, networking, security, scaling, upgrades and production operations.
- Strong expertise in Terraform/IaC, including reusable modules, large-scale infrastructure management,
governance and automation.
- Strong experience architecting and operating AWS cloud infrastructure and cloud-native platforms at scale.
- Solid expertise in observability, including metrics, logs, traces, SLOs/SLIs, distributed tracing, alerting and reliability analytics.
- Strong experience with CI/CD, GitOps, Jenkins/GitHub Actions/GitLab CI and automated software delivery.
- Strong understanding of cloud security, IAM/RBAC, networking, encryption, secrets management and secure infrastructure design.
- Proven experience designing DR, business continuity, resilience and multi-region/multi-AZ architectures for mission-critical systems.
- Strong understanding of capacity planning, performance engineering and cloud cost optimization.
- Excellent debugging, troubleshooting and incident leadership skills, with the ability to resolve ambiguous and highly complex production problems.
- Demonstrated ability to influence architecture and technical strategy beyond an individual team and drive cross-functional initiatives.
- Strong experience mentoring engineers, conducting design/architecture reviews and establishing engineering best practices.
- Ability to effectively evaluate and adopt AI/GenAI and AIOps to improve software engineering, automation, reliability and operational decision-making.
- Robust communication, stakeholder management and leadership skills with a high degree of ownership and autonomy.
- Demonstrated ability to operate as a technical leader and subject-matter expert, without requiring direct people-management responsibility.
What's in it for You?
- Flexible work model: Whether you're remote, hybrid, or in-office, we trust you to work where you thrive and deliver with impact.
- Empowered Growth: We invest in your growth and empower you to explore your full potential.
- Exceptional Perks: Our comprehensive benefits package includes health insurance, wellness programs, learning & development opportunities, and more.
At Angel One, our thriving culture is rooted in Diversity, Equity, and Inclusion (DEI).
As an Equal opportunity employer, we wholeheartedly welcome people from all backgrounds irrespective of caste, religion, gender, marital status, sexuality, disability, class or age to be part of our team. We believe that everyone's unique experiences and viewpoints make us stronger together. Come and be a part of #OneSpace*, where your individuality is celebrated and embraced.
📌 Site Reliability Engineer (Bengaluru)
🏢 Angel One
📍 Bengaluru