04 Sep
|
Sailpoint Technologies
|
India
04 Sep
Sailpoint Technologies
India
Job Summary
Job Title: Senior DevOps Engineer
Imagine a day where you transition from resolving an active system bottleneck to writing code that ensures nobody ever has to manually handle that operational task again. As a Senior DevOps Engineer on our team, you will spend your days turning manual runbooks into elegant, event-driven automations. You will design, build, and deploy production-grade software using Python and JavaScript/TypeScript. This role sits at the intersection of robust software development and systems reliability, meaning you will regularly evaluate blast radius, implement least-privilege IAM policies, and ensure every automated action is traceable and auditable.
Additionally, you will support our core database and streaming platforms in production and participate in the teams on-call rotation.
Responsibilities
- Turn manual runbooks into elegant, event-driven automations.
- Design, build, and deploy production-grade software using Python and JavaScript/TypeScript.
- This role sits at the intersection of robust software development and systems reliability, meaning you will regularly evaluate blast radius, implement least-privilege IAM policies, and ensure every automated action is traceable and auditable.
- You will join the Core Data Infrastructure team. We are responsible for running, maintaining, and scaling the critical data stores, distributed databases, and messaging systems that power our global Identity Security Cloud platform. Our technical ecosystem includes PostgreSQL, MySQL, Redis, Elasticsearch, DynamoDB, and Kafka.
- 30 Days Complete your engineering onboarding, gain a solid conceptual understanding of our cloud-native event-driven microservices architecture, and begin shadowing on-call engineers.
- You will identify your first high-frequency manual runbook candidate for automated remediation.
- 60 Days Design, build,
and deploy your first event-driven automation into our production environment using Python/TypeScript and AWS. You will also begin fully participating in the teams production on-call rotation.
- 6 Months Codify at least three major operational runbooks into safe, self-service or auto-triggered systems. You will establish shared deployment patterns and lead the team in utilizing LLM-augmented development tools like Cursor and Claude with proper code-review guardrails.
- 1 Year Own the automation roadmap for the Data Infrastructure team, leading to a measurable 50% reduction in repetitive operational tasks. You will partner closely with SRE, Security, and product teams to champion event-driven, self-healing systems across our cloud environment.
Requirements
- Bachelors degree in related field, or equivalent qualified experience
- 5+ years of qualified software engineering experience, with meaningful time spent building or operating production systems at a SaaS or cloud service provider
- Experience in 24x7 production operations for a highly available SaaS or cloud environment, including on-call responsibilities
- Strong programming ability in Python and JavaScript/TypeScript, with the ability to own services end-to-end from design to production
- Deep hands-on experience with AWS cloud infrastructure, especially Lambda, EventBridge, IAM, CloudWatch, and CloudTrail
- Experience deploying, operating, and troubleshooting databases and distributed data stores in production (such as PostgreSQL, MySQL, Redis,
Elasticsearch, DynamoDB, or Kafka)
- Experience designing event-driven architectures and monitoring-triggered automation rather than just scheduled scripts
- Solid working knowledge of Kubernetes in production, including deployments, RBAC, controllers, and safe operational actions
- Experience with Prometheus and alert-driven workflows (Alertmanager, webhook consumers, or equivalent)
- Strong security instincts, including least-privilege IAM design, secrets handling, and blast-radius reasoning
- Track record of building for auditability, including structured logs, correlation IDs, and immutable change records
- Demonstrated use of LLM-augmented development tools such as Cursor and Claude, with a point of view on how to review their output responsibly
- CI/CD experience (such as GitHub Actions or Jenkins) and comfort owning the pipeline for your services
- Strong understanding of systems, networking, and distributed systems troubleshooting
- Strong written communication, with the ability to turn complex operational incidents into clear runbooks and code
- Ability to operate autonomously across time zones and collaborate with a globally distributed team
Preferred Skills
- Experience replacing manual toil at scale with a measurable reduction in alerts, tickets, or MTTR
- Background in SRE, DevOps, or platform engineering for a large multi-tenant SaaS environment
- Familiarity with infrastructure-as-code (Terraform) and GitOps patterns
- Experience with data pipeline tooling (such as Apache Airflow) and streaming platforms (Kafka)
Disclaimer: This job posting & Location has been aggregated from external source. Role details, content, and availability are subject to change. Applicants are advised to confirm the latest information directly on the company website before applying.
📌 Sr DevOps Engineer (India)
🏢 Sailpoint Technologies
📍 India