04 Sep
|
Teciem
|
Bengaluru
Join Teciem, a global fintech company delivering cutting-edge Treasury & Capital Markets software solutions to financial institutions worldwide.
We're looking for a talented Site Reliability Engineer who is passionate about building resilient, scalable, and high-performing systems. If you thrive in a fast-paced setting and enjoy solving complex infrastructure challenges, we'd love to hear from you.
Required Skills & Experience
SRE & Production Operations (mandatory)
- 5+ years in a Site Reliability Engineer or production operations engineering role, operating a SaaS or always-on service at scale.
- Proven hands-on experience defining and operating against SLIs, SLOs, error budgets, and customer SLAs.
- Demonstrated incident management experience: 24/5 on-call, structured root-cause analysis, blameless post-mortems.
- Strong software engineering fundamentals and proficiency in at least one scripting/programming language (Python, Go, or Bash) for automation and operational tooling.
AWS & Kubernetes (mandatory)
- Hands-on experience operating production workloads on AWS: Amazon EKS, EC2, IAM, VPC, S3, RDS, Route 53, CloudWatch.
- Solid experience with Kubernetes and Helm for orchestration of containerised workloads in production.
- Working knowledge of Kubernetes networking (CNI, CoreDNS, NetworkPolicy) and Linux/Unix operating systems.
Observability
- Proficiency with Prometheus and Grafana (dashboards, recording rules, alerting).
- Experience with OpenTelemetry, distributed tracing (Jaeger or Tempo), and log aggregation (Loki or OpenSearch/ELK).
- Ability to define and instrument meaningful SLIs from application and infrastructure telemetry.
IaC & GitOps
- Experience with Terraform (modules, remote state, AWS provider) for infrastructure changes in production.
- Familiarity with ArgoCD or equivalent GitOps tooling for continuous delivery and drift detection.
Security & Compliance
- Knowledge of best practices for data encryption (KMS, TLS/mTLS), secrets management, and least-privilege IAM in production.
- Awareness of audit and compliance requirements for financial services (SoD, change records, DORA, ISO 27001).
Nice to Have
- Experience operating a service mesh (Istio) for traffic management, mTLS, and fine-grained observability.
- Experience with Kyverno or OPA for admission control and compliance guardrails in Kubernetes.
- Experience with HashiCorp Vault or AWS Secrets Manager / KMS for secrets management at scale.
- Familiarity with FinOps tooling and cloud cost optimisation for a multi-tenant SaaS platform.
- Experience with chaos engineering and disaster recovery practices (LitmusChaos, Gremlin, GameDay exercises).
- Exposure to treasury or capital markets systems (Kondor, Summit, or equivalent TMS/risk platforms).
- AWS certifications: Solutions Architect Professional, DevOps Engineer Professional, or SysOps Administrator.
- Kubernetes certifications: CKA (Certified Kubernetes Administrator) or CKAD.
Education
Engineering degree (Computer Science, Information Technology, Mathematics, or equivalent) or equivalent experience
Work model
- Hybrid - Bangalore (Teciem India office) + remote
- Participation in a 24/5 on-call rotation is a core requirement of this role
Why Teciem?
✔ Fast-growing global fintech organization
✔ Exposure to leading financial institutions worldwide
✔ Opportunity to work on innovative Treasury & Capital Markets solutions
✔ Collaborative and international work environment
✔ Flexible hybrid working model
✔ Excellent learning and career development opportunities
📌 Senior Site Reliability Engineer (Bengaluru)
🏢 Teciem
📍 Bengaluru