06 Aug
|
Uffizio
|
Ahmedabad
About Uffizio
Uffizio Technology is a SaaS product company based in Valsad, India, and part of the Uffizio Group. For over 25 years,
we've been building software platforms trusted by B2B customers in more than 100 countries.
We develop innovative SaaS products based on telematics and IoT solutions, helping businesses improve productivity,
operational efficiency, and sustainability. Our products are offered as both white-label solutions for technology partners and ready-to-deploy platforms for businesses worldwide. We foster a collaborative, innovation-driven culture where talented people build impactful technology and grow their careers.
About The Role
We are looking for a DevOps Team Lead who can build and scale a highly reliable infrastructure for globally deployed
SaaS platforms. This is not a maintenance-only role.
The expectation is to architect systems that improve uptime, deployment speed, scalability, observability, security, and operational efficiency across multiple products and business units.
The role requires someone who can think beyond servers and CI/CD pipelines. We need a systems thinker who understands how infrastructure impacts customer experience, release velocity, support load, product scalability, and business growth.
You'll own the automation, deployment, and reliability of our systems end to end — from CI/CD pipelines to production monitoring — and work closely with development and product teams to ship faster without compromising stability.
The candidate will lead the DevOps function across multiple SaaS products handling:
- Real-time telematics workloads
- IoT device communication
- High-ingestion APIs
- Live tracking systems
- Video and sensor-based platforms
- Multi-tenant SaaS deployments
Key Responsibilities
Infrastructure & Cloud Management
- Design, provision, and manage cloud and on-premise infrastructure (AWS) using Infrastructure as Code (Terraform,
CloudFormation, or Ansible).
- Ensure high availability, scalability, redundancy, and disaster recovery planning.
Manage
Linux-based production
environments.
- Optimize infrastructure cost without compromising reliability — through right-sizing, reserved capacity planning,
and usage monitoring.
- Handle scaling strategies for increasing device load and customer growth.
CI/CD & Release Engineering
- Build and maintain robust CI/CD pipelines (Jenkins, GitLab CI, GitHub Actions, or similar) to automate build, test,
deployment, rollback, and environment provisioning processes.
- Reduce deployment risks and deployment time.
- Standardize deployment practices across teams and products.
Monitoring & Reliability
- Establish strong monitoring, logging,
and alerting/observability systems (CloudWatch, Prometheus, Grafana,
ELK/EFK stack, Zabbix, Datadog).
- Implement proactive incident detection and root cause analysis.
- Reduce downtime and improve platform stability.
- Drive SRE-oriented operational maturity.
- Participate in on-call rotation, lead incident response, and drive root-cause analysis and post-mortems.
Security & Compliance
- Implement infrastructure security best practices — IAM policies, network security groups, secrets management,
SSL, firewall policies, backups, and vulnerability handling.
- Manage access control and compliance requirements.
- Ensure infrastructure hardening and operational compliance.
Containerization & Orchestration
- Deploy, manage, and scale containerized workloads using Docker and Kubernetes (EKS/AKS/GKE or self-managed
clusters).
- Improve deployment consistency and environment portability.
- Support microservices architecture where applicable.
Database & Performance Optimization
Work Closely With Backend And Database Teams On
- Performance tuning
- Query optimization support
- Load balancing
- Caching strategies
- Replication and failover systems
Team Leadership
- Lead and mentor DevOps engineers.
- Create operational SOPs and infrastructure standards.
- Build accountability, documentation culture, and ownership within the team.
- Coordinate with Development, QA, Support, and Product teams to improve deployment velocity, system reliability,
and developer experience.
Incident Management & Business Continuity
- Handle production incidents with urgency and ownership.
- Build escalation systems and incident response frameworks.
- Conduct postmortem analysis and preventive planning.
- Maintain disaster recovery, backup, and business continuity plans for production systems.
Required Technical Skills
Solid Expertise In
- Linux Server Administration (patching, configuration management, access control)
- AWS / GCP / Azure
- Infrastructure as Code (Terraform, CloudFormation, or Ansible)
- Docker & Kubernetes (EKS/AKS/GKE or self-managed clusters)
- Jenkins / GitHub Actions / GitLab CI
- Nginx / Apache, Load Balancers & Reverse Proxies
- Networking & Security (DNS, VPN, firewalls)
- Monitoring Tools (Prometheus, Grafana, ELK/EFK, CloudWatch, Zabbix,
Datadog)
- Infrastructure Automation
- Shell Scripting / Python / Bash
Good Understanding Of
- High-availability architecture
- Distributed systems
- Scaling real-time applications
- Database replication and clustering
- Message brokers (RabbitMQ, Kafka, Redis Streams, etc.)
- API infrastructure
- SSL, DNS, VPN, CDN, WAF
- Cloud security best practices (IAM, encryption, secrets management)
Nice to Have
- Experience in IoT or telematics platforms
- Experience managing large-scale real-time tracking systems
- SRE practices
- Cost optimization at scale
- Multi-region deployment experience
LEADERSHIP EXPECTATIONS
This role is not for someone who only executes tickets.
We expect the person to:
- Think proactively instead of reactively
- Build systems before problems become incidents
- Create operational leverage through automation
- Reduce dependency on manual intervention
- Build infrastructure that supports aggressive business growth
- Create visibility and measurable operational KPIs
KPIS / SUCCESS METRICS The DevOps TL Will Be Evaluated On
- Platform uptime
- Deployment frequency & stability
- MTTR (Mean Time to Recovery)
- Infrastructure scalability
- Security incident reduction
- Alert quality and monitoring maturity
- Automation coverage
- Infrastructure cost efficiency
- Team efficiency and operational discipline
Experience Required
- 8 years in DevOps / Infrastructure Engineering / Site Reliability Engineering / Cloud Infrastructure roles.
- 2 years leading teams or handling critical production infrastructure.
- Experience managing production SaaS environments at scale.
- Strong working knowledge of at least one major cloud platform (AWS), with proven hands-on experience in
Terraform, Docker, and Kubernetes.
IDEAL CANDIDATE PROFILE
We are not looking for a server administrator.
We are looking for someone who:
- Understands business impact of infrastructure decisions
- Can scale systems under uncertainty
- Handles pressure calmly during outages
- Builds processes, not heroics
- Has a strong ownership mindset
- Can challenge poor engineering practices
- Thinks in terms of reliability engineering, not firefighting
WHY THIS ROLE MATTERS For most SaaS companies, DevOps becomes a support function. For us, it is a growth constraint or growth accelerator.
A Weak DevOps Team Creates
- Slow releases
- Customer dissatisfaction
- Downtime
- Engineering bottlenecks
- Support overload
- Revenue risk
A strong DevOps function compounds the effectiveness of every other department. That is why this role is strategically important
📌 DevOps Team Lead (Ahmedabad)
🏢 Uffizio
📍 Ahmedabad