Position: Site Reliability Engineer (SRE) – Linux
Experience: 3+ Years
Location: Ahmedabad
Work Mode: Work From Office
Shift: Rotational Shifts – Day/Night
About the Role
We are looking for an experienced Site Reliability Engineer (SRE) with strong expertise in Linux system administration, troubleshooting, and production infrastructure support . The ideal candidate should have hands-on experience managing Linux servers, resolving critical production issues, monitoring system performance, and maintaining highly available and reliable environments.
Key Responsibilities
- Strong hands-on administration and troubleshooting of Linux servers .
- Monitor Linux systems, applications, services, CPU, memory, disk, network, and system performance.
- Troubleshoot and resolve production-level Linux issues and system failures.
- Perform server configuration, patching, upgrades, and maintenance.
- Analyze system logs and identify the root cause of incidents.
- Manage Linux processes, services, users, permissions, filesystems, and networking.
- Handle incident management, problem management, and root cause analysis (RCA).
- Ensure high availability, reliability, and performance of production infrastructure.
- Automate repetitive operational tasks using Shell/Bash scripting .
- Work with monitoring and alerting tools to identify and resolve issues proactively.
- Support deployment, release, and infrastructure-related activities.
- Collaborate with DevOps, Development, and IT teams to resolve infrastructure and application issues.
Required Skills
- 3+ years of strong hands-on experience in Linux Administration/SRE.
- Strong Linux knowledge is mandatory.
- Excellent troubleshooting skills in Linux/Unix environments .
- Strong understanding of Linux commands, processes, services, permissions, filesystems, and logs.
- Good knowledge of Linux networking – TCP/IP, DNS, HTTP/HTTPS, SSH, ports, routing, etc.
- Strong experience in system monitoring and performance troubleshooting.
- Hands-on experience with Shell/Bash scripting .
- Valuable understanding of production infrastructure and incident management.
- Experience with monitoring/logging tools such as Prometheus, Grafana, ELK, Nagios, Zabbix , or similar tools.
- Basic to good knowledge of Cloud platforms (AWS/Azure) .
- Understanding of Docker/Kubernetes is an added advantage.
- Familiarity with CI/CD and DevOps practices is a plus.
Interested Candidate Share Your CV on
[email protected] or Apply Here
📌 Senior Site Reliability Engineer (Ahmedabad)
🏢 Swaraa Tech Solutions
📍 Ahmedabad