24 Aug
|
TerraMD
|
Hyderabad
Role & responsibilities
- Manage and maintain Linux servers across production, staging, development and experimental environments.
- Keep infrastructure organized, documented and reproducible, including server inventory, access credentials, DNS, certificates, networking and service ownership.
- Deploy, configure and maintain Docker-based applications, databases, reverse proxies and supporting services.
- Implement monitoring, alerting, logging and uptime checks across all critical systems.
- Establish and maintain reliable backup, restore and disaster-recovery procedures, including regular recovery testing.
- Manage operating-system updates, security patches, SSH access, firewalls and server hardening.
- Automate repetitive administration work using tools such as Ansible, Terraform, shell scripting or equivalent.
- Maintain clear separation between production systems and experimental/development infrastructure.
- Monitor disk usage, memory, CPU, network traffic and infrastructure costs and proactively resolve issues.
- Troubleshoot server, networking, deployment and performance problems.
- Maintain documentation so another administrator can understand and recover the workplace without relying on tribal knowledge.
- Work closely with developers while taking primary responsibility for infrastructure stability, cleanliness and operational discipline.
- Help standardize deployments and gradually eliminate manually configured or undocumented infrastructure.
- Participate in incident response and perform root-cause analysis when systems fail.
- Maintain databases and storage infrastructure,
including PostgreSQL and other data services, without requiring application-development ownership.
Preferred candidate profile
- 36 years of hands-on Linux system administration, DevOps or infrastructure experience.
- Very comfortable administering Linux servers from the command line.
- Strong understanding of networking, DNS, TLS/SSL, SSH, firewalls, HTTP and reverse proxies.
- Practical experience with Docker and containerized production environments.
- Experience with Nginx, Caddy, HAProxy or similar technologies.
- Good understanding of backups, replication, disaster recovery and infrastructure security.
- Experience with monitoring tools such as Prometheus, Grafana, Uptime Kuma, Zabbix or similar.
- Familiarity with infrastructure automation such as Ansible, Terraform or shell scripting.
- Experience managing dedicated servers, VPS providers and cloud infrastructure.
- Comfortable working with Git and CI/CD pipelines.
- PostgreSQL/Linux database administration experience is a strong advantage.
- Kubernetes experience is helpful but not required.
- Strong organizational habits and an instinct for documentation, standardization and automation.
- Someone who notices an undocumented server, expiring certificate or missing backup and wants to fix it before being asked.
- Comfortable supporting a fast-moving engineering environment while being conservative and disciplined with production systems.
- We particularly value candidates who enjoy making infrastructure boring, predictable and reliable rather than constantly introducing new technologies.
📌 System Administrator (Hyderabad)
🏢 TerraMD
📍 Hyderabad