13 Sep
|
The Adlakha Group
|
Pune
13 Sep
The Adlakha Group
Pune
Job Summary
We are looking for a Data Center Engineer with 3–4 years of hands-on experience in Data Center operations, Virtual Machine (VM) management, Linux/Unix administration, and DC-DR (Data Center–Disaster Recovery) environments.
The candidate will be responsible for day-to-day infrastructure operations, VM provisioning and management, Linux server administration, system health monitoring, troubleshooting, backup/restore coordination, DR activities, and supporting application/deployment teams.
Key Responsibilities
- Manage and support Data Center and Disaster Recovery (DC-DR) infrastructure.
- Provision, configure, modify, and decommission Virtual Machines (VMs) based on project requirements.
- Perform VM activities such as:
- VM creation and deletion
- CPU, RAM, and disk allocation/extension
- Snapshot creation and restoration
- VM cloning
- Network configuration
- VM migration
- Backup and restore coordination
- Administer Linux/Unix servers, particularly Ubuntu and enterprise Linux distributions.
- Perform Linux system administration activities including:
- User and group management
- File system and disk management
- LVM configuration and extension
- Mount point management
- Package installation and updates
- Service/process management
- Permissions and ownership management
- OS patching and maintenance
- Troubleshoot server issues related to CPU, memory, disk, network, processes, services, and file systems.
- Monitor infrastructure availability, capacity, utilization, and server health.
- Support DC-to-DR replication, DR drills, failover, failback, and disaster recovery testing.
- Validate application and infrastructure availability during planned DR exercises.
- Perform regular health checks of servers and virtualization infrastructure.
- Coordinate with Network, Security, Storage, Database, Application, and DevOps teams for infrastructure-related activities.
- Support application deployment teams with server prerequisites, ports, storage, users, permissions, and connectivity.
- Troubleshoot network connectivity using tools such as ping, traceroute, telnet, nc, curl, ss, and netstat.
- Perform log analysis and troubleshoot Linux system/service failures.
- Manage server storage, disk utilization, NFS mounts, and file-system-related issues.
- Participate in incident, change, problem, and service-request management processes.
- Maintain infrastructure inventory and technical documentation.
- Follow organizational security, access control, backup, patching, and compliance standards.
- Provide support during planned maintenance activities, production deployments, migrations, and critical incidents.
Required Technical Skills
Operating Systems
- Ubuntu Linux
- RHEL / CentOS / Rocky Linux or equivalent
- Good understanding of Unix/Linux concepts
Virtualization
- Hands-on experience in VM provisioning and lifecycle management
- VMware / vSphere / ESXi, KVM, or similar virtualization platforms
- VM snapshots, cloning, migration, resource allocation, and troubleshooting
Linux Administration
- User/group and permission management
- LVM and filesystem management
- Disk partitioning and mount management
- Systemd and service management
- Package management using apt, yum, and dnf
- SSH configuration and troubleshooting
- Cron jobs and basic shell scripting
- Performance and log troubleshooting
Data Center / DR
- Understanding of Data Center infrastructure and operations
- Hands-on exposure to DC and DR environments
- DR drills, failover/failback, and recovery validation
- Backup and restoration concepts
- Server availability and capacity monitoring
Networking
- TCP/IP fundamentals
- DNS, Gateway, Routing, VLAN, Firewall, and Proxy concepts
- Basic understanding of ports and network connectivity troubleshooting
- Experience troubleshooting server-to-server and application connectivity
Valuable to Have
- Basic knowledge of AWS/Azure or private cloud environments
- Knowledge of Docker and Kubernetes
- Basic Bash/Shell scripting
- Exposure to monitoring tools such as Zabbix, Nagios, Prometheus, Grafana, or similar
- Understanding of ITIL processes
- Exposure to ticketing tools such as ServiceNow or Jira
- Knowledge of backup and storage technologies
- Understanding of security tools such as EDR, vulnerability scanners, and SIEM agents
📌 Data Center Engineer (Pune)
🏢 The Adlakha Group
📍 Pune