Server & Infrastructure Administrator
Mumbai, Lower Parel
Key Responsibilities
A. Physical Server Administration
- Install, configure, and maintain physical servers across data centers and branch locations.
- Manage server hardware including CPU, RAM, RAID, disks, NICs, HBA,
firmware and BIOS.
- Monitor server health, hardware alerts, capacity, performance, and availability.
- Coordinate hardware replacement, warranty, AMC, and vendor support.
- Perform server firmware, BIOS, driver, and hardware lifecycle upgrades.
- Maintain server documentation, inventory, configuration records, and asset information.
- Troubleshoot hardware and operating-system-related issues.
- Ensure appropriate redundancy and high-availability configurations.
B. VMware / Virtualization Administration
- Administer and maintain VMware vSphere / ESXi / vCenter / VMware Cloud
Foundation (VCF) environments.
- Manage virtual machines, clusters, resource pools, templates, snapshots, and virtual networking.
- Perform VM provisioning, migration, resizing, cloning, and decommissioning.
- Monitor CPU, memory, datastore, network, and VM performance.
- Manage ESXi host lifecycle, upgrades, patches, and firmware compatibility.
- Troubleshoot virtualization and VM performance issues.
- Maintain high availability, resource utilization, and capacity planning.
- Support VMware upgrades and infrastructure lifecycle management.
C. SAN Storage Administration
- Administer enterprise SAN storage infrastructure and associated components.
- Manage LUNs, storage pools, volumes, RAID, host groups, mappings and capacity.
- Monitor storage capacity, IOPS, latency, throughput, and performance.
- Troubleshoot storage connectivity and performance issues.
- Manage SAN zoning and connectivity in coordination with network teams.
- Perform storage provisioning and reclamation based on business requirements.
- Plan storage capacity and lifecycle upgrades.
- Maintain storage configuration and operational documentation.
D. AWS Cloud Infrastructure
- Administer and support AWS infrastructure across multiple accounts/environments.
- Manage EC2, EBS, VPC, subnets, security groups, IAM, load balancers and related AWS services.
- Support AWS infrastructure deployed using AWS Control Tower and multi-
account architecture.
- Manage cloud compute, storage, networking, and availability.
- Monitor AWS resource utilization, health, performance, and cost.
- Implement appropriate security controls and least-privilege access.
- Support EKS and cloud-hosted application infrastructure as required.
- Manage backup and recovery for AWS workloads.
- Support migration of on-premises workloads to AWS.
- Participate in AWS architecture reviews, capacity planning, and optimization.
E. Backup & Disaster Recovery
- Administer Veeam Backup & Replication infrastructure.
- Configure and maintain backup jobs, repositories, retention policies, and backup schedules.
- Monitor daily backup status and investigate failures.
- Perform file-level, VM-level, and application-level restoration as required.
- Maintain backup infrastructure and repositories.
- Perform regular restore testing and backup verification.
- Participate in DR drills and business continuity exercises.
- Maintain and test RPO/RTO requirements.
- Support replication and cloud-based DR solutions.
- Ensure backup infrastructure follows security and compliance requirements,
including appropriate protection against ransomware.
F. Server Patch Management
- Manage operating-system patching for Windows and Linux servers.
- Plan and execute monthly patching cycles.
- Use enterprise endpoint/server management tools such as ManageEngine
Endpoint Central.
- Perform patch assessment, deployment, validation, and reporting.
- Coordinate patching activities with application and business teams.
- Maintain patch compliance dashboards and reports.
- Handle patch-related incidents and rollback activities.
- Ensure critical and security patches are prioritized based on risk.
G. Vulnerability Management / VAPT Remediation
- Review periodic VAPT, vulnerability assessment, and security scanning reports.
- Identify vulnerabilities affecting servers, operating systems, applications, and infrastructure components.
- Coordinate remediation with application, network, security, and infrastructure teams.
- Perform OS upgrades, configuration changes, security hardening, and patching required for remediation.
- Track vulnerabilities from identification through closure.
- Validate remediation and provide evidence for closure.
- Maintain vulnerability remediation reports and compliance records.
- Prioritize vulnerabilities based on CVSS, business criticality, exposure, and regulatory requirements.
- Support internal and external audits related to infrastructure security.
H. PIM / PAM Administration
- Administer and support Privileged Identity Management / Privileged Access
Management (PIM/PAM) solutions.
- Manage privileged accounts, service accounts, administrative accounts, and privileged credentials.
- Implement least-privilege and role-based access controls.
- Manage privileged access workflows and approvals.
- Configure and monitor privileged sessions.
- Support session recording and audit requirements.
- Perform periodic privileged-account reviews.
- Ensure administrative access is provided through approved PAM mechanisms.
- Support PAM onboarding of servers, applications, databases, network devices,
and cloud environments.
- Assist with PAM audits, access reviews, and compliance requirements.
I. Server Security & Hardening
- Implement server hardening based on organizational security standards and industry best practices.
- Maintain secure configurations for Windows and Linux servers.
- Manage endpoint/server security agents and security policies.
- Disable unnecessary services, ports, protocols, and accounts.
- Ensure appropriate firewall and access controls are implemented.
- Support security audits and compliance assessments.
- Coordinate with the cybersecurity team for security incidents and remediation.
J. Monitoring & Incident Management
- Monitor infrastructure availability and performance.
- Respond to infrastructure alerts and incidents.
- Troubleshoot server, VM, storage, backup, cloud, and operating system issues.
- Perform root-cause analysis for recurring incidents.
- Participate in major incident resolution and post-incident reviews.
- Maintain incident and problem records in the organization's ITSM platform.
- Develop and maintain operational runbooks and troubleshooting procedures.
K. Capacity & Performance Management
- Monitor infrastructure utilization and trends.
- Perform capacity planning for:
- CPU
- Memory
- Storage
- Network
- Virtual machines
- AWS resources
- Backup infrastructure
• Identify potential capacity bottlenecks before they impact production.
- Recommend infrastructure upgrades and optimization opportunities.
L. Change & Configuration Management
- Follow formal ITIL-based Change Management processes.
- Raise and implement changes through the organization's ITSM platform.
- Participate in CAB activities where required.
- Maintain configuration records and infrastructure documentation.
- Ensure production changes are properly planned, tested, approved,
implemented, and documented.
- Maintain rollback and recovery procedures for infrastructure changes.
1. Compliance & Documentation
• Maintain accurate infrastructure documentation, diagrams, asset inventories, and configuration records.
- Maintain SOPs, MOPs, and operational runbooks.
- Provide infrastructure evidence for internal, external, regulatory, and security audits.
- Ensure infrastructure activities comply with organizational policies and applicable regulatory requirements.
- Maintain records for patching, backups, vulnerability remediation, privileged access, and DR testing.
- Support implementation of IT controls and audit recommendations.
1. Experience
Experience: 5–8 years in IT Infrastructure / Server Administration. Candidates should have hands-on experience managing enterprise production infrastructure with responsibility for server availability, virtualization, storage, backup,
cloud infrastructure, security remediation, and privileged access.
Experience in a regulated financial-services setting will be an advantage.
1. Key Performance Indicators (KPIs)
• Server availability and uptime
- VMware infrastructure availability
- SAN availability and storage performance
- AWS infrastructure availability
- Backup success rate
- Backup restore success rate
- DR drill success rate
- Server patch compliance
- Critical vulnerability remediation within SLA
- VAPT remediation closure rate
- PIM/PAM privileged-account compliance
- Infrastructure incident resolution within SLA
- Change success rate
- Infrastructure capacity utilization
- Audit and compliance findings
- Infrastructure documentation accuracy
If interested, please share your resume at
[email protected]
📌 Job Opening For Sr. Server Administrator - Mumbai
🏢 Team Computers
📍 Mumbai