The Server & Storage Expert will be responsible for the design, implementation, administration, monitoring, troubleshooting, backup and optimization of mission-critical server, storage and virtualization infrastructure. The role will provide L3/L5 technical support for complex infrastructure incidents, performance issues, capacity planning, high availability, disaster recovery and data center operations.
YOU MUST HAVE
- Minimum 2 years of experience in advanced technical support or related roles
- Strong leadership and team management skills
- Excellent problem-solving and decision-making abilities
WE VALUE
- Bachelor's degree in Engineering or related field
- Experience with quality management systems and processes
- Passion for innovation and continuous learning
Key Responsibilities
- Manage enterprise physical and virtual server infrastructure.
- Administration of Windows Server 2019/2022/2025 and Linux environments.
- Manage Dell/HPE/Lenovo or equivalent enterprise servers.
- Perform server provisioning, configuration, hardening and decommissioning.
- Troubleshoot CPU, memory, disk, RAID, firmware and hardware issues.
- Manage server BIOS, firmware, drivers and hardware lifecycle.
- Perform OS patching and vulnerability remediation.
- Analyze Windows/Linux system logs and performance counters.
- Handle critical production incidents and provide RCA.
- Administration of Microsoft Hyper-V / VMware vSphere environments.
- Manage virtualization clusters, hosts, VMs, templates and virtual switches.
- Configure and troubleshoot: HA/Failover Clustering, Live Migration, Cluster Shared Volumes, Virtual networking, VM performance, VM failover and recovery, Troubleshoot VM freeze, pause, crash and unexpected shutdown issues. Perform VM migration and workload balancing. Capacity planning for CPU, memory and storage.
- Administration of enterprise SAN/NAS storage.
- Hands-on experience with NetApp ONTAP, Dell EMC, HPE, IBM or equivalent storage platforms.
- Configure and manage : LUNs, Volumes, Aggregates, Storage pools, Shares, Snapshots, Storage replication, Manage FC/iSCSI connectivity. Configure and troubleshoot SAN zoning and multipathing. Monitor storage capacity, latency, IOPS and throughput. Troubleshoot storage path failures and degraded storage performance. Perform storage expansion, migration and optimization.
- Windows Failover Clustering
- Administration and troubleshooting of Windows Server Failover Clustering.
- Manage cluster nodes, roles and resources.
- Troubleshoot: Cluster node failures, Quorum issues, CSV problems, Cluster communication failures, Storage connectivity, Network heartbeat failures. Perform controlled node maintenance using Drain/Resume operations. Manage SQL/Hyper-V clustered workloads.
- Backup & Disaster Recovery
- Administration of enterprise backup solutions such as Veeam Backup & Replication.
- Configure and monitor: VM backups, Incremental/full backups, Backup repositories, Backup copy jobs, Replication, Retention policies, Troubleshoot failed backup and replication jobs. Perform VM/file/database recovery. Participate in DR drills and BCP activities. Validate RPO/RTO requirements.
- Monitoring & Capacity Management
- Monitor server, storage and virtualization infrastructure.
- Analyze: CPU utilization, Memory utilization, Disk latency, IOPS, Network throughput, Storage capacity, Identify capacity and performance bottlenecks. Prepare infrastructure health and capacity reports. Maintain infrastructure dashboards and operational KPIs.
- Implement server hardening according to security standards.
- Remediate VAPT vulnerabilities.
- Manage OS security patches and firmware updates.
- Support vulnerability assessment and penetration testing activities.
- Incident & Problem Management – L3/L5
- Act as the final technical escalation point for complex server/storage incidents.
- Analyze major incidents and identify root causes.
- Coordinate with OEMs, vendors and technology partners.
- Provide RCA, corrective action and preventive action (CAPA).
- Participate in 24×7 critical incident support when required.
- Lead technical bridge calls during P1/P2 incidents.
- Review recurring incidents and implement permanent fixes.
- OEM & Vendor Coordination
- Coordinate with Dell, HPE, NetApp, Microsoft, VMware/Broadcom, Veeam and other OEMs.
- Open and manage critical support cases.
- Provide logs and diagnostic information to OEM support.
- Validate OEM recommendations before production implementation.
- Track hardware replacement, firmware updates and warranty activities.
- Coordinate with application vendor.
Honeywell Technologies is a global, pure-play automation company with a legacy of innovating to help solve the world’s most mission-critical challenges, enhancing the quality of life for people and communities around the world. We serve the building, industrial and process sectors with a broad portfolio of services, solutions and products, underpinned by our Honeywell Technologies Accelerator operating system and Honeywell Technologies Forge intelligence layer. By combining the deep domain expertise of our more than 50,000 employees with decades of data from our global installed base, we are uniquely positioned to lead the industrial sector’s transition from automation to autonomy.