01Key Responsibilities
Incident Remediation & Operational Support
Support remediation of storage and backup incidents including outages, degraded performance, failed backups, replication issues, and data access issues
Participate in incident bridges and assist with troubleshooting across storage, backup, compute, virtualization, network, and application teams
Perform initial triage, evidence gathering, log review, and escalation for production-impacting issues
Assist with root cause analysis and document corrective actions to reduce repeat incidents
Monitor platform health and respond to alerts, failed jobs, capacity thresholds, and service-impacting conditions
Maintain and follow operational runbooks, escalation paths, and recovery procedures
Storage Platform Operations
Administer and support storage platforms including NetApp ONTAP, Pure Storage FlashArray, Dell PowerStore, HPE Allatra, and IBM Storwize / FlashSystem
Perform routine storage provisioning, volume and LUN management, capacity expansion, mapping, masking, and decommissioning activities
Support NAS services including CIFS/SMB shares, NFS exports, permissions coordination, and access troubleshooting
Assist with replication, availability, and resiliency activities across supported storage platforms
Review capacity, performance, and health metrics to identify risks and support proactive remediation
SAN/NAS Operations
Support SAN and NAS connectivity across enterprise infrastructure environments
Assist with Fibre Channel zoning, host connectivity, multipathing validation, and storage presentation troubleshooting
Work with compute, virtualization, and network teams to resolve connectivity, latency, pathing, and access issues
Maintain operational standards for storage provisioning, naming, documentation, and configuration hygiene
Backup & Recovery Operations
Administer and support enterprise backup and recovery operations with a focus on Rubrik
Investigate and remediate failed backups, missed SLAs, replication issues, policy gaps, and restore failures
Support backup policy configuration, retention management, workload onboarding, and operational reporting
Perform restore testing and assist with validation of recovery procedures
Contribute to backup reliability, cyber recovery readiness, and disaster recovery preparedness
Upgrade & Lifecycle Execution
Support lifecycle management for storage and backup platforms including software upgrades, firmware updates, hardware refreshes, and platform maintenance
Assist with planning, testing, scheduling, execution, and validation of storage and backup upgrade activities
Follow documented implementation plans, pre-checks, post-checks, and rollback procedures
Track platform versions, support status, known issues, and remediation needs across supported technologies
Coordinate with vendors and senior engineers during maintenance windows, escalations, and lifecycle events
Documentation, Monitoring & Operational Hygiene
Maintain accurate documentation for storage allocations, backup policies, platform configurations, operational procedures, and support contacts
Support CMDB accuracy and relationship mapping for storage, backup, host, and application .