Search by job, company or skills

Storage and Resiliency Operations

  • Posted 21 hours ago
  • Be among the first 10 applicants

Job Description

Job Title: Storage and Resiliency

Overview

We are seeking a highly skilled and motivated engineer to serve in our Infrastructure as a Service (IaaS) organization supporting Storage and Resiliency. This role is ideal for a seasoned professional with deep expertise in enterprise storage, backup, replication, disaster recovery, cyber recovery, and recoverability operations across a large scale environment. The ideal candidate will guide junior engineers and drive operational excellence across storage, backup, and resiliency services.

Key Responsibilities

Storage & Backup Administration

  • Install, configure, and maintain enterprise storage, backup, replication, and recovery platforms across private and public cloud environments
  • Manage lifecycle activities including provisioning, capacity management, upgrades, technology currency, and decommissioning
  • Monitor storage and backup platform health, performance, availability, recoverability, and capacity through observability capabilities
  • Perform troubleshooting and root cause analysis for storage, backup, replication, and recovery incidents
  • Implement platform security hardening, retention standards, access controls, and compliance requirements
  • Manage storage provisioning, snapshot policies, backup schedules, replication policies, recovery tests, and service reporting

Resiliency & Cyber Recovery

  • Lead disaster recovery, recoverability validation, backup restoration, replication testing, and operational readiness activities
  • Partner with Cybersecurity and application teams to support cyber recovery capabilities and recoverability objectives
  • Drive improvement plans for backup success, restore performance, data protection coverage, and resiliency gaps

24x7 Operations & Incident Management

  • Support 24x7 storage and resiliency operations including capacity, backup operations, restore support, and vulnerability remediation
  • Oversee incident response, root cause analysis, problem management, and service restoration for storage, backup, and recovery services

Mentorship & Collaboration

  • Mentor junior engineers and foster a culture of continuous learning and technical excellence
  • Collaborate with cross-functional teams including compute, network, security, application, disaster recovery, and service management teams

Operational Excellence

  • Ensure high availability, scalability, security, recoverability, and compliance of storage and resiliency environments
  • Develop metrics for capacity, performance, backup success, restore performance, replication health, and recoverability compliance

Qualifications

  • Strong enterprise storage and backup operations experience in large scale environments
  • Experience with NetApp, SAN, NAS, object storage, backup platforms, replication, and disaster recovery technologies
  • Knowledge of backup policies, retention standards, recovery testing, cyber recovery, and recoverability objectives
  • Scripting and automation experience with PowerShell, Python, Ansible, Terraform, or vendor automation tools
  • Storage and backup capacity management, performance troubleshooting, vulnerability remediation, and technology currency experience
  • Networking fundamentals including TCP/IP, DNS, NFS, SMB, Fibre Channel, iSCSI, and firewall concepts
  • Logging and monitoring tools such as Splunk, vendor management tools, Prometheus, or equivalent

Experience with virtualization and/or cloud platforms including VMw

More Info

Job Type:
Industry:
Employment Type:

About Company

Job ID: 152531381

Beware of Scammers

We don’t charge money for job offers