Job Opportunity Posted 10 days ago

Storage and Resiliency Operations

The Hartford India
Hyderabad

Job Description

Job Title: Storage and Resiliency


Overview

We are seeking a highly skilled and motivated engineer to serve in our Infrastructure as a Service (IaaS) organization supporting Storage and Resiliency. This role is ideal for a seasoned professional with deep expertise in enterprise storage, backup, replication, disaster recovery, cyber recovery, and recoverability operations across a large scale environment. The ideal candidate will guide junior engineers and drive operational excellence across storage, backup, and resiliency services.


Key Responsibilities

Storage & Backup Administration

  • Install, configure, and maintain enterprise storage, backup, replication, and recovery platforms across private and public cloud environments
  • Manage lifecycle activities including provisioning, capacity management, upgrades, technology currency, and decommissioning
  • Monitor storage and backup platform health, performance, availability, recoverability, and capacity through observability capabilities
  • Perform troubleshooting and root cause analysis for storage, backup, replication, and recovery incidents
  • Implement platform security hardening, retention standards, access controls, and compliance requirements
  • Manage storage provisioning, snapshot policies, backup schedules, replication policies, recovery tests, and service reporting

Resiliency & Cyber Recovery

  • Lead disaster recovery, recoverability validation, backup restoration, replication testing, and operational readiness activities
  • Partner with Cybersecurity and application teams to support cyber recovery capabilities and recoverability objectives
  • Drive improvement plans for backup success, restore performance, data protection coverage, and resiliency gaps

24x7 Operations & Incident Management

  • Support 24x7 storage and resiliency operations including capacity, backup operations, restore support, and vulnerability remediation
  • Oversee incident response, root cause analysis, problem management, and service restoration for storage, backup, and recovery services

Mentorship & Collaboration

  • Mentor junior engineers and foster a culture of continuous learning and technical excellence
  • Collaborate with cross-functional teams including compute, network, security, application, disaster recovery, and service management teams

Operational Excellence

  • Ensure high availability, scalability, security, recoverability, and compliance of storage and resiliency environments
  • Develop metrics for capacity, performance, backup success, restore performance, replication health, and recoverability compliance

Qualifications

  • Strong enterprise storage and backup operations experience in large scale environments
  • Experience with NetApp, SAN, NAS, object storage, backup platforms, replication, and disaster recovery technologies
  • Knowledge of backup policies, retention standards, recovery testing, cyber recovery, and recoverability objectives
  • Scripting and automation experience with PowerShell, Python, Ansible, Terraform, or vendor automation tools
  • Storage and backup capacity management, performance troubleshooting, vulnerability remediation, and technology currency experience
  • Networking fundamentals including TCP/IP, DNS, NFS, SMB, Fibre Channel, iSCSI, and firewall concepts
  • Logging and monitoring tools such as Splunk, vendor management tools, Prometheus, or equivalent

Experience with virtualization and/or cloud platforms including VMw

About this job listing
This job opportunity is provided through our external job listing network. MyJobAlerts helps you discover job opportunities and redirects you to the original listing to apply.