Overview
The Storage and Backup Engineer protects and recovers data across on-premises and cloud environments, administering data protection platforms, enterprise SAN and NAS storage, supporting compute and virtualization platforms, and disaster recovery capabilities. The role works within a shift-based infrastructure team in Navi Mumbai and provides technical guidance to less experienced engineers.
What you'll do
- Administer Rubrik, Veeam, NetBackup and tape-based data protection, including policies, protection groups, SLA domains, scheduling, agent health, libraries, media pools, offsite rotation and vaulting.
- Monitor backup and replication outcomes, investigate failures to root cause, remediate issues, maintain immutable and air-gapped copies, upgrade platforms and manage vendor cases.
- Ensure protected workloads have appropriate policies and onboard unprotected workloads across on-premises and cloud estates.
- Administer AWS backup and archive using S3 and Glacier storage classes, lifecycle policies, versioning, retention, expiry and tier transitions.
- Configure and verify S3 Object Lock, AWS Backup Vault Lock, WORM retention, IAM, KMS, cross-account separation and least-privilege access.
- Administer AWS Backup, EBS and EC2 snapshots, RDS snapshots, and EFS or FSx backups.
- Manage replication from on-premises Rubrik clusters to Rubrik Cloud, including archival locations, retention, seeding, bandwidth and reconciliation.
- Maintain backup connectivity and throughput, including VPC endpoints, Direct Connect or VPN capacity, and shared-bandwidth impact.
- Execute cloud recovery, including Glacier restores, on-premises workload recovery into AWS and cloud-native workload recovery, while assessing retrieval time and cost.
- Monitor cloud protection costs, including storage, retrieval, egress, orphaned snapshots and unattached volumes, and report budget variance.
- Execute restores and recoveries across file, virtual machine, database and application workloads against agreed targets.
- Provide second- and third-level support for data protection, storage and compute recovery issues across backup agent, array, hypervisor, database and network layers.
- Perform data copy and migration between arrays, sites, tiers and cloud targets, including cutover planning, integrity validation and rollback.
- Support legal hold, eDiscovery and audit retrieval requests and participate in major incidents involving data loss, corruption or platform failure.
- Operate restoration tests across file, virtual machine, database and application workloads; agree schedules, scope and acceptance criteria with infrastructure, service and application owners.
- Document restoration-test outcomes, failures, corrective actions and recovery estimates; re-test after remediation and maintain audit evidence.
- Maintain disaster recovery capability, replication topology, failover design, runbooks and per-workload RTO and RPO positions.
- Participate in and execute disaster recovery drills covering failover, failback and validation, track findings and coordinate remediation to verified closure.
- Follow and enforce change control, rollback planning and coordination of maintenance, drill and test windows.
- Administer SAN and NAS storage, including volume and LUN provisioning, snapshots, replication, zoning support and performance troubleshooting.
- Administer supporting VMware and Nutanix platforms, Windows and Linux workloads, and SQL Server or Oracle backup interfaces.
- Maintain the physical storage, tape and server estate, including racking, cabling, firmware and OEM support status.
- Manage capacity planning, forecasting, procurement requirements, lifecycle, refresh planning, media retirement and secure disposal.
- Optimize retention, deduplication, compression, tiering, archive placement, cloud storage classes, orphaned snapshots and unattached volumes.
- Produce recurring backup, restore, restoration-test, capacity and compliance reports for IT leadership, information security, audit and service owners.
- Develop PowerShell, Python or Bash scripting and platform API automation for administration and protection reporting.
- Maintain runbooks, procedures, configuration and retention documentation and provide technical guidance and mentoring.
- Work the assigned shift, complete and verify documented handovers, and coordinate with other shifts and infrastructure teams.
What you'll need
- Bachelor’s degree in computer science, Information Technology, Engineering or a closely related field, or equivalent professional experience.
- 7+ years of experience in systems engineering, storage administration or data protection, including substantial hands-on responsibility for a production backup estate.
- Hands-on administration of at least two enterprise data protection platforms from Rubrik, Veeam, NetBackup or comparable products, including policy design and failure remediation.
- Hands-on administration of enterprise SAN and NAS storage, including volume and LUN provisioning, snapshots, replication and capacity management.
- Hands-on experience with compute and virtualization platforms related to protection and recovery, including VMware or Nutanix, Windows and Linux workloads, and SQL Server or Oracle backup interfaces.
- Experience with tape infrastructure, including libraries, media management and offsite rotation.
- Experience executing restores and recoveries across files, virtual machines, databases and applications.
- Experience operating a scheduled restoration testing program, including test design, execution with service owners, evidence capture and remediation.
- Experience maintaining per-workload RTO and RPO objectives against a defined organizational standard, including measurement and escalation of shortfalls.
- Active participation in disaster recovery drills as an executing engineer, including failover, failback and validation.
- Experience producing recurring backup, recovery, capacity and compliance reporting for IT leadership, security or audit.
- Hands-on AWS backup and archive experience, including S3 and Glacier storage classes, bucket and lifecycle policy, versioning, retention and cross-region or cross-account copies.
- Hands-on cloud immutability experience, including S3 Object Lock or AWS Backup Vault Lock, and IAM roles, policies and KMS encryption for backup access and recovery.
- Hands-on Rubrik replication and archival to Rubrik Cloud, or equivalent vendor cloud replication and archive from an on-premises backup platform.
- Experience protecting AWS-native workloads, including AWS Backup, EBS and EC2 snapshots, and RDS, EFS or FSx backups.
- Experience executing cloud recovery, including archive-tier restores with retrieval time and cost understood in advance.
- Experience managing cloud protection cost, including storage class placement, retrieval and egress charges, and removal of orphaned snapshots.
- Experience working within a shift-based infrastructure team, including structured handover of running jobs and open recovery work.
- Scripting proficiency in PowerShell, Python or Bash, with practical use of platform APIs for reporting and bulk administration.
- Professional proficiency in spoken and written English.
Nice to have
- Certification in a relevant platform such as Rubrik, Veeam Certified Engineer (VMCE), NetBackup, or a storage vendor credential from NetApp, Dell, Pure Storage or equivalent.
- AWS certification such as AWS Certified Solutions Architect – Associate or AWS Certified SysOps Administrator.
- ITIL 4 Foundation, or equivalent demonstrated knowledge of incident, change and problem practice.
Details
- Location: Navi Mumbai.
- The role works within a shift-based infrastructure team and includes assigned-shift reporting to the Manager IT Infrastructure.
- Shift handover must cover running jobs, failed backups, open restores, in-flight migrations, active incidents and pending actions.
- The role requires coordination across shifts and regions.
Read the full description and apply on the company’s own careers page.