Overview
The Storage and Backup Engineer protects and recovers data across on-premises and cloud environments, administering Rubrik, Veeam, NetBackup, tape, AWS, Rubrik Cloud, enterprise SAN and NAS storage, and supporting compute and virtualization platforms. The role operates on an assigned shift in Navi Mumbai and is accountable for backup success, restore performance, disaster recovery capability, RTO/RPO adherence and recoverability testing.
What you'll do
- Administer Rubrik, Veeam and NetBackup policy configuration, protection groups, SLA domains, scheduling, and agent or connector health.
- Administer tape libraries, drives, media pools, barcodes, slots, and offsite media rotation and vaulting.
- Monitor backup and replication outcomes, investigate failures to root cause, and remediate them.
- Maintain immutable and air-gapped protection copies and verify that immutability is enforced.
- Maintain platform currency through software and firmware upgrades and manage vendor cases to closure.
- Ensure protected workloads have appropriate policies and onboard unprotected workloads across on-premises and cloud estates.
- Administer AWS backup and archive using S3 and Glacier storage classes, bucket and lifecycle policies, versioning, retention, expiry and tier transitions.
- Configure and maintain S3 Object Lock and AWS Backup Vault Lock immutability and verify WORM retention.
- Administer AWS Backup, EBS and EC2 snapshots, RDS snapshots, and EFS or FSx backups where used.
- Manage replication from on-premises Rubrik clusters to Rubrik Cloud, including archival locations, replication and retention policies, seeding, bandwidth and reconciliation.
- Maintain IAM roles, policies and KMS keys, applying least privilege and cross-account separation.
- Maintain backup connectivity and throughput, including VPC endpoints, Direct Connect or VPN capacity, and shared-bandwidth impact.
- Execute and test cloud recovery, including Glacier restores, on-premises workload recovery into AWS, and cloud-native workload recovery.
- Monitor cloud protection costs, including storage, retrieval, egress, orphaned snapshots and unattached volumes, and report budget variance.
- Execute restores and recoveries across file, virtual machine, database and application workloads.
- Provide second- and third-level support for complex protection, storage and compute recovery issues across backup agent, array, hypervisor, database and network layers.
- Perform data copy and migration between arrays, sites, tiers and cloud targets, including cutover planning, integrity validation and rollback.
- Support legal hold, eDiscovery and audit retrieval requests.
- Participate in priority and major incidents involving data loss, corruption or platform failure and provide recovery estimates.
- Operate a defined restoration-testing cadence across file, virtual machine, database and application recovery.
- Agree test schedules, scope and acceptance criteria with infrastructure teams and service and application owners.
- Document test outcomes, failures, corrective actions and revised recovery estimates, and retest after remediation.
- Publish restoration-test results and maintain audit evidence until findings are closed.
- Maintain disaster recovery capability, replication topology, failover design and runbooks.
- Maintain per-workload RTO and RPO positions against global standards and escalate shortfalls.
- Execute disaster recovery drills covering failover, failback and validation.
- Plan and execute DR tests, report remediation actions and track findings to verified closure.
- Follow and enforce change control, including rollback planning and coordination of maintenance, drill and test windows.
- Administer enterprise SAN and NAS storage, including volumes, LUNs, snapshots, replication, zoning support and performance troubleshooting.
- Administer VMware and Nutanix hosts and clusters, Windows and Linux workloads, and SQL Server or Oracle backup interfaces as they relate to protection and recovery.
- Maintain the physical estate, including racking, cabling, firmware and OEM support status for storage, tape and server hardware.
- Manage capacity planning and forecasting across storage, compute and protection estates.
- Optimize retention, deduplication, compression, tiering, archive placement, cloud storage classes, orphaned snapshots and unattached volumes.
- Manage storage, compute and backup hardware, media and licensing lifecycle, including refresh planning, media retirement and secure disposal.
- Produce recurring backup, restore, restoration-test, capacity and compliance reports.
- Develop PowerShell, Python or Bash scripting and platform API automation.
- Maintain runbooks, procedures, configuration and retention documentation and mentor less experienced engineers.
- Complete documented shift handovers covering running jobs, failed backups, open restores, migrations, incidents and pending actions, and verify received handovers.
- Coordinate across shifts and infrastructure teams so protection and recovery activity continues without loss of context.
What you'll need
- Bachelor’s degree in computer science, Information Technology, Engineering or a closely related field, or equivalent professional experience.
- 7+ years of experience in systems engineering, storage administration or data protection, including substantial hands-on responsibility for a production backup estate.
- Hands-on administration of at least two enterprise data protection platforms from Rubrik, Veeam, NetBackup or comparable products, including policy design and failure remediation.
- Hands-on administration of enterprise SAN and NAS storage, including volume and LUN provisioning, snapshots, replication and capacity management.
- Hands-on experience with compute and virtualization platforms related to protection and recovery, including VMware or Nutanix, Windows and Linux workloads, and database backup interfaces such as SQL Server or Oracle.
- Experience with tape infrastructure, including libraries, media management and offsite rotation.
- Experience executing restores and recoveries across files, virtual machine, database and application workloads.
- Experience operating a scheduled restoration-testing program, including test design, execution with service owners, evidence capture and remediation of findings.
- Experience maintaining per-workload recovery time and recovery point objectives against a defined organizational standard, including measurement and escalation of shortfalls.
- Active participation in disaster recovery drills as an executing engineer, including failover, failback and validation.
- Experience producing recurring backup, recovery, capacity and compliance reporting for IT leadership, security or audit.
- Advanced, practical administration of enterprise data protection platforms, tape infrastructure and media management.
- Hands-on command of enterprise SAN and NAS storage and supporting compute and virtualization platforms.
- Practical command of AWS backup and archive, including S3 and Glacier, lifecycle and retention policy, Object Lock and Vault Lock, AWS Backup, native snapshots and Rubrik Cloud replication.
- Ability to apply least privilege through IAM roles and policies, manage KMS keys and cross-account separation, and protect backup copies.
- Strong analytical skills with proven ability to identify root cause across backup agent, array, hypervisor, database and network layers.
- Excellent written and verbal communication in English.
- Ability to work independently, make technical decisions with limited information, and provide technical guidance to less experienced engineers.
Details
- Location: Navi Mumbai.
- Work within a shift-based infrastructure team on the assigned shift.
- Report to the Manager IT Infrastructure for the shift worked.
- Coordinate with other shifts and infrastructure teams across regions.
Read the full description and apply on the company’s own careers page.