Overview
The Storage and Backup Engineer protects and recovers data across on-premises and cloud environments, administering enterprise data protection, SAN/NAS storage, compute and virtualization platforms. The role works in a shift-based infrastructure team in Navi Mumbai and is accountable for backup success, restore performance, disaster recovery capability, RTO/RPO adherence and recoverability testing.
What you'll do
- Administer Rubrik, Veeam, NetBackup, tape infrastructure and associated policies, protection groups, SLA domains, schedules, media and connectors.
- Monitor backup and replication outcomes, investigate failures to root cause, remediate issues and maintain platform currency.
- Maintain immutable and air-gapped protection copies and verify enforced immutability.
- Ensure protected workloads are covered by appropriate policies and onboard unprotected workloads.
- Administer AWS backup and archive, including S3, Glacier, AWS Backup, EBS, EC2, RDS, EFS and FSx protection.
- Manage S3 lifecycle, versioning, retention, expiry, tier transitions, Object Lock, AWS Backup Vault Lock, IAM, KMS, cross-region and cross-account copies.
- Manage replication from on-premises Rubrik clusters to Rubrik Cloud, including archival configuration, retention, seeding, bandwidth and reconciliation.
- Maintain backup connectivity and throughput and manage shared-bandwidth impact.
- Execute and test cloud recovery, including Glacier retrievals, on-premises recovery into AWS and cloud-native workload recovery.
- Monitor cloud protection costs, retrieval and egress charges, orphaned snapshots and unattached volumes.
- Execute restores and recoveries across file, virtual machine, database and application workloads.
- Provide second- and third-level support for complex recovery issues across backup agent, array, hypervisor, database and network layers.
- Perform data copy and migration between arrays, sites, tiers and cloud targets, including cutover planning, integrity validation and rollback.
- Support legal hold, eDiscovery and audit retrieval requests.
- Participate in priority and major incidents involving data loss, corruption or platform failure and provide recovery estimates.
- Operate scheduled restoration tests across file, virtual machine, database and application workloads.
- Agree test schedules, scope and acceptance criteria with infrastructure teams and service and application owners.
- Document test outcomes, failures, corrective actions and revised recovery estimates; re-test after remediation.
- Publish restoration results and maintain audit evidence until findings are closed.
- Maintain disaster recovery capability, replication topology, failover design and runbooks.
- Maintain per-workload RTO and RPO positions against the Institute's global standards and escalate shortfalls.
- Plan and execute disaster recovery drills covering failover, failback and validation.
- Follow change control, rollback planning and maintenance, drill and test-window coordination.
- Administer enterprise SAN and NAS storage, including volume and LUN provisioning, snapshots, replication, zoning support and performance troubleshooting.
- Administer VMware and Nutanix platforms, Windows and Linux workloads, and SQL Server and Oracle backup interfaces.
- Manage physical storage, tape and server hardware, including racking, cabling, firmware and OEM support.
- Forecast capacity, raise procurement requirements, optimize retention, deduplication, compression, tiering and archive placement, and manage hardware, media and licensing lifecycle.
- Produce backup, restore, restoration-test, capacity and compliance reports.
- Develop PowerShell, Python or Bash scripting and platform API automation for administration and reporting.
- Maintain runbooks, procedures, configuration and retention documentation and provide technical guidance to less experienced engineers.
- Complete documented shift handovers covering running jobs, failed backups, open restores, migrations, incidents and pending actions.
- Coordinate with other shifts and infrastructure teams to preserve operational context across regions.
What you'll need
- Bachelor’s degree in computer science, Information Technology, Engineering or a closely related field, or equivalent professional experience.
- 7+ years of experience in systems engineering, storage administration or data protection, including substantial hands-on responsibility for a production backup estate.
- Hands-on administration of at least two enterprise data protection platforms from Rubrik, Veeam, NetBackup or comparable products, including policy design and failure remediation.
- Hands-on administration of enterprise SAN and NAS storage, including volume and LUN provisioning, snapshots, replication and capacity management.
- Hands-on experience with compute and virtualization platforms related to protection and recovery, including VMware or Nutanix, Windows and Linux workloads, and database backup interfaces such as SQL Server or Oracle.
- Experience with tape infrastructure, including libraries, media management and offsite rotation.
- Experience executing restores and recoveries across files, virtual machine, database and application workloads.
- Experience operating a scheduled restoration testing program, including test design, execution with service owners, evidence capture and remediation of findings.
- Experience maintaining per-workload recovery time and recovery point objectives against a defined organizational standard, including measurement and escalation of shortfalls.
- Active participation in disaster recovery drills as an executing engineer, including failover, failback and validation.
- Experience producing recurring backup, recovery, capacity and compliance reporting for IT leadership, security or audit.
- Hands-on experience using AWS as a backup and archive target, including S3 and Glacier storage classes, bucket and lifecycle policy, versioning, retention and cross-region or cross-account copies.
- Hands-on experience with cloud immutability, including S3 Object Lock or AWS Backup Vault Lock, and IAM roles, policies and KMS encryption for backup access and recovery.
- Hands-on experience with Rubrik replication and archival to Rubrik Cloud, or equivalent vendor cloud replication and archive from an on-premises backup platform.
- Experience protecting AWS-native workloads, including AWS Backup, EBS and EC2 snapshots, and RDS, EFS or FSx backups.
- Experience executing cloud recovery, including archive-tier restore with retrieval time and cost understood in advance.
- Experience managing cloud protection cost, including storage class placement, retrieval and egress charges, and removal of orphaned snapshots.
- Experience working within a shift-based infrastructure team, including structured handover of running jobs and open recovery work.
- Scripting proficiency in PowerShell, Python or Bash, with practical use of platform APIs for reporting and bulk administration.
- Professional proficiency in spoken and written English sufficient to coordinate with global infrastructure and service teams and communicate with senior stakeholders.
Nice to have
- Certification in a relevant platform such as Rubrik, Veeam Certified Engineer (VMCE), NetBackup, or a storage vendor credential from NetApp, Dell, Pure Storage or equivalent.
- AWS certification such as AWS Certified Solutions Architect – Associate or AWS Certified SysOps Administrator.
- ITIL 4 Foundation, or equivalent demonstrated knowledge of incident, change and problem practice.
Details
- Location: Navi Mumbai.
- Work in a shift-based infrastructure team and work the assigned shift.
- Report to the Manager IT Infrastructure for the shift worked.
- Complete documented handovers at shift boundaries and coordinate with other shifts.
Read the full description and apply on the company’s own careers page.