Overview
Manage IT infrastructure engineering and operations across on-premises, cloud, identity, endpoint, storage and backup disciplines in support of a 24x7 operating model. Lead shift coverage, operational resilience, compliance, reporting and major incident coordination for the hybrid estate in Navi Mumbai.
What you'll do
- Lead the infrastructure engineering team across Wintel, Messaging, SCCM and Intune, Linux, Cloud Platform, and Storage and Backup roles across shifts.
- Keep each discipline represented on every shift within the 24x7 operating model.
- Set performance expectations, conduct evaluations, build technical depth and cross-skilling, support recruitment and onboarding, and maintain a current skills matrix.
- Run documented handovers at every shift boundary, transferring open incidents, in-flight changes, active priority incidents and pending actions.
- Report availability, incidents, change, patch compliance, capacity, cloud cost and risk to infrastructure leadership.
- Coordinate infrastructure effort during priority and major incidents.
- Oversee Nutanix, hyperconverged infrastructure, VMware, on-premises compute and server hardware operations.
- Oversee Dell Isilon and HPE storage provisioning, capacity, performance, replication and tiering.
- Maintain firmware, hypervisor and storage software currency, plan capacity and end-of-life replacement, manage OEM and partner vendors, and provide input to hardware refresh and support renewal budgets.
- Oversee AWS and Microsoft Azure workloads, including compute, storage, virtual networking, load balancing, monitoring and alerting.
- Maintain AWS IAM, Azure RBAC, federation, privileged access controls, landing zones, subscription and account conventions, tagging, naming standards and guardrails.
- Manage cloud cost and consumption, including rightsizing, commitment planning and orphaned resource cleanup.
- Support migration, modernization and exit activity, hybrid connectivity, DNS and certificate services.
- Manage Microsoft 365, messaging, Active Directory, Entra ID, conditional access, multi-factor authentication and privileged identity management.
- Manage license assignment and entitlement, and coordinate asset and license records with Service Asset and Configuration Management.
- Manage SCCM and Intune endpoint baselines, application packaging and deployment, co-management, endpoint compliance and protection coverage.
- Execute the defined patching cycle across servers, endpoints, hypervisors, storage and cloud workloads, documenting risk acceptance and remediation dates for exceptions.
- Maintain configuration and hardening baselines, remediate drift, and partner with Information Security on vulnerability remediation, certificate lifecycle and privileged access.
- Maintain backup and disaster recovery, including immutable or offsite copies, RTOs, RPOs and periodic restore and DR testing.
- Operate approved change processes with CAB submission, tested rollback and post-implementation review.
- Maintain monitoring coverage, ServiceNow configuration item accuracy, automation, runbooks, SOPs and the infrastructure risk register.
- Support internal and external audits with evidence, control walkthroughs and corrective action closure.
- Apply scripting and declarative tooling to remove repetitive operational work and reduce configuration variance.
- Coordinate with leadership, security, applications and business stakeholders across time zones, including during priority and major incidents.
- Apply root cause analysis and convert recurring incidents into permanent engineering fixes.
What you'll need
- Engineering, Computer Science, or a related field.
- 10+ years of experience in IT infrastructure engineering and operations across virtualisation, storage, identity and cloud.
- 3+ years in a formal supervisory or management role leading infrastructure engineers, including shift-based teams.
- Demonstrated enterprise-scale experience operating Nutanix, hyperconverged infrastructure and VMware virtualization.
- Demonstrated experience operating enterprise storage and backup, including Dell Isilon and HPE storage platforms, with capacity, performance and restore accountability.
- Demonstrated experience administering Windows Server and Linux estates and on-premises compute and server hardware from major OEMs.
- Demonstrated experience operating production workloads in AWS, including compute, storage, virtual networking, IAM and native backup services.
- Demonstrated experience operating production workloads in Microsoft Azure, including IaaS, virtual networking, RBAC, Azure Backup and Azure Site Recovery.
- Experience with hybrid cloud connectivity and identity integration, including VPN or direct interconnect, certificate services and federation with Active Directory and Entra ID.
- Experience managing cloud cost and consumption, including rightsizing, reserved or savings-plan commitments, tagging governance and budget reporting.
- Demonstrated experience with Microsoft 365, Active Directory and Entra ID in a hybrid identity environment, including conditional access and privileged identity management.
- Demonstrated experience with SCCM and Intune endpoint management, including configuration baselines and application deployment.
- Demonstrated experience owning patch management across servers, endpoints and cloud workloads, including exception governance and compliance reporting against defined standards.
- Experience with configuration hardening baselines and vulnerability remediation in partnership with a security function.
- Experience owning backup and disaster recovery, including documented RTO and RPO for business-critical services and periodic recovery testing.
- Experience with infrastructure monitoring and observability tooling across on-premises and cloud estates.
- Working experience with automation and infrastructure as code, such as PowerShell, Bash, Ansible or Terraform.
- Experience operating infrastructure change through a formal process, including Change Advisory Board and post-implementation review.
- Experience operating within a 24x7 or shift-based support model, including structured handover between shifts and regions.
- Experience working with geographically distributed leadership and teams across multiple time zones, including follow-the-sun models.
- Experience supporting internal or external audit, including evidence preparation, control walkthroughs and corrective action closure.
- Working knowledge of enterprise networking fundamentals, including DNS, DHCP, load balancing and firewall change requests.
- Clear, concise written and verbal communication in English across executive, technical and end-user audiences.
Nice to have
- Working knowledge of ServiceNow for incident, change and configuration management.
- Familiarity with IT infrastructure operations delivered from an India-based capability center or offshore delivery organizations.
- ITIL 4 Foundation certification.
- Cloud certification such as AWS Certified Solutions Architect or SysOps Administrator, or Microsoft Azure Administrator (AZ-104) or Azure Solutions Architect (AZ-305).
- Platform certification such as Nutanix NCP, VMware VCP, Microsoft 365 administration, or Red Hat RHCSA.
- Experience in regulated or audit-driven environments, including exposure to SOX ITGC, NIST, ISO/IEC 27001 or equivalent control frameworks.
Details
- Location: Navi Mumbai, India.
- Day shift role within a 24x7 operating model.
- Works effectively within shift rotation, maintenance windows outside standard hours and overlap with US business hours.
- Reports to the Sr. Manager IT Infrastructure.
Read the full description and apply on the company’s own careers page.