Overview
Manage IT infrastructure engineering and operations across a 24x7 operating model on the swing shift, supporting Wintel, Messaging, SCCM and Intune, Linux, Cloud Platform, and Storage and Backup disciplines. Own day-to-day operation of the hybrid estate across on-premises infrastructure, AWS, Microsoft Azure, Microsoft 365, identity, endpoint management, patching, resilience, security compliance, audit readiness, and shift handover.
What you'll do
- Lead the infrastructure engineering team across shifts, covering Wintel, Messaging, SCCM and Intune, Linux, Cloud Platform, and Storage and Backup roles while keeping each discipline represented on every shift within the 24x7 operating model.
- Set performance expectations, conduct evaluations, build technical depth and cross-skilling, support recruitment and onboarding, and maintain a current skills matrix.
- Run documented handovers at every shift boundary, transferring open incidents, in-flight changes, active priority incidents, and pending actions.
- Report availability, incidents, change, patch compliance, capacity, cloud cost, and risk to infrastructure leadership, and coordinate infrastructure effort during priority and major incidents.
- Oversee Nutanix, hyperconverged infrastructure, VMware, on-premises compute, server hardware, Dell Isilon, and HPE storage operations.
- Maintain firmware, hypervisor, and storage software currency; plan capacity and end-of-life replacement; manage OEM and partner vendors; and provide input to hardware refresh, support renewal, capital, and operating budget planning.
- Oversee AWS and Microsoft Azure workloads, including compute, storage, virtual networking, load balancing, monitoring, alerting, identity integration, backup, resilience, and cost management.
- Maintain AWS IAM, Azure RBAC, federation, privileged access controls, landing zone, subscription and account conventions, resource tagging, naming standards, and guardrails.
- Manage cloud rightsizing, commitment planning, orphaned resource cleanup, migration, modernization, exit activity, hybrid connectivity, DNS, and certificate services.
- Manage Microsoft 365, messaging, Active Directory, Entra ID, conditional access, multi-factor authentication, privileged identity management, licensing, and entitlement records.
- Manage SCCM and Intune, including image and configuration baselines, application packaging and deployment, co-management, endpoint compliance, and protection coverage.
- Execute patching across servers, endpoints, hypervisors, storage, and cloud workloads, documenting risk acceptance and remediation dates for exceptions.
- Maintain configuration and hardening baselines, remediate drift, and partner with Information Security on vulnerability remediation, certificate lifecycle, and privileged access.
- Maintain backup and disaster recovery, including immutable or offsite copies, recovery time and recovery point objectives, restore testing, and disaster recovery testing.
- Operate approved change processes with CAB submission, tested rollback, and post-implementation review.
- Maintain monitoring coverage, ServiceNow configuration item accuracy, automation of routine work, runbooks, SOPs, and the infrastructure risk register.
- Support internal and external audits through evidence preparation, control walkthroughs, and corrective action closure.
- Work with leadership, security, applications, and business stakeholders across time zones; perform root cause analysis; coordinate major incidents; and provide operational reporting in English.
- Work within a 24x7 operating model, including shift rotation, maintenance windows outside standard hours, and overlap with US business hours.
What you'll need
- Engineering, Computer Science, or a related field.
- 10+ years of experience in IT infrastructure engineering and operations across virtualisation, storage, identity, and cloud.
- 3+ years in a formal supervisory or management role leading infrastructure engineers, including shift-based teams.
- Demonstrated enterprise-scale experience operating Nutanix and hyperconverged infrastructure and VMware virtualization.
- Demonstrated experience operating enterprise storage and backup, including Dell Isilon and HPE storage platforms, with capacity, performance, and restore accountability.
- Demonstrated experience administering Windows Server and Linux estates and on-premises compute and server hardware from major OEMs.
- Demonstrated experience operating production workloads in AWS, including compute, storage, virtual networking, IAM, and native backup services.
- Demonstrated experience operating production workloads in Microsoft Azure, including IaaS, virtual networking, RBAC, Azure Backup, and Azure Site Recovery.
- Experience with hybrid cloud connectivity and identity integration, including VPN or direct interconnect, certificate services, and federation with Active Directory and Entra ID.
- Experience managing cloud cost and consumption, including rightsizing, reserved or savings-plan commitments, tagging governance, and budget reporting.
- Demonstrated experience with Microsoft 365, Active Directory, and Entra ID in a hybrid identity environment, including conditional access and privileged identity management.
- Demonstrated experience with endpoint management through SCCM and Intune, including configuration baselines and application deployment.
- Demonstrated experience owning a patch management cycle across servers, endpoints, and cloud workloads, including exception governance and compliance reporting against defined standards.
- Experience with configuration hardening baselines and vulnerability remediation in partnership with a security function.
- Experience owning backup and disaster recovery, including documented RTO and RPO for business-critical services and periodic recovery testing.
- Experience with infrastructure monitoring and observability tooling across on-premises and cloud estates.
- Working experience with automation and infrastructure as code, such as PowerShell, Bash, Ansible, or Terraform.
- Experience operating infrastructure change through a formal change management process, including Change Advisory Board and post-implementation review.
- Experience operating within a 24x7 or shift-based support model, including structured handover between shifts and regions.
- Experience working with geographically distributed leadership and teams across multiple time zones, including follow-the-sun models.
- Experience supporting internal or external audit, including evidence preparation, control walkthroughs, and corrective action closure.
- Working knowledge of enterprise networking fundamentals, including DNS, DHCP, load balancing, and firewall change requests.
- Clear, concise written and verbal communication in English across executive, technical, and end-user audiences.
Nice to have
- Working knowledge of ServiceNow for incident, change, and configuration management.
- Familiarity with IT infrastructure operations delivered from an India-based capability center or offshore delivery organizations.
- ITIL 4 Foundation certification.
- Cloud certification such as AWS Certified Solutions Architect or SysOps Administrator, or Microsoft Azure Administrator (AZ-104) or Azure Solutions Architect (AZ-305).
- Platform certification such as Nutanix NCP, VMware VCP, Microsoft 365 administration, or Red Hat RHCSA.
- Experience in regulated or audit-driven environments, including exposure to SOX ITGC, NIST, ISO/IEC 27001, or equivalent control frameworks.
Details
- Location: Navi Mumbai, India.
- The role supports engineering and operations in Navi Mumbai.
- Shift: Swing shift within a 24x7 operating model.
- Working conditions include shift rotation, maintenance windows outside standard hours, and overlap with US business hours.
- The role reports to the Sr. Manager IT Infrastructure.
Read the full description and apply on the company’s own careers page.