Overview
As a Platform Engineer, you will modernize and automate virtual machine and core infrastructure services supporting data center and hybrid-cloud environments. The role focuses on reliability, scalability, security, Infrastructure as Code, GitOps, VM and hypervisor management, and automated lifecycle management of COTS management tools.
What you'll do
- Support hypervisor operations across a global datacenter footprint.
- Provide VM management services, such as patching, backups, and monitoring, through automation and at scale.
- Create and manage reusable Terraform modules and Ansible playbooks for consistent, repeatable, and version-controlled deployments of core services.
- Apply GitOps principles to infrastructure and service configurations, using pull requests for changes and enabling automated reconciliation.
- Lead the architecture, automation, and operational management of enterprise DNS, Public Key Infrastructure (PKI), and NTP.
- Design, build, and maintain centralized log aggregation for the observability platform covering infrastructure and applications.
- Build and maintain CI/CD pipelines to automate testing, validation, and promotion of changes to core service configurations and infrastructure.
- Embed security and compliance into the service lifecycle, including secret management, certificate lifecycle automation, and vulnerability scanning.
- Implement monitoring, alerting, and dashboards for managed services.
- Apply SRE principles to improve service reliability and performance.
- Collaborate with network, security, and application teams to ensure seamless integration and consumption of core services.
What you'll need
- +5 year experience working as a systems engineer or site reliability engineer supporting on-premise datacenter infrastructure and core infrastructure services.
- Proficiency with hypervisor technologies that support operating systems and containers, such as VMware, Nutanix, or Openshift.
- Proficiency with automation and management of Windows Server and Linux VMs, such as Ubuntu, Redhat, or Debian.
- Deep hands-on experience engineering and automating core infrastructure services within a large-scale data center environment, such as DNS/BIND or Active Directory Certificate Services/PKI.
- Experience building container platform capabilities in an on-premise ecosystem.
- Strong proficiency with Infrastructure as Code (IaC) and configuration management tools, particularly Terraform and Ansible.
- Proven experience building CI/CD pipelines for infrastructure automation, such as using GitLab CI, Jenkins, or similar tools.
- Comfort with Git-based workflows, including PR reviews, branching strategies, and versioning, and GitOps principles.
- Strong scripting and programming ability, such as Python, Bash, or PowerShell, to build automation, integrations, and tooling.
- Solid understanding of data center networking concepts, including TCP/IP, routing, and firewalls.
- Solid understanding of security principles, including least privilege, encryption, and network security zones.
- Experience with centralized logging and observability platforms, such as Splunk, Elastic Stack, or LogicMonitor.
Nice to have
- Familiarity with security and control frameworks such as NIST CSF, SOX, and ISO 27002.
- Experience working in hybrid-cloud environments and connecting on-premises services to public cloud providers such as AWS or Azure.
- Knowledge of containerization and Kubernetes.
- Experience with platform engineering concepts such as building internal tools and APIs to improve the developer/operator experience.
Details
- Location: Hyderabad, India.
Read the full description and apply on the company’s own careers page.