Overview
Join NCR VOYIX as a Site Reliability Engineer II (GCP & Azure), supporting reliable software systems through CI/CD, infrastructure automation, monitoring, and collaboration between development and operations teams.
What you'll do
- Design, develop, and maintain scalable and reliable CI/CD pipelines using tools like Jenkins, GitLab CI/CD, or similar.
- Implement and manage infrastructure as code using tools such as Terraform, CloudFormation, or Ansible to provision and configure Azure and GCP cloud resources.
- Monitor system performance, troubleshoot issues, and implement solutions to optimize reliability, availability, and efficiency.
- Collaborate with development teams to integrate automated testing, security scanning, and deployment processes into the CI/CD workflow.
- Manage and administer containerization technologies such as Docker and Kubernetes and container orchestration platforms.
- Develop and maintain comprehensive documentation for infrastructure, processes, and tools.
- Participate in on-call rotations to support critical production systems.
- Automate routine tasks and workflows to improve operational efficiency.
- Contribute to the selection and evaluation of new technologies and tools to enhance DevOps capabilities.
- Ensure compliance with security best practices and company policies.
- Undertake additional duties and projects consistent with your skills, capabilities, and the overall purpose of the position.
What you'll need
- Bachelor's degree in Computer Science, Information Technology, or a related field, or equivalent practical experience.
- 3 to 6 years of experience in a DevOps, SRE, or similar role.
- Proficiency in scripting languages such as Python, Bash, or PowerShell.
- Strong experience with at least one major cloud provider (Azure & GCP).
- Solid understanding of CI/CD principles and hands-on experience with relevant tools such as Jenkins, GitLab CI/CD, or Azure DevOps.
- Experience with infrastructure as code tools such as Terraform, Ansible, or CloudFormation.
- Familiarity with containerization technologies such as Docker and Kubernetes.
- Experience with monitoring and logging tools such as Prometheus, Grafana, ELK stack, or Splunk.
- Strong understanding of networking concepts, operating systems (Linux/Windows), and database systems.
- Excellent problem-solving skills and the ability to troubleshoot complex technical issues.
- Strong communication and collaboration skills, with the ability to work effectively in a team environment.
- Experience with version control systems, particularly Git.
Details
- Location: Chennai, India.
- Participate in on-call rotations to support critical production systems.
Read the full description and apply on the company’s own careers page.