Overview
As a Software Engineer – Cloud Infrastructure, you will design, operate, and improve cloud infrastructure and platform engineering systems supporting production applications.
What you'll do
- Design, operate, and continuously improve Kubernetes infrastructure on Amazon EKS, including Karpenter-based autoscaling, workload scheduling, and cluster lifecycle management.
- Administer and tune PostgreSQL databases on RDS that back platform applications such as Airflow, including configuration, parameter tuning, maintenance, and reliability.
- Own Infrastructure as Code using Terraform to provision, modify, and audit AWS resources including S3, EC2, RDS, EKS, and IAM.
- Manage IAM roles, policies, and access patterns across the AWS environment with a security-first mindset.
- Maintain artifact registries including ECR and Artifactory by enforcing retention policies, image versioning standards, and responding to infosec requests around package vulnerabilities and container image remediation.
- Own and evolve the monitoring and observability stack, including Prometheus, Grafana, and OpenSearch, by improving coverage, reducing noise, and evaluating the stack as needs change.
- Support the CI/CD and source control environment across Bitbucket, Jenkins, GitHub, and GitHub Actions during migration.
- Collaborate with data engineers to define infrastructure patterns, containerization standards, and deployment best practices.
- Participate in an on-call rotation for platform-level incidents.
What you'll need
- Bachelor's degree in Computer Science, Engineering, or a related field, or equivalent professional experience.
- 4+ years of hands-on experience in cloud infrastructure, platform engineering, or DevOps in production environments.
- Deep expertise with Kubernetes in production, ideally Amazon EKS, including autoscaling with Karpenter or Cluster Autoscaler, networking, and day-to-day operations.
- Strong proficiency in Terraform for infrastructure as code.
- Solid AWS experience across EKS, EC2, S3, RDS, and IAM.
- Experience administering PostgreSQL or similar relational databases in a cloud environment, including configuration, tuning for specific workloads, and operational maintenance.
- Hands-on experience with Helm, ArgoCD, and GitOps delivery workflows.
- Experience with observability tooling including Prometheus, Grafana, and log aggregation using OpenSearch or similar.
- Familiarity with artifact management, container image lifecycle, and vulnerability remediation workflows.
- Experience with code repository and CI/CD solutions such as Bitbucket, Jenkins, GitHub, and GitHub Actions.
- Scripting and automation proficiency in Python and/or Bash.
- Strong Linux fundamentals and comfort operating cloud-native production systems.
- Commitment to the highest ethical standards.
Details
- Location: Bengaluru, India.
- Participate in an on-call rotation for platform-level incidents.
Read the full description and apply on the company’s own careers page.