Overview
Senior DevOps Engineer responsible for building and operating Expedia Group’s core CI infrastructure and enabling fast, reliable, and scalable software delivery across multiple engineering domains. This high-performing individual contributor leads complex, vaguely defined projects, mentors junior engineers, interfaces with local technology leadership, and develops team leadership skills through current projects.
What you'll do
- Design, implement, and operate native cloud infrastructure, CI/CD pipelines, and platform services that enable rapid, reliable application delivery across multiple domains.
- Build, configure, and maintain automated deployment, monitoring, and observability solutions to ensure high availability, scalability, and performance of critical services.
- Drive improvements in infrastructure-as-code, configuration management, and environment standardization to increase reliability, reduce manual work, and improve developer productivity.
- Partner with engineering and architecture teams to define system design, including low-level design, API integration patterns, and data flows aligned with operational and security best practices.
- Safely integrate and operate AI/ML-enabled solutions that improve outcomes.
- Lead incident response, problem management, and post-incident reviews, using data and metrics to drive continuous operational excellence across services and platforms.
- Lead complex and vaguely-defined projects.
- Mentor junior engineers.
- Interface consistently with technology leadership in the local organization.
What you'll need
- Relevant technical degree or equivalent practical experience in software engineering, computer science, or a related field with a focus on infrastructure, systems, or DevOps.
- 9+ years of extensive hands-on experience in native cloud environments operating production systems, including ownership of one or more services or platforms end to end.
- Demonstrated expertise in CI/CD tooling, infrastructure-as-code, and automation, including build, test, deployment, and environment provisioning workflows.
- Strong proficiency in system design, including low-level design of infrastructure components, API integration, networking, security controls, and data modeling for operational systems.
- Proven experience monitoring, scaling, and troubleshooting distributed systems in production using logs, metrics, and alerts to maintain reliability and performance.
Nice to have
- Experience implementing and tuning remote build caching, including Gradle Enterprise and Bazel Remote Cache.
- Experience tuning macOS build performance, concurrency, resource allocation, filesystem tuning, or virtualization.
- Strong AWS experience across compute, storage, networking, and autoscaling.
- Experience with Terraform, Packer, and the GitHub suite, including GH Actions, workflows, and runners.
- Experience driving large-scale CI improvements across multiple teams or organizations, with the ability to lead ambiguous technical initiatives with minimal direction.
- Deep experience designing and operating complex, large-scale cloud-native platforms or multi-service environments, with strong emphasis on resilience, capacity planning, and cost optimization.
- Track record of leading architecture and design for deployment pipelines, infrastructure platforms, and operational tooling, including defining standards and best practices across teams.
- Advanced expertise in observability, incident management, and reliability engineering, using data-driven approaches to improve SLIs/SLOs and reduce operational toil.
- Experience integrating and operating AI/ML-enabled or AI-assisted systems in production, such as intelligent monitoring, automated remediation, or optimization pipelines, including safely leveraging AI tooling to enhance DevOps workflows.
- Familiarity with AI-driven systems, tools, or workflows and applying AI/ML concepts to real-world products in the context of deployment automation, infrastructure optimization, or reliability engineering.
Details
- Location: Gurgaon, India.
Read the full description and apply on the company’s own careers page.