Overview
Senior Manager, Data Reliability Engineering (Site Reliability Engineering) for Visa’s data platforms and data-intensive production systems.
What you'll do
- Lead reliability strategy and operational standards for data pipelines, platforms, and distributed processing systems.
- Oversee observability practices, incident management, disaster recovery, and reliability engineering for data services.
- Drive automation and operational excellence practices including monitoring, alerting, and service-level objectives.
- Guide troubleshooting of complex production issues across data pipelines, streaming, storage, and infrastructure services.
- Partner with engineering, product, security, risk, compliance, and business stakeholders to meet regulatory and governance expectations.
- Coach and mentor SRE/infrastructure/platform/operations engineering teams supporting the Data organization.
What you'll need
- 8+ years of relevant work experience (or 11+ years with a Bachelor’s OR other stated education combination).
- Experience leading teams designing, developing, and deploying large-scale engineering solutions.
- Experience with cloud infrastructure, distributed systems, observability platforms, and reliability engineering practices.
- Experience building scalable infrastructure and automation frameworks, including incident response processes and service reliability programs.
- Programming/scripting and infrastructure tool experience including Python, Go, Bash, Terraform, Kubernetes, CI/CD, and Git.
- Experience implementing security, compliance, access control, disaster recovery, and operational risk management standards.
Details
- Hybrid position; number of office days to be confirmed by the hiring manager.
- Location: Bengaluru, India.
Read the full description and apply on the company’s own careers page.