Overview
Lead production support quality engineering for critical production systems in a highly regulated financial services environment. Oversee quality, stability, performance, technical leadership, people management, incident prevention, and operational excellence.
What you'll do
- Lead, mentor, and manage a team of Production Support Quality Engineers while fostering collaboration, accountability, and continuous improvement.
- Develop and implement testing strategies for production support, including monitoring, alerting, incident response validation, performance, and resilience.
- Lead production incident analysis, identify root causes, and implement preventive measures and automated solutions to reduce recurrence.
- Partner with development, operations, security, and business teams to understand system architecture, data flows, and critical business processes.
- Define and maintain quality metrics, dashboards, and reporting that measure production system health, performance, reliability, and support effectiveness.
- Champion incident management, problem management, change management, and continuous improvement practices.
- Provide leadership and technical guidance during high priority incidents.
- Participate in an on call rotation as needed.
- Evaluate and implement tools and technologies that strengthen monitoring, automation, troubleshooting, and production support efficiency.
- Ensure production support practices comply with applicable regulatory requirements and internal security policies.
- Conduct performance reviews, provide regular feedback, and support team members’ professional development.
What you'll need
- Bachelor’s degree in Computer Science, Engineering, or a related field.
- 8 or more years of experience in quality assurance, production support, or a related technical role.
- At least three years of leadership or people management experience.
- Experience supporting production systems in financial services or another highly regulated industry.
- Demonstrated expertise with software development lifecycle practices.
- Demonstrated expertise with testing methodologies.
- Demonstrated expertise with production monitoring tools.
- Demonstrated expertise with incident management systems.
- Demonstrated expertise with at least one scripting or programming language.
- Experience with relational or nonrelational databases, including writing complex queries for data analysis and troubleshooting.
- Technical leadership and team development.
- Analytical problem solving and sound decision making.
- Clear communication with technical and nontechnical stakeholders.
- Operational discipline, accountability, and attention to detail.
- Collaboration and continuous improvement.
Nice to have
- Master’s degree in Computer Science, Engineering, or a related field.
- Experience with cloud platforms such as AWS, Azure, or Google Cloud Platform.
- Experience with containerization and orchestration technologies such as Docker and Kubernetes.
- Hands on experience with monitoring platforms such as Splunk, Dynatrace, Prometheus, or Grafana.
Details
- Location: Hyderabad.
- Participate in an on call rotation as needed.
Read the full description and apply on the company’s own careers page.