Overview
Senior NOC Engineer role focused on maintaining health, stability, and uptime of production systems. You’ll monitor infrastructure in a 24/7 setup, handle incidents, and drive resolution and continuous improvements.
What you'll do
- Monitor production systems and applications for uptime, performance, and availability.
- Respond to and manage real-time incidents, alerts, and outages.
- Conduct root cause analysis and implement corrective and preventive actions.
- Troubleshoot system, application, and network issues from monitoring and support escalations.
- Participate in 24/7 shift rotations including weekends and holidays.
- Develop and update SOPs, runbooks, and internal knowledge bases.
- Drive post-incident reviews and blameless postmortems for process improvement.
What you'll need
- 2+ years of hands-on Linux/Unix systems administration and network troubleshooting experience.
- Solid understanding of internet/network protocols including DNS, DHCP, TCP/IP, NTP, SMTP, VPNs, HTTPS, TLS, and IPSec.
- Experience monitoring and managing applications like Apache, Tomcat, and MySQL.
- Proficiency scripting in Shell, Python, or Ruby for automation.
- Experience with monitoring/logging tools such as Nagios, Datadog, New Relic, ELK, Splunk, or Sumo Logic.
- Familiarity with incident management platforms like PagerDuty, JIRA, or ServiceNow.
- Hands-on experience with Docker and Kubernetes.
Details
- Location: Hyderabad, India.
- 24/7 shift rotations including weekends and holidays.
Read the full description and apply on the company’s own careers page.