R

Site Reliability Engineer

Ryan Specialty
NewPosted today

LOCATION

Mumbai · Onsite

EXPERIENCE

5+ Years

TYPE

FullTime

SALARY

Negotiable

SKILLS REQUIRED

Site Reliability EngineeringObservabilityLinux System AdministrationIncident responseKubernetes

Job description

Overview

Join an agile team as a Site Reliability Engineer supporting the reliability, scalability, and efficiency of The Connector, an InsurTech platform for obtaining quotes, issuing policies, creating documents, electronic signing, and data collection. Work at the intersection of software engineering and systems engineering to design, deploy, operate, and continuously improve scalable systems.

What you'll do

  • Design and implement scalable, resilient systems using serverless architecture, event stream processing, Kubernetes, and other modern technologies.
  • Modernize legacy systems by containerizing applications and services and helping migrate them to scalable, cloud-native environments.
  • Develop and maintain monitoring, alerting, and incident-response tools to ensure system reliability and uptime.
  • Collaborate with development teams to use observability feedback to create more robust and dependable solutions.
  • Identify and address potential reliability risks to keep systems secure and performant.
  • Enable industry best practices, including branch by abstraction, safe testing in production, and continuous delivery.
  • Debug and diagnose complex issues across distributed systems and perform root cause analysis.
  • Automate manual and repetitive tasks to improve operational efficiency and reduce potential human error.
  • Contribute to post-incident reviews and retrospectives to improve processes and systems.
  • Champion a culture of reliability and continuous learning.

What you'll need

  • Minimum of 5 years of cumulative experience in Site Reliability Engineering, DevOps, Systems Engineering/Ops, or Software Development roles.
  • 1-2 years of experience programming in one or more programming languages, such as Bash Scripting.
  • 1-2 years of experience working with systems administration (Ops/Infrastructure) or networking.
  • Demonstrated experience working with Linux systems and troubleshooting performance or application-related issues.
  • Experience building and managing containers using Docker, Podman, and Kubernetes, focused on developing scalable, portable, and maintainable applications.
  • Hands-on experience managing large-scale distributed systems.
  • Strong collaboration and communication skills, including the ability to work effectively across time zones and teams.
  • Proven ability to design and implement solutions from inception to production.
  • Experience owning issues and troubleshooting them without an overreliance on others.
  • A Bachelor’s degree is preferred; Software Engineering, Computer Science, or related disciplines are preferred.
  • Equivalent work experience and professional certifications in relevant areas will be considered.

Nice to have

  • Cross-functional experience.
  • Familiarity with cloud platforms such as Google Cloud Platform and Azure.

Details

  • Location: Mumbai, India.

Read the full description and apply on the company’s own careers page.

Stay safe

Hiring on Abekus is free for applicants

We never charge a fee, and employers are prohibited from doing so. If a recruiter asks for payment, please report them right away.

Site Reliability Engineer