FIS logo

Production Support Analyst I(AWS and Linux/UNIX administration)

FIS
NewPosted today

LOCATION

PUNE FL7 · Onsite

EXPERIENCE

3+ Years

TYPE

FullTime

SALARY

Negotiable

SKILLS REQUIRED

AWSLinux/UNIX AdministrationProduction MonitoringIncident responseRoot Cause AnalysisObservability

Job description

Overview

Engineer reliable, secure and scalable cloud services by combining production operations with software engineering, automation, observability and cloud platform expertise. Own service reliability outcomes across the application lifecycle and partner with development, infrastructure, security and support teams to prevent incidents, reduce toil and improve customer experience.

What you'll do

  • Own the reliability, availability, performance and operational readiness of cloud-hosted applications and platform services.
  • Define and maintain service level indicators (SLIs), service level objectives (SLOs), availability targets and actionable service-health dashboards.
  • Build monitoring and alerting around customer-impacting symptoms, golden signals and service dependencies; reduce alert noise and improve diagnostic quality.
  • Automate repeatable operational work, remediation, deployments, configuration, evidence collection and environment validation using code and pipelines.
  • Design, build and maintain cloud infrastructure using infrastructure as code, reusable modules, policy guardrails and secure engineering standards.
  • Participate in incident response and on-call support, including rapid triage, stabilisation, technical escalation and clear stakeholder communication.
  • Lead or contribute to blameless post-incident reviews; identify root causes, track corrective actions and engineer controls that prevent recurrence.
  • Partner with application engineering teams on architecture, capacity planning, performance engineering, release readiness and production operability.
  • Improve deployment safety through CI/CD controls, automated testing, progressive validation, rollback strategies and change-risk reduction.
  • Engineer resilience through high-availability patterns, backup and restore validation, disaster-recovery runbooks, failover testing and dependency mapping.
  • Manage production risks including vulnerabilities, patching, certificates, secrets, access controls, audit evidence and cloud security findings.
  • Create and maintain runbooks, standard operating procedures, architecture records and operational knowledge that support consistent 24x7 service delivery.
  • Analyse operational data and trends to reduce mean time to detect and restore service, eliminate recurring failure modes and improve capacity and cost efficiency.
  • Mentor support and engineering colleagues in SRE practices, automation, observability, troubleshooting and operational ownership.

What you'll need

  • Three or more years of experience in cloud engineering, production engineering, DevOps, platform engineering or Site Reliability Engineering.
  • Hands-on experience operating production workloads on AWS and/or Microsoft Azure, including compute, networking, identity, storage, managed databases and monitoring services.
  • Strong Linux/UNIX administration and troubleshooting skills.
  • Practical experience with Kubernetes and containers, including EKS and/or AKS, Docker, deployment troubleshooting and workload reliability.
  • Infrastructure-as-code experience using Terraform; ability to build reusable, governed and maintainable modules.
  • CI/CD experience with tools such as Harness, Azure DevOps, GitHub Actions, Jenkins or equivalent, including deployment and rollback controls.
  • Programming or advanced automation capability using Python, PowerShell, Bash or a comparable language; coding experience beyond simple one-off scripts.
  • Experience with observability platforms and practices covering metrics, logs, traces, dashboards, alerting and application performance monitoring.
  • Strong incident and problem management experience, including technical triage, root cause analysis, corrective actions and production communications.
  • Understanding of distributed systems, scalability, high availability, capacity management, performance bottlenecks and failure modes.
  • Experience supporting SQL Server and cloud database services, including connectivity, performance diagnostics, backup/restore and operational monitoring.
  • Working knowledge of cloud security, least privilege, certificate and secrets management, vulnerability remediation, auditing and compliance controls.
  • Experience with ServiceNow or a comparable IT service management platform for incidents, problems, changes and operational work tracking.
  • Clear written and verbal communication, disciplined documentation, strong ownership and the ability to work across engineering and business teams.

Nice to have

  • Cloud certification in AWS or Microsoft Azure; Kubernetes or Terraform certification is advantageous.
  • Experience with Akamai, API gateways, web application delivery, DNS, load balancing, VPNs and enterprise network connectivity.
  • Knowledge of event-driven and service-oriented architectures, domain-driven design and messaging platforms.
  • Experience in regulated financial services or another environment with formal change, risk, audit, resilience and data-protection obligations.
  • Experience implementing policy as code, security scanning, automated compliance controls, FinOps or cloud cost optimisation.
  • Experience supporting globally distributed services, customer onboarding and follow-the-sun operational models.
  • Working knowledge of Windows Server is beneficial.

Details

  • Location: Pune, India.
  • Flexible to work on rotational shifts and weekends too.
  • Participate in an agreed on-call or out-of-hours support rotation where required for production services.
  • Work closely with global engineering, production support, security and service-management teams.
  • Support consistent 24x7 service delivery.

Read the full description and apply on the company’s own careers page.

Stay safe

Hiring on Abekus is free for applicants

We never charge a fee, and employers are prohibited from doing so. If a recruiter asks for payment, please report them right away.

Production Support Analyst I(AWS and Linux/UNIX administration)