Site Reliability Engineer (Azure Preferred)

FIS · Pune/Pimpri-Chinchwad Area

  • Experience5–6 yrs
  • SalaryNot disclosed
  • Work modeonsite
  • Levelsenior
  • Posted16 Sept 2026

About FIS

FIS is hiring in Pune/Pimpri-Chinchwad Area in financial services. This role looks for around 5+ years of experience.

Skills

  • Infrastructure as Code
  • Terraform
  • Ansible
  • AWS
  • Microsoft Azure
  • Google Cloud Platform
  • Prometheus
  • Grafana
  • Datadog
  • Splunk
  • ELK Stack
  • Python
  • Bash
  • Jenkins
  • GitLab CI/CD
  • Azure DevOps
  • Docker
  • incident management
  • root cause analysis
  • production support

The role

A site reliability engineer at a banking and payments technology company designs reliable cloud-native platforms, strengthens observability, and automates infrastructure with Terraform and Kubernetes. The role improves incident management and CI/CD pipelines for distributed systems and production services.

Full job description

Position Type

Full time

Type Of Hire

Experienced (relevant combo of work and education)

Site Reliability Engineer (Azure Preferred) – 5+ Yrs – Pune Location

About The Role

As a Site Reliability Engineer, you will be responsible for ensuring the reliability, scalability, performance, and availability of mission-critical banking, payments, and capital markets platforms. You will drive automation, strengthen operational resilience, and improve observability across cloud-native and distributed systems. Working closely with engineering, DevOps, security, QA, and product teams, you will help deliver highly available services while reducing operational risk. Success in this role is measured through platform stability, service reliability, incident reduction, and continuous operational improvement.

What You Will Be Doing

Design, implement, and maintain monitoring and observability solutions for infrastructure, applications, and customer experienceBuild and enhance automation frameworks to improve operational efficiency and reduce manual processesEnsure high availability, reliability, scalability, and performance of critical production systemsLead incident management activities, including triage, root cause analysis, recovery, and post-incident reviewsPerform capacity planning, performance tuning, and infrastructure optimization to support business growthDevelop and manage Infrastructure as Code solutions for consistent and scalable cloud deploymentsMaintain and optimize CI/CD pipelines to enable reliable and secure software deliveryCollaborate with security teams to implement platform security controls and compliance best practicesDevelop, validate, and improve disaster recovery, backup, and business continuity strategiesPartner with engineering, DevOps, QA, and product teams to achieve service-level objectives and operational excellenceParticipate in on-call rotations and provide support for critical production environments

What You Bring

5+ years of experience in Site Reliability Engineering, Production Support, Platform Engineering, DevOps, Cloud Operations, or a related fieldHands-on experience with cloud platforms including AWS, Microsoft Azure, or Google Cloud PlatformStrong knowledge of Infrastructure as Code and automation tools such as Terraform and AnsibleExperience supporting web applications, APIs, distributed systems, and modern software architecturesProficiency with monitoring and observability tools including Prometheus, Grafana, Datadog, or similar platformsExperience with logging and analytics solutions such as Splunk, ELK Stack, or equivalent technologiesStrong scripting and automation skills using Python, Bash, or similar programming languagesExperience designing, maintaining, and optimizing CI/CD pipelines using Jenkins, GitLab CI/CD, Azure DevOps, or related toolsKnowledge of containerization technologies such as Docker and container orchestration platformsDemonstrated experience in incident management, root cause analysis, and production support in enterprise environments.

Preferred Qualifications

Experience in applying SRE practices, reliability engineering principles, and service-level management in large-scale environmentsKnowledge of Kubernetes, cloud-native architectures, and microservices-based platformsExperience conducting operational readiness assessments and post-mortem reviewsUnderstanding of disaster recovery, high-availability design patterns, and resiliency engineeringIndustry certifications in AWS, Azure, Google Cloud, Kubernetes, DevOps, or Site Reliability Engineering

What We Offer You

A work environment built on collaboration, flexibility and respectCompetitive salary and attractive range of benefits designed to help support your lifestyle and wellbeing Varied and challenging work to help you grow your technical skillset

Privacy Statement

FIS is committed to protecting the privacy and security of all personal information that we process in order to provide services to our clients. For specific information on how FIS protects personal information online, please see the Online Privacy Notice.

Sourcing Model

Recruitment at FIS works primarily on a direct sourcing model; a relatively small portion of our hiring is through recruitment agencies. FIS does not accept resumes from recruitment agencies which are not on the preferred supplier list and is not responsible for any related fees for resumes submitted to job postings, our employees, or any other part of our company.

#pridepass