SRE/Senior SRE

Saviynt · Bengaluru

  • Experience5+ yrs
  • SalaryNot disclosed
  • Work modeunknown
  • Levelsenior
  • Posted17 Sept 2026

About Saviynt

Saviynt is hiring in Bengaluru in technology software. This role looks for around 5+ years of experience.

The role

Senior SRE Engineer role at an AI security/identity platform company. You own uptime, reliability, and performance for AWS + Kubernetes services, build self-healing infrastructure with LLM-powered automation, and implement monitoring/alerting, incident response, and observability. Required skills include AWS, Kubernetes, Python or Go, OpenAI API, LangChain or AutoGen, Prometheus, Grafana, ELK/OpenSearch, OpenTelemetry, Terraform, CI/CD, and incident management. Location: Bengaluru, Karnataka, India; work mode not specified.

Full job description

Saviynt's AI-powered identity platform manages and governs human and non-human access to all of an organization's applications, data, and business processes. Customers trust Saviynt to safeguard their digital assets, drive operational efficiency, and reduce compliance costs. Built for the AI age, Saviynt is today helping organizations safely accelerate their deployment and usage of AI. Saviynt is recognized as the leader in identity security, with solutions that protect and empower the world’s leading brands, Fortune 500 companies and government institutions. For more information, please visit www.saviynt.com.

We’re a fast-moving AI Security Company building AI-native infrastructure and applications powered by LLMs and autonomous agents. Our stack is deeply integrated with AWS, Kubernetes, and OpenAI-based systems, and we’re rethinking reliability in a world where software can reason, adapt, and self-heal.

We’re hiring a Senior SRE Engineer to own reliability across our cloud-native and AI-driven platform. You’ll work at the intersection of distributed systems, Kubernetes operations, and LLM-powered automation, building systems that don’t just scale—but think and fix themselves.

WHAT YOU BRING

5+ years in SRE / DevOps / Platform Engineering

Strong hands-on experience with:

AWS infrastructure at scale

Kubernetes (production-grade clusters)

Proven ability to debug complex distributed systems under pressure

Strong coding skills (Python or Go)—you build internal platforms and tools

Experience implementing monitoring, alerting, and incident management systems

Bonus (AI / LLM Focus)

Experience working with LLM APIs such as the OpenAI API

Familiarity with agent frameworks like:

LangChain

AutoGen

Built or experimented with:

AI agents for DevOps / SRE workflows

Retrieval-Augmented Generation (RAG) systems

Vector databases (Pinecone, Weaviate, etc.)

Exposure to AIOps or intelligent automation systems

WHAT YOU WILL BE DOING

Own uptime, reliability, and performance of services running on AWS + Kubernetes (EKS)

Design and implement self-healing infrastructure using automation and AI agents

Build LLM-powered operational tooling using APIs such as the OpenAI API for:

Intelligent alert triage

Incident summarization

Root cause analysis

Runbook automation

Manage and scale Kubernetes workloads:

Deployments, autoscaling, resource optimization

Cluster reliability and cost efficiency

Build and evolve observability systems:

Metrics (Prometheus), dashboards (Grafana)

Logs (ELK / OpenSearch)

Tracing (OpenTelemetry)

Define and enforce SLOs, SLAs, and error budgets tied to business metrics

Automate infrastructure using Terraform and CI/CD pipelines

Lead incident response, postmortems, and continuous reliability improvements

Introduce chaos engineering practices to proactively test system resilience

If required for this role, you will:- Complete security & privacy literacy and awareness training during onboarding and annually thereafter- Review (initially and annually thereafter), understand, and adhere to Information Security/Privacy Policies and Procedures such as (but not limited to):

> Data Classification, Retention & Handling Policy > Incident Response Policy/Procedures > Business Continuity/Disaster Recovery Policy/Procedures > Mobile Device Policy > Account Management Policy > Access Control Policy > Personnel Security Policy > Privacy Policy

Saviynt is an amazing place to work. We are a high-growth, Platform as a Service company focused on Identity Authority to power and protect the world at work. You will experience tremendous growth and learning opportunities through challenging yet rewarding work which directly impacts our customers, all within a welcoming and positive work environment. If you're resilient and enjoy working in a dynamic environment you belong with us!

Saviynt is an equal opportunity employer and we welcome everyone to our team.  All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or veteran status.

We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.