DevOps lead
Comviva · Bengaluru
- Experience8–9 yrs
- SalaryNot disclosed
- Work modeonsite
- Levelexecutive
- Posted10 Sept 2026
About Comviva
Comviva is hiring in Bengaluru in technology software. This role looks for around 8+ years of experience.
Skills
- CI/CD
- Jenkins
- GitLab CI/CD
- GitHub Actions
- Azure DevOps
- Infrastructure as Code
- Terraform
- Kubernetes
- Helm
- Containerization
- AWS
- Microsoft Azure
- Google Cloud Platform
- Linux
- Bash
- Python
- Groovy
- High Availability
- Disaster Recovery
- Observability
- HashiCorp Vault
- DevSecOps
- SAST
- Software Composition Analysis
- DAST
- DORA metrics
The role
A site reliability and platform engineer at a technology software company designs and operates CI/CD platforms, Infrastructure as Code, and Kubernetes deployments for enterprise applications, with high availability and disaster recovery. The role also applies DevSecOps and observability practices to improve platform reliability and software delivery.
Full job description
Key Responsibilities
Own and drive the end-to-end DevOps and Platform Engineering strategy for building, releasing, deploying, and operating enterprise applications across environments.Design, implement, and manage scalable CI/CD pipelines using industry-standard tools, ensuring consistency, reusability, reliability, and deployment efficiency.Establish and maintain Infrastructure as Code (IaC) practices using tools such as Terraform or equivalent, including infrastructure provisioning, environment management, and governance.Define and implement deployment standards, release management processes, versioning strategies, promotion workflows, rollback mechanisms, and deployment automation.Manage containerized and non-containerized application deployments across Kubernetes, OpenShift, cloud, and on-premises environments.Own configuration management, secrets management, certificate lifecycle management, and automated database schema deployment processes.Design and maintain observability frameworks, including monitoring, logging, tracing, alerting, dashboards, and operational health reporting.Define and monitor Service Level Indicators (SLIs), Service Level Objectives (SLOs), and operational KPIs to improve platform reliability.Design, implement, and operate High Availability (HA) and Disaster Recovery (DR) solutions, including Active-Active and Active-Passive architectures.Define, monitor, and achieve Recovery Time Objectives (RTO) and Recovery Point Objectives (RPO) through effective backup, replication, recovery, and failover strategies.Conduct periodic disaster recovery drills, failover testing, restoration exercises, and resilience validation activities.Integrate security controls into CI/CD pipelines, including code quality checks, vulnerability scanning, software composition analysis, and container security validation.Enforce security best practices, access controls, credential management, and compliance requirements across the DevOps ecosystem.Measure, analyze, and improve software delivery performance through DORA metrics and platform engineering KPIs.Lead technical design reviews, establish engineering standards, drive platform modernization initiatives, and mentor engineers across the organization.Collaborate with Development, QA, Security, Architecture, Infrastructure, and Product teams to deliver scalable, secure, and reliable platforms.
Mandatory Skills
8+ years of experience in DevOps, Platform Engineering, Site Reliability Engineering (SRE), or related domains.Strong hands-on experience in designing and managing enterprise-scale CI/CD platforms using Jenkins, GitLab CI/CD, GitHub Actions, Azure DevOps, or similar tools.Deep expertise in Infrastructure as Code (IaC) using Terraform or equivalent technologies.Hands-on experience with Kubernetes orchestration platforms and Helm-based application deployments.Strong understanding of containerization technologies, deployment automation, and platform operations.Experience working with public cloud platforms such as AWS, Azure, or GCP and/or private cloud and on-premises environments.Strong knowledge of cloud infrastructure, networking, load balancing, DNS, IAM, security controls, and environment management.Proven experience in designing and managing High Availability (HA) and Disaster Recovery (DR) solutions.Practical experience with Active-Active and/or Active-Passive deployment topologies, failover, failback, backup, and restore processes.Hands-on experience with observability platforms including monitoring, logging, tracing, alerting, and dashboarding solutions.Experience with secrets management solutions such as HashiCorp Vault or cloud-native secret management services.Strong Linux administration, troubleshooting, and automation skills.Proficiency in scripting languages such as Bash, Python, Groovy, or similar.Experience implementing DevSecOps practices including SAST, SCA, DAST, container security scanning, and vulnerability management.Strong understanding and practical application of DORA metrics and software delivery performance measurement.Proven ability to simplify, standardize, optimize, and scale engineering platforms and deployment processes.Excellent problem-solving, stakeholder management, communication, and technical leadership capabilities.
Desirable Skills
Experience with database change management and migration tools such as Liquibase or Flyway.Exposure to multiple database technologies such as PostgreSQL, Oracle, MySQL, or SQL Server.Experience with messaging and event-streaming platforms such as Kafka, RabbitMQ, or equivalent technologies.Hands-on experience managing API gateways and ingress platforms such as Kong, NGINX, APISIX, KrakenD, or similar solutions.Knowledge of service mesh technologies and network policy frameworks such as Istio, Linkerd, or Cilium.Experience with GitOps practices and tools such as ArgoCD, FluxCD, or equivalent platforms.Experience implementing progressive deployment and automated rollback strategies.Understanding of multi-region, multi-cluster, and multi-cloud deployment architectures.Experience working within regulated industries such as Banking, Financial Services, FinTech, Telecommunications, Healthcare, or other compliance-driven environments.Exposure to chaos engineering, resilience testing, game-day exercises, and platform reliability programs.Knowledge of cloud cost optimization, FinOps practices, and operational efficiency management.Relevant certifications in Cloud, Kubernetes, DevOps, Terraform, or Platform Engineering technologies.