Senior Manager - Reliability Engineering
London Stock Exchange Group · Hyderabad
- Experience0–4 yrs
- SalaryNot disclosed
- Work modeonsite
- Posted29 Sept 2026
About London Stock Exchange Group
London Stock Exchange Group is hiring in Hyderabad in financial services. This role looks for around 0+ years of experience.
Skills
- Site Reliability Engineering
- Service Level Objectives
- Service Level Indicators
- error budgets
- automation
- observability
- Incident Management
- Problem Management
- Root Cause Analysis
- ITSM
- Change Management
The role
A reliability engineering manager at a financial markets infrastructure company drives Site Reliability Engineering through Service Level Objectives, observability, and Incident Management, improving resilience and operational effectiveness across enterprise platforms. The role also applies automation and ITSM.
Full job description
ABOUT US:
LSEG (London Stock Exchange Group) is more than a diversified global financial markets infrastructure and data business. We are dedicated, open-access partners with a dedication to excellence in delivering the services our customers expect from us. With extensive experience, deep knowledge and worldwide presence across financial markets, we enable businesses and economies around the world to fund innovation, manage risk and create jobs. Its how weve contributed to supporting the financial stability and growth of communities and economies globally for more than 300 years. Through a comprehensive suite of trusted financial market infrastructure services and our open-access model we provide the flexibility, stability and trust that enable our customers to pursue their ambitions with confidence and clarity.
LSEG is headquartered in the United Kingdom, with significant operations in 65 countries across EMEA, North America, Latin America and Asia Pacific. We employ 25,000 people globally, more than half located in Asia Pacific. LSEGs ticker symbol is LSEG.
OUR PEOPLE:
People are at the heart of what we do and drive the success of our business. Our values of Integrity, Partnership, Excellence and Change shape how we think, how we do things and how we help our people fulfil their potential. We embrace diversity and actively seek to attract individuals with unique backgrounds and perspectives. We break down barriers and encourage teamwork, enabling innovation and rapid development of solutions that make a difference. Our workplace generates an enriching and rewarding experience for our people and customers alike. Our vision is to build an inclusive culture in which everyone feels encouraged to fulfil their potential.
We know that real personal growth cannot be achieved by simply climbing a career ladder which is why we encourage and enable a wealth of avenues and interesting opportunities for everyone to broaden and deepen their skills and expertise. As a global organisation spanning 65 countries and one rooted in a culture of growth, opportunity, diversity and innovation, LSEG is a place where everyone can grow, develop and fulfil your potential with meaningful careers.
ROLE PROFILE:
The Senior Manager - Reliability Engineering is accountable for the reliability, resilience, and operational effectiveness of critical enterprise applications and platforms throughout the service lifecycle.
The role will drive the evolution of the L2/SRE function from reactive production support towards proactive reliability engineering, using automation, observability, operational learning, and strong engineering partnerships to reduce toil, incidents, and operational risk.
Working across engineering, product, infrastructure, security, and support, the role will promote shared ownership of production reliability and ensure services are continuously improved to meet agreed levels of availability, performance, resilience, and customer experience.
What you'll be doing
Lead the L2/SRE function and its continuous improvement, with accountability for improving service reliability, resilience, operational efficiency, and engineering maturity.
Lead, mentor, and develop a high-performing L2/SRE team, fostering accountability, technical capability, operational excellence, and continuous improvement.
Define and mature Service Level Objectives (SLOs), Service Level Indicators (SLIs), and error budgets to support clear, data-led reliability and engineering decisions.
Reduce operational toil through automation, improved tooling, self-service capabilities, and simplification of repetitive support activities.
Establish strong observability and alerting standards across logging, metrics, tracing, dashboards, and telemetry to improve detection, diagnosis, and recovery.
Define operational readiness standards, ensuring services are supportable, resilient, appropriately monitored, well documented, and have clear recovery and escalation procedures before entering production.
Lead effective Incident Management within L2/SRE, including response, escalation, post-incident review, and identification of recurring issues. Partner with L3 on Problem Management and RCA, ensuring corrective actions and reliability improvements are tracked through to completion.
Champion strong ITSM practices and operational hygiene across Incident, Problem, Change, service ownership, documentation, and continuous improvement.
Oversee proportionate, risk-based Change Management and approval processes, ensuring changes are appropriately assessed, dependencies and conflicts are understood, implementation and rollback plans are robust, and higher-risk activity receives the right level of scrutiny while avoiding unnecessary overhead for well-understood, low-risk change.
Partner with L3 engineering and other technology teams to ensure production learnings, recurring incidents, support .