Jobs Companies Lloyds Banking Group Principal SRE

Über diese Principal SRE Stelle bei Lloyds Banking Group

Lloyds Banking Group · Vor Ort · Hyderabad Knowledge Park Tower 2

End Date

Tuesday 29 September 2026

We Support Flexible Working – Click here for more information on flexible working options

Flexible Working Options

Hybrid Working

Job Description Summary

A Lead SRE is accountable for a complex area of the cloud infrastructure resources managing the SLOs through the work of their product team. Advocate for best approach to apply SRE for their technical resources and collaborating with the product teams and the application teams consuming them

Job Description

Job Description: Lead Site Reliability Engineer (F)

Lead Site Reliability Engineer (F)

15+ Years of experience

Hyderabad Location



Job Description

A Lead Site Reliability Engineer (SRE) proactively ensures the reliability, availability, scalability, and performance of products deployed in production environments. The role combines software engineering and operational expertise to build, operate, and continuously improve highly resilient digital services.

The Lead SRE is accountable for defining and managing Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Error Budgets to ensure services consistently meet business and customer expectations. The role applies software engineering principles to operations, leveraging automation, cloud-native technologies, observability platforms, and AI-powered operational capabilities to improve reliability, reduce operational toil, and enhance customer experience.

The Lead SRE acts as a senior technical leader across one or more products, partnering with engineering, platform, and architecture teams to embed reliability, resilience, observability, and operational excellence into solution design and delivery. The role provides technical leadership during major incidents, problem investigations, and service recovery activities, driving improvements that increase Mean Time To Failure (MTTF) and reduce Mean Time To Restore (MTTR).


Role Responsibilities


In addition to the responsibilities of a Senior Site Reliability Engineer:

  • Own and drive the reliability strategy for one or more business-critical products and services.
  • Define, implement, and govern Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Error Budget policies.
  • Lead major incident management, service restoration, and post-incident reviews to drive continual service improvement.
  • Develop and maintain automation, operational tooling, integrations, and self-healing capabilities using Java and/or Python.
  • Drive adoption of cloud-native engineering practices, Infrastructure as Code, CI/CD, and operational automation.
  • Champion observability through monitoring, logging, tracing, telemetry analytics, dashboards, and performance engineering.
  • Leverage AI-powered operational capabilities to improve incident detection, root cause analysis, predictive reliability, and service restoration.
  • Evaluate and recommend new tools, technologies, and engineering practices that improve reliability and operational efficiency.
  • Collaborate with Engineering, Platform, Architecture, and Product teams to embed reliability, resilience, and operability into solution designs.
  • Act as a technical leader and mentor, sharing knowledge and developing engineering capability across teams.
  • Communicate effectively with technical and business stakeholders, influencing reliability investment and prioritisation decisions.
  • Support resilience testing, disaster recovery exercises, operational readiness reviews, and capacity planning activities

Skill


Description

Weighting

SRE & Service Engineering

Uses deep expertise in reliability engineering, Service Level Objectives (SLOs), Service Level Indicators (SLIs), Error Budgets, incident management, problem management, resilience engineering, and continuous improvement to enhance product reliability and customer experience.

30%

Software Engineering & Automation (Java/Python)

Develops automation, operational tooling, APIs, integrations, and self-healing capabilities using Java and/or Python. Applies software engineering principles to reduce operational toil and improve service reliability and efficiency.

25%

Cloud Platform Engineering &

Designs, operates, and optimises cloud-native platforms and services using Kubernetes/OpenShift, Infrastructure as Code, CI/CD, and modern operational practices to deliver scalable, secure, and highly available solutions.

20%

Observability

Leverages observability platforms, telemetry analytics, AIOps, and AI-assisted operational capabilities to improve service visibility, incident detection, root cause analysis, predictive insights, and automated remediation.

15%

Technical

Provides technical leadership, mentoring, and strategic direction across engineering teams. Influences reliability roadmaps, engineering standards, and adoption of SRE best practices whilst fostering a culture of operational excellence.

10%

Bereit, sich bei Lloyds Banking Group zu bewerben?
Bei Lloyds Banking Group bewerben

Über Lloyds Banking Group

With 320 years under our belt, we're used to change, and today is no different. Join us and help drive this change, shaping the future of finance whilst working at pace to deliver for our customers. Here, you'll do the best work of your career. Your impact will be amplified by our scale as you learn and develop, gaining skills for the future.

Alle Jobs bei Lloyds Banking Group ansehen →

Ähnliche Jobs

Lloyds Banking Group
Senior SRE
Lloyds Banking Group
⚡ Früh bewerben Hyderabad Knowledge Park Tower... Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 1 Mon.
Lloyds Banking Group
Senior Site Reliability Engineer
Lloyds Banking Group
⚡ Früh bewerben Hyderabad Knowledge Park Tower... Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 1 Mon.
Lloyds Banking Group
Senior Site Reliability Engineer
Lloyds Banking Group
⚡ Früh bewerben Hyderabad Knowledge Park Tower... Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 2 Mon.
Raiffeisen Bank Ukraine
Middle Site Reliability Engineer
Raiffeisen Bank Ukraine
⚡ Früh bewerben Kyiv, Kyiv city, Ukraine Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 12 Min.
GoReel
Junior Site Reliability Engineer
GoReel
⚡ Früh bewerben Warsaw, Masovian Voivodeship,... · standortgebunden
● Neu 👁 Gesehen ✓ Beworben vor 44 Min.
Federal Reserve System
Cloud AWS Site Reliability Engineer (SRE)
Federal Reserve System
⚡ Früh bewerben New York, NY Vor Ort $160,000–$230,000
● Neu 👁 Gesehen ✓ Beworben vor 1 Std.
StraitsX
Senior Site Reliability Engineer
StraitsX
⚡ Früh bewerben Jakarta, Jakarta, Indonesia Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 2 Std.
GoGuardian
Senior Site Reliability Engineer
GoGuardian
⚡ Früh bewerben India Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 2 Std.
NICE
Senior Site Reliability Engineer
NICE
⚡ Früh bewerben United Kingdom - London; Unite... Hybrid
● Neu 👁 Gesehen ✓ Beworben vor 2 Std.

Registrieren für Vorschläge, die auf die von Ihnen geöffneten Jobs und gespeicherten Suchen zugeschnitten sind.

Mehr Jobs bei Lloyds Banking Group

Alle Jobs bei Lloyds Banking Group ansehen →

Jetzt bewerben
🤖

Moment — langsam

JobsRadar wurde für echte Menschen gebaut, die eine schwere Zeit bei der Jobsuche haben — nicht für automatisierte Anfragen. Sie klicken viel zu schnell und sind jetzt vorübergehend blockiert.

Kommen Sie später wieder. Wenn Sie wirklich auf Jobsuche sind, stehen wir hinter Ihnen — verhalten Sie sich einfach wie ein Mensch.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Verschaffe dir einen Vorsprung bei der Jobsuche.

Tritt unserem Telegram-Kanal bei für das, was dir hilft, die Stelle zu bekommen — Gehaltsbenchmarks, den wöchentlichen Marktpuls und neue Feature-Drops. Kein Spam, nur Signal.

Dem Kanal beitreten — kostenlos