Jobs Companies Lloyds Banking Group Senior Site Reliability Engineer

Sobre este puesto de Senior Site Reliability Engineer en Lloyds Banking Group

Lloyds Banking Group · Presencial · Hyderabad Knowledge Park Tower 2

End Date

Friday 11 September 2026

We Support Flexible Working – Click here for more information on flexible working options

Flexible Working Options

Hybrid Working

Job Description Summary

Our Site Reliability Engineering (SRE) team plays a key role in improving the reliability, resilience and operational excellence of the Analytics & AI Platform.

As a Site Reliability Engineer, you’ll work closely with Product, Engineering and Production Support teams to improve service reliability, reduce operational risk and embed Site Reliability Engineering principles across our platforms. You’ll use engineering, automation and observability to improve service resilience, reduce operational toil and support the delivery of highly available services.

As an experienced engineer, you’ll have opportunities to mentor colleagues, lead technical initiatives and, where appropriate, take on line management responsibilities to support the growth and development of engineers within the team

Job Description

  • Role: Senior Site Reliability Engineer
    Experience: 9-15 years
    Location: Hyderabad
    Job Type: Full Time
     
  • What you’ll do 

    • Design, implement and continuously improve monitoring, alerting and observability solutions that enable reliable production services. 

    • Define, measure and continuously improve Service Level Indicators (SLIs), Service Level Objectives (SLOs) and Error Budgets to drive service reliability. 

    • Work collaboratively with Production Support and engineering teams to investigate complex production incidents and support service restoration. 

    • Support Post Incident Reviews (PIRs) and Problem Management activities by identifying reliability improvements and driving preventative engineering actions. 

    • Reduce operational toil through automation, tooling and engineering improvements, enabling teams to focus on higher-value engineering work. 

    • Work closely with Production Support teams to develop and continuously improve operational runbooks, support processes and service readiness. 

    • Partner with engineering teams to embed reliability, resilience and operational best practices throughout the software development lifecycle. 

    • Collaborate with Product and Engineering teams to improve service operability, ensuring applications are designed with appropriate monitoring, alerting, diagnostics and operational documentation before entering production support. 

    • Identify reliability risks and proactively deliver engineering improvements that reduce incidents and improve service resilience. 

    • Support the onboarding of new applications into the SRE operating model by ensuring agreed reliability, observability and operational standards are achieved. 

    • Coach and mentor engineers, sharing knowledge and promoting engineering excellence across the platform. 

    • Contribute to the evolution of SRE standards, tooling and engineering practices across the Analytics & AI Platform. 

    • Where appropriate, provide line management, coaching and performance development for engineers within the team. 

     

    What you’ll need 

     

    Essential 

    • Strong understanding of Site Reliability Engineering principles and practices. 

    • Experience designing and implementing observability solutions, including monitoring, logging and alerting. 

    • Experience defining and improving Service Level Indicators (SLIs), Service Level Objectives (SLOs) and Error Budgets. 

    • Strong troubleshooting and technical investigation skills across complex production environments. 

    • Experience supporting incident management, Post Incident Reviews and Problem Management activities. 

    • Experience identifying operational toil and delivering automation to improve reliability and efficiency. 

    • Experience using Infrastructure as Code and CI/CD tooling. 

    • Strong scripting or programming skills in one or more languages such as Python, Java, JavaScript, PowerShell or Bash. 

    • Experience working with cloud technologies and modern application platforms. 

    • Strong understanding of cloud security, networking and operational resilience. 

    • Excellent stakeholder management, communication and collaboration skills. 

    • Ability to work effectively across Product, Engineering and Production Support teams. 

     

    Desirable 

    • Experience with Kubernetes and containerised platforms. 

    • Experience with Azure, Google Cloud Platform or AWS. 

    • Experience using observability platforms such as Dynatrace, Grafana, Prometheus, ELK or Splunk. 

    • Experience working within Financial Services or another regulated industry. 

    • Experience using Jira, Confluence and Agile delivery practices. 

    • Experience mentoring or coaching engineers. 

    • Previous line management experience or a desire to develop people leadership capability. 

    • Relevant Cloud, DevOps or SRE certifications

¿Listo para postularte en Lloyds Banking Group?
Postúlate en Lloyds Banking Group

Sobre Lloyds Banking Group

With 320 years under our belt, we're used to change, and today is no different. Join us and help drive this change, shaping the future of finance whilst working at pace to deliver for our customers. Here, you'll do the best work of your career. Your impact will be amplified by our scale as you learn and develop, gaining skills for the future.

Ver todos los empleos en Lloyds Banking Group →

Empleos similares

Lloyds Banking Group
Senior SRE
Lloyds Banking Group
⚡ Postúlate pronto Hyderabad Knowledge Park Tower... Presencial
● Nuevo 👁 Visto ✓ Postulado hace 1 mes
Lloyds Banking Group
Senior Site Reliability Engineer
Lloyds Banking Group
⚡ Postúlate pronto Hyderabad Knowledge Park Tower... Presencial
● Nuevo 👁 Visto ✓ Postulado hace 1 mes
Lloyds Banking Group
Site Reliability Engineer
Lloyds Banking Group
⚡ Postúlate pronto Hyderabad Knowledge Park Tower... Presencial
● Nuevo 👁 Visto ✓ Postulado hace 1 mes
Genomics
Staff Site Reliability Engineer
Genomics
⚡ Postúlate pronto London Híbrido
● Nuevo 👁 Visto ✓ Postulado hace 2h
NXP Semiconductors
Non-Volatile Memory Reliability Engineer
NXP Semiconductors
⚡ Postúlate pronto Hsinchu Presencial
● Nuevo 👁 Visto ✓ Postulado hace 4h
Planet
Senior Site Reliability Engineer
Planet
⚡ Postúlate pronto Berlin, Germany; Haarlem, Neth... Híbrido €77,000–€96,300
● Nuevo 👁 Visto ✓ Postulado hace 4h
SL
Staff Site Reliability Engineer
Sumo Logic
⚡ Postúlate pronto Bangalore, Karnataka, India Presencial
● Nuevo 👁 Visto ✓ Postulado hace 5h
SL
Staff Site Reliability Engineer
Sumo Logic
⚡ Postúlate pronto Noida, Uttar Pradesh, India Presencial
● Nuevo 👁 Visto ✓ Postulado hace 5h
Roblox
Senior Site Reliability Engineer, Compute
Roblox
⚡ Postúlate pronto San Mateo, CA, United States Presencial $243,290–$295,250
● Nuevo 👁 Visto ✓ Postulado hace 6h

Regístrate para recibir sugerencias adaptadas a los empleos que abres y las búsquedas que guardas.

Más empleos en Lloyds Banking Group

Ver todos los empleos en Lloyds Banking Group →

Postúlate ahora
🤖

Un momento — para

JobsRadar se creó para personas reales que están pasando un mal momento en su búsqueda de empleo — no para solicitudes automatizadas. Estás haciendo clic demasiado rápido y ahora estás bloqueado temporalmente.

Vuelve más tarde. Si de verdad estás buscando empleo, cuentas con nosotros — solo compórtate como una persona.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Toma ventaja en tu búsqueda de empleo.

Únete a nuestro canal de Telegram para lo que te ayuda a conseguir el puesto — referencias salariales, el pulso semanal del mercado y avisos de nuevas funciones. Sin spam, solo señal.

Únete al canal — es gratis