Jobs Companies CVS Health Site Reliability Engineer - Observability

Sobre este puesto de Site Reliability Engineer - Observability en CVS Health

CVS Health · Presencial · IRL - Galway

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time.

Position Summary:

Join Fortune 6 CVS Health as a Software Engineer to lead and advance our DevOps, Site Reliability Engineering (SRE), AIOps, Observability, and Monitoring capabilities in the CVS Digital team. This role is critical in advancing intelligent, automated, and scalable reliability practices across our platforms. You will drive the evolution from traditional monitoring to AI-driven operations (AIOps) leveraging automation, machine learning, and advanced analytics to improve system resilience, reduce operational toil, and accelerate incident detection and resolution. As a technical leader, you will influence architecture, build platforms, and mentor teams to embed reliability, observability, and automation into the software delivery lifecycle.

Key Responsibilities:

·DevOps & Platform Engineering:

  • Drive adoption of CI/CD pipelines, Infrastructure as Code (IaC), and GitOps practices.
  • Lead the design and evolution of scalable, automated, and secure platform engineering solutions.
  • Standardize development and deployment workflows across teams.
  • Champion DevOps maturity, developer productivity, and release automation.
  • SRE Strategy & Reliability Engineering
    • Define and implement enterprise-wide SRE practices, including SLIs, SLOs, error budgets, and reliability governance.
    • Drive a culture of reliability, automation, and continuous improvement across engineering teams.
    • Establish metrics-driven approaches to measure system health, availability, and performance.
  • AIOps & Intelligent Operations
    • Lead adoption of AIOps solutions to enable predictive monitoring, anomaly detection, and automated root cause analysis.
    • Integrate machine learning models and analytics into monitoring pipelines to proactively detect and prevent incidents.
    • Develop intelligent alerting systems to reduce noise and improve signal quality.
  • Observability & Monitoring Platforms
    • Design and build scalable observability frameworks covering metrics, logs, traces, and events.
    • Define standards for instrumentation, telemetry collection, and distributed tracing.
    • Enable real-time insights into system performance across microservices and cloud-native architectures.
  • Incident Management & Automation
    • Participate into incident triage, including on-call readiness, RCA, postmortems, and continuous learning loops.
    • Implement runbooks, playbooks, and automated escalations.
  • Platform Engineering & Tooling
    • Develop internal platforms and tools for observability, monitoring, and performance optimization.
    • Integrate observability into CI/CD pipelines to enable proactive quality and reliability checks.
    • Drive infrastructure automation using IaaC frameworks and GitOps principles.
  • Collaboration & Technical Leadership
    • Partner with engineering, platform, and product teams to embed reliability and observability into system design.

Required Qualifications:

  • 1+ years of experience in software engineering, SRE, or production engineering in large-scale distributed systems.
  • Hands-on experience with Observability tools such as AppDynamics, Grafana, Prometheus, Datadog, OpenTelemetry, or similar.
  • Experience with AIOps or intelligent monitoring platforms, including anomaly detection and event correlation.
  • Strong expertise in cloud platforms (AWS, Azure, or GCP), cloud-native architectures (Kubernetes, containers, microservices), and CI/CD pipelines (GitHub Actions, Jenkins).
  • Proficiency in at least one programming language (e.g., Python, Java, Go).
  • Strong understanding of distributed systems, resiliency patterns, and fault tolerance.
  • Experience implementing incident management, on-call processes, and root cause analysis.
  • Hands-on expertise with Infrastructure as Code (Terraform, ARM, CloudFormation) and CI/CD pipelines.
  • Experience using GenAI/Automation tools and frameworks such as OpenAI, CoPilot, Gemini, Claude, MCP etc.
  • Proven ability to design scalable, reliable, and observable systems.

Preferred Qualifications:

  • Strong knowledge of machine learning applications in IT operations (e.g., anomaly detection, forecasting, clustering).
  • Experience defining and managing SLIs/SLOs and error budgets at scale.
  • Experience with OpenTelemetry and modern observability standards.
  • Familiarity with chaos engineering, resilience testing, and fault injection frameworks.
  • Exposure to GenAI-driven operations or AI-assisted troubleshooting tools.
  • Demonstrated leadership in driving cross-functional initiatives and influencing senior stakeholders.
  • Contributions to open-source projects in SRE, observability, or AIOps domains.

Education:

  • Bachelor’s degree or equivalent work experience in Computer Science, Engineering, or related discipline.
  • Certifications in AIOps, SRE, OpenTelemetry, cloud platforms, or DevOps are a plus.

Leadership Competencies:

  • Strategic thinking and execution excellence
  • Strong communication and stakeholder influence
  • Data-driven decision making
  • Continuous improvement mindset

Pay Range

The typical pay range for this role is:

€35,000.00 - €90,000.00

  

We anticipate the application window for this opening will close on: 15/09/2026
¿Listo para postularte en CVS Health?
Postúlate en CVS Health

Cómo se compara este salario de SRE

Este puesto paga $72,637/yrpor debajo de el rango típico para los puestos de SRE.

$96,136 la mediana de $155,527 $228,747

Rango típico $121,488–$193,000/yr, a partir de 815 ofertas comparables de SRE en JobsRadar (salario anualizado en USD). Ver datos salariales de SRE →

Sobre CVS Health

Our Work Experience is the combination of everything that's unique about us: our culture, our core values, our company meetings, our commitment to sustainability, our recognition programs, but most importantly, it's our people. Our employees are self-disciplined, hard working, curious, trustworthy, humble, and truthful. They make choices according to what is best for the team, they live for opportunities to collaborate and make a difference, and they make us the #1 Top Workplace in the area.

Ver todos los empleos en CVS Health →

Empleos similares

Fivetran
Senior Site Reliability Engineer
Fivetran
⚡ Postúlate pronto Dublin, Dublin, Ireland, EMEA Híbrido
● Nuevo 👁 Visto ✓ Postulado hace 2d
Fivetran
Staff Site Reliability Engineer
Fivetran
⚡ Postúlate pronto Dublin, Dublin, Ireland, EMEA Híbrido
● Nuevo 👁 Visto ✓ Postulado hace 2d
NTT
Senior SRE Architect
NTT
⚡ Postúlate pronto hyderabad Presencial
● Nuevo 👁 Visto ✓ Postulado hace 12h
Xcel Energy
Performance Optimization Reliability Engineer Intern- TX
Xcel Energy
⚡ Postúlate pronto Earth, TX, 79031 Híbrido $44,096–$48,048
● Nuevo 👁 Visto ✓ Postulado hace 12h
Airbus
Developer - SRE ( Digital Archiving )
Airbus
⚡ Postúlate pronto Bangalore Area Presencial
● Nuevo 👁 Visto ✓ Postulado hace 12h
Omnissa
Senior Site Reliability Engineer
Omnissa
⚡ Postúlate pronto Bengaluru, India Presencial
● Nuevo 👁 Visto ✓ Postulado hace 12h
CrowdStrike
Engineer II - Site Reliability (Hybrid, IND)
CrowdStrike
⚡ Postúlate pronto India - Bangalore Presencial
● Nuevo 👁 Visto ✓ Postulado hace 12h
Pfizer
Senior Associate, AI Ops Site Reliability Engineer
Pfizer
⚡ Postúlate pronto Greece-Thessaloniki Chortiatis Presencial
● Nuevo 👁 Visto ✓ Postulado hace 12h
NTT
Senior Associate Site Reliability Engineer
NTT
⚡ Postúlate pronto hyderabad Presencial
● Nuevo 👁 Visto ✓ Postulado hace 12h

Regístrate para recibir sugerencias adaptadas a los empleos que abres y las búsquedas que guardas.

Más empleos en CVS Health

Ver todos los empleos en CVS Health →

Postúlate ahora
🤖

Un momento — para

JobsRadar se creó para personas reales que están pasando un mal momento en su búsqueda de empleo — no para solicitudes automatizadas. Estás haciendo clic demasiado rápido y ahora estás bloqueado temporalmente.

Vuelve más tarde. Si de verdad estás buscando empleo, cuentas con nosotros — solo compórtate como una persona.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Toma ventaja en tu búsqueda de empleo.

Únete a nuestro canal de Telegram para lo que te ayuda a conseguir el puesto — referencias salariales, el pulso semanal del mercado y avisos de nuevas funciones. Sin spam, solo señal.

Únete al canal — es gratis