Jobs Companies MeridianLink Sr. Site Reliability Engineer

À propos de ce poste Sr. Site Reliability Engineer chez MeridianLink

MeridianLink · Télétravail · US Remote

About the Role

We are seeking a Senior Site Reliability Engineer to join our cloud engineering team. You will own the reliability, scalability, and observability of our critical financial SaaS applications and infrastructure, working across cloud platforms to ensure our customers experience is seamless, secure, and performant services. This is a high-impact role for someone who is passionate about building resilient systems and preventing outages before they happen.

Key Responsibilities

  • Design, implement, and maintain Service Level Objectives (SLOs) and Service Level Indicators (SLIs) across all critical systems; ensure we meet or exceed targets consistently

  • Lead observability strategy by designing comprehensive monitoring, logging, and tracing architectures; select and deploy observability tools that provide deep visibility into system behavior

  • Build and own runbooks, incident response procedures, and post-incident review processes; mentor the team on incident management and blameless postmortems

  • Architect and deploy cloud infrastructure on AWS or Azure; implement infrastructure-as-code practices and ensure high availability, disaster recovery, and business continuity

  • Develop automation and AIOps capabilities to reduce toil, accelerate incident detection, and enable self-healing systems; implement intelligent alerting to minimize false positives

  • Drive reliability improvements through load testing, chaos engineering, and failure scenario analysis; identify and eliminate single points of failure

  • Partner with application and backend teams to design reliable systems from inception; conduct architecture reviews and reliability assessments

  • Write production-grade Python tooling for automation, metrics collection, alert management, and operational workflows

  • Champion security and compliance in infrastructure; implement defense-in-depth principles for a regulated fintech environment

Required Qualifications

  • 7+ years in Site Reliability Engineering, DevOps, platform engineering, or closely related roles with significant responsibility for production systems

  • Expert-level experience with Azure or AWS (or both); deep knowledge of compute, networking, storage, and managed services; experience managing infrastructure at scale

  • Demonstrated expertise in observability: designing and implementing monitoring, alerting, logging, and distributed tracing solutions; hands-on with observability platforms (e.g., Prometheus, Grafana, ELK, Datadog, New Relic, or similar)

  • Strong background in SLOs, SLIs, and SLAs; experience defining meaningful objectives and building systems to meet them; understanding of error budgets and their role in prioritization

  • Proven experience designing and troubleshooting highly available, resilient, and scalable systems; deep understanding of distributed systems concepts and failure modes

  • Proficiency in Python, PowerShell, bash, etc. scripting languages for production automation, tooling, and systems programming; ability to write clean, maintainable code for operational workflows

  • Hands-on experience with AIOps practices: event correlation, intelligent alerting, predictive analytics, and automated remediation; familiarity with AIOps platforms is a plus

  • Experience with infrastructure-as-code tools (e.g., Terraform, CloudFormation, Ansible); version control and CI/CD pipeline design

  • Track record of incident management and on-call ownership; comfort with incident response and the ability to remain calm under pressure

  • Excellent communication skills; ability to work cross-functionally and influence without authority; comfort mentoring junior engineers

Preferred Qualifications

  • Experience in the fintech, payments, banking, or other regulated industries; understanding of compliance requirements (SOC 2, PCI-DSS, etc.)

  • Experience with Kubernetes and container orchestration; deep knowledge of containerized application deployment and management

  • Proficiency with observability as code; experience building custom metrics, dashboards, and alerts programmatically

  • Background in chaos engineering or reliability testing; experience using tools like Gremlin or similar platforms

  • Contribution to open-source observability or infrastructure projects

  • Expertise in network security, application security, or infrastructure hardening

  • Experience with database optimization, query performance tuning, and backup/recovery strategies

Prêt à postuler chez MeridianLink ?
Postuler chez MeridianLink

Comment se compare ce salaire pour SRE

Ce poste paie $140,874/yren dessous de la fourchette habituelle pour les postes SRE.

$142,012 la médiane $185,000 $228,750

Fourchette typique $170,002–$214,689/yr, à partir de 18 annonces SRE comparables sur JobsRadar (rémunération annualisée en USD). Voir les aperçus de salaire pour SRE →

Emplois similaires

Affirm
Senior Software Engineer, Backend (Reliability Platform)
Affirm
⚡ Postuler tôt Remote Canada · lieu restreint $153,000–$213,000
● Nouveau 👁 Vu ✓ Postulé il y a 1 j
Affirm
Senior Software Engineer, Backend (Reliability Platform)
Affirm
⚡ Postuler tôt Remote US · lieu restreint $195,000–$255,000
● Nouveau 👁 Vu ✓ Postulé il y a 1 j
Pinterest
Sr. Site Reliability Engineer, tvScientific
Pinterest
⚡ Postuler tôt San Francisco, CA, US; Remote,... · lieu restreint $139,764–$287,749
● Nouveau 👁 Vu ✓ Postulé il y a 2 j
GitLab
Site Reliability Engineer, Cloud Cost Utilization
GitLab
⚡ Postuler tôt Remote, US · lieu restreint
● Nouveau 👁 Vu ✓ Postulé il y a 2 j
GitLab
Site Reliability Engineer, Intermediate to Senior Staff — Infrastructure Platforms
GitLab
⚡ Postuler tôt Remote, Canada; Remote, United... · lieu restreint $126,400–$314,400
● Nouveau 👁 Vu ✓ Postulé il y a 2 j
CR
Senior Infrastructure Engineer/SRE
Cresta
⚡ Postuler tôt United States (Remote) · lieu restreint $205,000–$270,000
● Nouveau 👁 Vu ✓ Postulé il y a 2 j
CR
Infrastructure Engineer/SRE
Cresta
⚡ Postuler tôt Canada (Remote) · lieu restreint
● Nouveau 👁 Vu ✓ Postulé il y a 2 j
Block
Senior Site Reliability Engineer
Block
⚡ Postuler tôt Bay Area, CA, United States of... Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 2 j
Block
Senior Site Reliability Engineer
Block
⚡ Postuler tôt New York, NY, United States of... Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 2 j

Inscrivez-vous pour des suggestions adaptées aux emplois que vous ouvrez et aux recherches que vous enregistrez.

Plus d’emplois chez MeridianLink

Voir tous les emplois chez MeridianLink →

Postuler maintenant
🤖

Doucement — un instant

JobsRadar a été conçu pour de vraies personnes qui traversent une période difficile dans leur recherche d’emploi — pas pour des requêtes automatisées. Vous cliquez beaucoup trop vite et vous êtes maintenant temporairement bloqué.

Revenez plus tard. Si vous cherchez réellement un emploi, nous sommes de votre côté — agissez simplement comme un être humain.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Prenez une longueur d’avance dans votre recherche d’emploi.

Rejoignez notre canal Telegram pour ce qui vous aide à décrocher le poste — références salariales, le pouls hebdomadaire du marché et les annonces de nouveautés. Pas de spam, que du signal.

Rejoindre le canal — c’est gratuit