Jobs Companies The Access Group Senior Site Reliability Engineer

Über diese Senior Site Reliability Engineer Stelle bei The Access Group

The Access Group · Hybrid · United States Remote

We’re looking for people to join the Access family, who share our passion for believing in better, and who will help us continue to grow.   

Love Work. Love Life. Be You. - is central to our success and how we give our customers the freedom to do more of what's important to them. 

What does Access offer you? 

We offer a blended approach to office working, encouraging you to collaborate and connect in one of our thriving offices. We deliver on what we say, taking the development of our people seriously. We’ll work with you to progress your success plan and provide opportunities to accelerate your career. 

On top of a competitive salary, you’ll receive 22 days paid time off, plus 11 company paid holidays. Also, medical, dental & vision insurance, 5% 401(k) company match, plus a range of other benefits that you can choose from.

About You:

You think in systems, not just technologies. You are the engineer your peers escalate to when the problem is hard, the blast radius is unclear, and the path forward requires both depth and judgment. You are equally at home designing cloud architecture, leading a complex P1 incident response, and writing the Terraform module that ensures it never happens again.

You bring discipline to post-incident reviews, urgency to production issues, and genuine satisfaction to the quieter work — a clean runbook, a well-structured SLO, a piece of toil that no longer exists. You don't separate strategy from operations. You understand that the best Site Reliability Engineers do both, every day, at a senior level.

You are a clear communicator who can lead a technical architecture review with engineers and then explain the same decision to an executive stakeholder without losing either audience. You lead through credibility, influence without authority, and take ownership of outcomes — not just tasks. If this sounds like the way you already work, we'd love to talk.

Day-to-day, you will:

  • Serve as the senior escalation point for complex production incidents (P1/P2), owning cross-system triage and leading permanent architectural remediation — not just tactical fixes.

  • Lead platform-level architecture reviews, ensuring cloud infrastructure designs meet reliability, scalability, security, and operational standards before implementation.

  • Identify systemic failure patterns across incidents and translate them into architectural changes, design standards, and lasting platform improvements.

  • Own the availability, reliability, performance, and scalability of production systems on a daily basis.

  • Define, track, and improve Service Level Objectives (SLOs), Service Level Indicators (SLIs), and operational KPIs across critical services.

  • Develop and maintain Infrastructure-as-Code (IaC) solutions using Terraform, including module design, state management, and governance standards.

  • Identify and systematically eliminate operational toil through automation, self-service capabilities, and platform-level tooling.

  • Build and maintain automation frameworks and operational tooling using Bash, PowerShell, and related scripting technologies.

  • Administer and architect solutions within Microsoft Azure as the primary cloud platform, with working knowledge of AWS.

  • Operate Kubernetes in production, including cluster management, workload operations, and platform-level maintenance.

  • Manage hybrid-cloud environments including virtual machines, networking, and distributed infrastructure.

  • Maintain and improve Datadog observability and PagerDuty alerting configurations, championing observability standards across metrics, logging, tracing, and alerting.

  • Design infrastructure controls that meet PCI-DSS, SOC 1/2, and ISO 27001 compliance requirements, and support evidence collection during audits.

  • Partner with Engineering, Product, Security, and Operations teams to improve CI/CD pipelines, release processes, and overall DevOps maturity.

  • Mentor peers and junior engineers, and influence organizational engineering standards as the internal technical authority on infrastructure design.

Your skills and experiences might also include:

Required

  • 8+ years in Site Reliability Engineering, Platform Engineering, or Infrastructure Engineering, with direct ownership of complex production platforms at scale.

  • Proven track record as a senior technical escalation point for cross-team incidents requiring architectural-level decision-making and permanent remediation.

  • Expert-level experience designing and operating cloud infrastructure in Microsoft Azure, with working knowledge of AWS.

  • Deep expertise running Kubernetes in production — cluster design, workload operations, and platform-level maintenance.

  • Advanced Terraform and Infrastructure-as-Code (IaC) skills, including module design, state management, and governance.

  • Strong Bash scripting and automation development for operational tooling and self-service platform capabilities.

  • Solid networking fundamentals: firewalls, DNS, routing, VPN, and network troubleshooting, including Cloudflare edge services.

  • Active Directory administration and hybrid identity experience.

  • Hands-on CI/CD pipeline design and deployment workflow improvement in a DevOps environment.

  • Datadog, PagerDuty, or equivalent observability and alerting platform experience.

  • Compliance-regulated environment experience, including PCI-DSS, SOC 1/2, and ISO 27001.

  • Strong ownership mentality with the ability to influence across organizational boundaries without direct authority.

Preferred

  • Puppet or equivalent configuration management platform administration.

  • Microsoft SQL Server environment administration.

  • Meraki firewall policy management.

  • Background with AI-driven operational workflows and Model Context Protocol (MCP) development.

  • Internal developer platform (IDP) or platform engineering initiative leadership.

  • Large-scale SaaS or high-availability platform support.

  • Scala and/or Java application ecosystem knowledge at the infrastructure level.

  • Azure Solutions Architect Expert, Azure Administrator Associate, AWS Solutions Architect, or CKA certification.

Applicants must reside within the Eastern or Central time zones to ensure alignment with our core business hours and effective collaboration with our team.

Authorization to work in the U.S. without employer sponsorship is required for this opportunity.

The anticipated base salary range for this position is $165,000 to $185,000 annually. Final compensation will be determined based on a variety of factors, including location, qualifications, experience, and skill set. Any compensation outside the stated range will be determined in accordance with applicable laws and company policy.

What are we all about?

The Access Group is one of the largest UK-headquartered business management software providers. It provides solutions that empower more than 160,000 small and mid-sized organisations in commercial and non-profit sectors across Europe, USA and APAC, giving every employee the freedom to do more of what's important. Its innovative cloud solutions and integrated AI software experience across multiple Access products transform how business technology is used.

With over 9,300 talented individuals driving innovation and customer excellence, we’re shaping the future of work. And we want you to be part of it. At Access, people are at the heart of everything we do. We’re committed to creating an inclusive, high-performing culture where everyone feels valued, respected, and empowered to thrive. If you’re excited about this role - even if your experience doesn’t tick every box - you might be exactly who we’re looking for.

We believe in equality for all and the transformative power of diversity. So why not join our vibrant team, where you can love what you do, love how you live, and most importantly, be authentically you?

Let’s make a difference together.

Love Work. Love Life. Be You.

Bereit, sich bei The Access Group zu bewerben?
Bei The Access Group bewerben

Wie sich dieses Gehalt für SRE vergleicht

Diese Stelle zahlt $175,000/yrim Einklang mit der üblichen Spanne für SRE Stellen.

$161,899 dem Median $191,000 $245,233

Übliche Spanne $174,525–$218,360/yr, aus 38 vergleichbaren SRE Anzeigen auf JobsRadar (Vergütung auf USD hochgerechnet). Gehaltseinblicke für SRE ansehen →

Über The Access Group

Hello! We are thrilled you are considering Access as your next career move. Why? Because At Access, people are at the heart of everything we do. Our people are game-changers, life-lovers, culture-builders, and free-thinkers. We strive to create an environment where everyone feels valued, respected, and empowered to contribute their best. We believe in hiring creative, curious, and caring future ready people who bring their unique diversity and a learn-it-all mindset to our global team. We believe in working hard as a collective team to win with our customers, give back to the communities we represent, and celebrate successes together. We can’t wait to meet you.

Alle Jobs bei The Access Group ansehen →

Ähnliche Jobs

Babylist
Staff Engineer, Site Reliability
Babylist
⚡ Früh bewerben United States Vor Ort $226,673–$271,991
● Neu 👁 Gesehen ✓ Beworben vor 4 Tg.
Teleport
Site Reliability Engineer (Forward Deployed)
Teleport
⚡ Früh bewerben United States (Remote) · standortgebunden $189,000–$326,000
● Neu 👁 Gesehen ✓ Beworben vor 4 Tg.
Teleport
Senior Site Reliability Engineer - US
Teleport
⚡ Früh bewerben United States (Remote) · standortgebunden $222,000–$326,000
● Neu 👁 Gesehen ✓ Beworben vor 4 Tg.
Astronomer
Customer Reliability Engineer (Airflow)
Astronomer
⚡ Früh bewerben Remote (United States) · standortgebunden $125,000–$130,000
● Neu 👁 Gesehen ✓ Beworben vor 5 Tg.
IV
Site Reliability Engineer (Tu-Sat, night shift)
Ivanti
⚡ Früh bewerben United States, Remote · standortgebunden
● Neu 👁 Gesehen ✓ Beworben vor 5 Tg.
GoDaddy
Lead Site Reliability Engineer - Ceph Storage
GoDaddy
⚡ Früh bewerben United States Hybrid $154,000–$231,000
● Neu 👁 Gesehen ✓ Beworben vor 5 Tg.
Reddit
Senior Site Reliability Engineer, Ads
Reddit
⚡ Früh bewerben San Francisco, CA Vor Ort $190,800–$267,100
● Neu 👁 Gesehen ✓ Beworben vor 6 Tg.
Reddit
Staff Site Reliability Engineer, Ads
Reddit
⚡ Früh bewerben San Francisco, CA Vor Ort $217,000–$303,900
● Neu 👁 Gesehen ✓ Beworben vor 6 Tg.
Nebius
Site Reliability Engineer
Nebius
⚡ Früh bewerben Remote - United States · standortgebunden $130,000–$180,000
● Neu 👁 Gesehen ✓ Beworben vor 1 Wo.

Registrieren für Vorschläge, die auf die von Ihnen geöffneten Jobs und gespeicherten Suchen zugeschnitten sind.

Mehr Jobs bei The Access Group

Alle Jobs bei The Access Group ansehen →

Jetzt bewerben
🤖

Moment — langsam

JobsRadar wurde für echte Menschen gebaut, die eine schwere Zeit bei der Jobsuche haben — nicht für automatisierte Anfragen. Sie klicken viel zu schnell und sind jetzt vorübergehend blockiert.

Kommen Sie später wieder. Wenn Sie wirklich auf Jobsuche sind, stehen wir hinter Ihnen — verhalten Sie sich einfach wie ein Mensch.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Verschaffe dir einen Vorsprung bei der Jobsuche.

Tritt unserem Telegram-Kanal bei für das, was dir hilft, die Stelle zu bekommen — Gehaltsbenchmarks, den wöchentlichen Marktpuls und neue Feature-Drops. Kein Spam, nur Signal.

Dem Kanal beitreten — kostenlos