Sobre este puesto de Staff Site Reliability Engineer en Assured
Assured is on a mission to modernize insurance. Claims processing (i.e. should we pay this claim?), while often overlooked, is the foundation of the entire industry. It’s currently highly manual, involving phone calls, faxes, and gut instinct, costing tens of billions of dollars a year. We can do better.
At Assured, we provide large insurers with the software solutions they need to win in a modern, technology-driven world. From self-service claim-filing software to backend fraud detection, we’re the engine that powers claims processing for some of the largest insurers in the world.
The challenges we face are deep and diverse, from creating digital experiences that provide comfort and clarity to claimants at their most stressed and vulnerable to orchestrating large-scale ML-driven decision-making on billions of dollars of claims payments, life at Assured is dynamic, collaborative, and rewarding.
As a Staff Site Reliability Engineer, you'll define the standards, patterns, and platforms that let every engineering team at Assured run their services reliably — and partner directly with the teams adopting them.
Some of the Problems You'll Solve:
🎯 Define what reliability means across a growing engineering organization.
Set the standards, patterns, and reference implementations teams adopt for SLOs, error budgets, instrumentation, alerting, and incident practice — and make them easy enough to adopt that teams actually do.
📊 Measure reliability the way customers experience it.
Move us beyond component-level availability targets to end-to-end SLOs for the claims journeys insurers and claimants depend on, spanning many services and teams, and extend that into how we measure and report SLA compliance.
🔭 Unify a fragmented observability picture.
Help drive our consolidation onto OpenTelemetry as a single instrumentation standard across shared services and product applications, so signal is consistent and comparable wherever it comes from.
🧱 Turn scattered reliability signal into decisions.
Build on and refine the reporting layer that pulls incident, alerting, and coverage data into one place, surfacing where risk actually lives across the platform and where we're flying blind.
🚨 Shorten the distance between an incident and a lasting improvement.
Improve how we detect, respond to, and learn from failure — incident tooling and automation, post-incident review practice, and making sure action items get closed rather than quietly aging out.
How You'll Make an Impact:
🤝 Help other teams run their own systems well.
Embed with product teams for a period at a time: specify what reliability looks like for their most critical paths, help them build it, then hand it over with them as the durable owner.
🔍 Find the risk before it finds us.
Surface coverage gaps, weak signals, and single points of failure across the platform, and make the case for fixing them before they become incidents.
📟 Support engineering when things go wrong.
Share an interrupt rotation with the rest of the SRE team, triaging reliability escalations and requests from across the organization.
🧑🏫 Raise the technical bar around you.
Mentor engineers across the organization through design review, written guidance, and hands-on collaboration on the problems they own.
⚡ Use AI to work faster and more effectively.
Use tools such as Claude, Codex, Cursor, and similar platforms to support tooling development, incident analysis, debugging, documentation, and operational work.
You'll Probably Thrive Here If You:
📐 Have deep site reliability and systems expertise.
You bring 10+ years of site reliability, production, or platform engineering experience, ideally within SaaS platforms or high-scale distributed systems environments.
🛠 Work across a modern reliability and observability stack.
You're comfortable with OpenTelemetry, metrics, traces and logs, AWS, Kubernetes, PostgreSQL, and modern incident tooling. Experience with every tool isn't required — we value strong fundamentals and the ability to learn quickly. Platform provisioning and cloud infrastructure sits with a separate Infrastructure team, and you'll work closely with a dedicated Database Reliability Engineering function.
📈 Know how to make SLOs stick.
You've designed and landed SLOs and error budgets that teams genuinely use to make decisions, rather than dashboards nobody opens.
💬 Lead through influence rather than ownership.
Our SRE function advises and enables rather than executing on other teams' behalf. You can bring a product team along with you, and redirect work that genuinely belongs elsewhere.
🚨 Stay effective when both systems and people are under pressure.
You've run incidents and post-incident reviews, and you're as comfortable coordinating people mid-incident as you are debugging the failure itself.
🔧 Build, rather than only configure.
You write real tooling and services, and you can reason about failure modes in systems you didn't build and debug across team boundaries.
🔄 Adapt quickly to new tools and technologies.
Your engineering judgment and ability to learn matter more than experience with a particular framework or platform. Great engineers learn new technologies. Great systems thinking is harder to teach.
Benefits:
🤑 Competitive Compensation: Competitive salary and equity packages for all employees
🏥 Healthcare Plan: Platinum medical, dental, and vision
🛡️ Free life insurance: Including long-term disability & short-term disability
🏄 Unlimited PTO: Uncapped vacation days & paid holidays
👶 Family Leave: Maternity & paternity
📈 401(k) Contribution: Assured contributes 3% of your income, even if you don't contribute
🏠 WFH Benefits: Lunch on us 2x/week, monthly phone stipend & other home office perks
👪 Health FSAs & HSAs: Pre-tax accounts for out-of-pocket medical expenses
🤝 Team events & Offsites: We're remote, but we regularly get together
**We have been made aware of individuals falsely posing as recruiters from Assured Insurance Technologies Inc. Please note that we only contact candidates from official @assured.claims email addresses and all interviews are conducted through verified company channels. If you are unsure whether a message is legitimate, please contact us directly at [email protected] before sharing any personal information**
Our Commitment:
We are an equal opportunity employer and value diversity at our company. We do not discriminate on the basis of race, religion, color, national origin, sex, gender, gender expression, sexual orientation, age, marital status, veteran status, or disability status. We will ensure that individuals with disabilities are provided reasonable accommodation to participate in the job application or interview process, perform essential job functions, and receive other benefits and privileges of employment. Please contact us to request accommodation.