Jobs Companies Filevine Sr. Site Reliability Engineer

Über diese Sr. Site Reliability Engineer Stelle bei Filevine

Filevine · Remote · United States
Filevine is a Legal AI company delivering Legal Operating Intelligence for the future of legal work. Grounded in a singular system of truth, Filevine brings together data, documents, workflows, and teams into one unified platform—where modern legal work happens with clarity and consistency.
 
Powered by LOIS, the Legal Operating Intelligence System, Filevine connects context across every matter to transform legal operations from reactive to proactive. LOIS reads, understands, and reasons across your data to surface insight, automate complexity, and give professionals the clarity and confidence to see more, know more, and do more. Fueled by a team of exceptional collaborators and innovators, Filevine’s rapid growth has earned AI awards and recognition from Deloitte and Inc. as one of the most innovative and fastest-growing technology companies in the country.

Role Summary:
 
As a Senior Site Reliability Engineer at Filevine, you will be a well-rounded reliability leader
embedded with a cross-functional team — and a primary driver of Observability excellence. You
bring broad, deep SRE expertise across the full software development lifecycle, and you bring
mastery in making complex systems visible, understandable, and actionable through best-in-class
monitoring, alerting, incident management, and platform tooling.

You will establish and own the observability posture for your team’s systems, serve as an
authoritative voice for reliability, and invest meaningfully in the engineers around you. As you ramp
up and gain context in the Filevine environment, you’ll grow into increasingly impactful work —
taking on mission-critical objectives that build out and improve our autonomous, observable
systems while cementing your reputation as an exceptional engineer who designs and maintains
systems that perform reliably at enterprise scale.

Responsibilities

  • Own and continuously advance the observability strategy for your team — including
    monitoring, alerting, dashboards, distributed tracing, log aggregation, and SLI/SLO/SLA
    frameworks.
  • Lead incident management end-to-end: detection, triage, communication, resolution, and
    blameless post-mortems that drive lasting improvements.
  • Serve as a Reliability Engineering leader on your team — providing strong technical
    leadership, sound judgment, and a clear voice on reliability across the SDLC.
  • Design and maintain autonomous systems for building, deploying, testing, and operating all
    Filevine products with minimal human intervention.
  • Continuously enhance CI/CD pipelines, automation scripts, playbooks, and tooling to reduce
    toil and accelerate resolution time.
  • Proactively identify and resolve gaps in system availability, performance, and security while
    defending overall security posture.
  • Mentor engineers — dedicating meaningful time to growing those around you, sharing
    knowledge proactively, and leaving people more capable than before.
  • Document processes, architecture, procedures, and best practices; take full ownership of
    documentation for the technologies in your domain and actively close gaps for fellow SREs.
  • Participate in 24/7 on-call rotation for production support and emergency response;
    communicate clearly with technical and management stakeholders at all levels.
  • Build roadmaps for the technologies and workflows you own, giving the team a clear direction
    for ongoing improvement.
  • Observability & Incident Management

    This role places particular emphasis on Observability and Incident Management. The ideal
    candidate will bring hands-on, expert-level experience in these areas:
     
    • Deep, practical experience with New Relic — or a comparable enterprise observability
    platform (e.g., Datadog, Dynatrace, Grafana/Prometheus) — including dashboarding, alerting,
    APM, infrastructure monitoring, and log management.
    • Proven ability to design and implement comprehensive observability strategies: defining
    meaningful SLIs and SLOs, building actionable alert hierarchies, and instrumenting services
    for full-stack visibility.
    • Strong incident management discipline: structured on-call practices, runbooks, escalation
    paths, stakeholder communication under pressure, and rigorous post-incident review.
    • Experience integrating observability tooling into CI/CD pipelines to surface reliability signals
    earlier in the development lifecycle.
    • Ability to translate complex system behavior into clear, consumable signals for both technical
    teams and non-technical stakeholders

    Qualifications

  • 8+ years of hands-on technical experience in software engineering, infrastructure, or
    operations roles, including a minimum of 5 years dedicated to Site Reliability Engineering.• Expert-level, well-rounded SRE skill set — proficient across monitoring/alerting, incident
    response, capacity planning, performance optimization, CI/CD, and reliability engineering best
    practices.
    • Deep hands-on expertise with New Relic or a comparable observability platform; strong
    preference for candidates who have led observability platform adoption or migration at scale.
    • Demonstrated experience owning incident management programs: on-call processes,
    escalation design, post-mortem culture, and measurable MTTR/MTTD improvement.
    • Strong proficiency in Python, Bash, PowerShell, and other common SRE scripting and
    automation technologies.
    • Expert-level experience designing, building, and maintaining autonomous systems that handle
    software build, deployment, testing, monitoring, and operations.
    • Proficient hands-on experience with AWS (EC2, EKS/Kubernetes, CloudWatch, Lambda, S3,
    IAM) and the broader cloud-native ecosystem.
    • Strong communicator who proactively informs stakeholders, operates transparently, and can
    bridge technical complexity for product and management audiences.
    • Proven track record of mentoring engineers, leading initiatives to completion, and making
    those around them measurably better.
    • Bachelor's degree in Computer Science, Information Systems, or a related field; equivalent
    certifications (e.g., AWS certifications, Google Cloud Professional); or substantial comparable
    direct work experience

  • Cool Company Benefits:
    - A dynamic, rapidly growing company, focused on helping organizations thrive 
    - Medical, Dental, & Vision Insurance (for full-time employees)
    - Competitive & Fair Pay
    - Maternity & paternity leave (for full-time employees)
    - Short & long-term disability
    - Opportunity to learn from a dedicated leadership team
    - Top-of-the-line company swag
     
    Privacy Policy Notice
    Filevine will handle your personal information according to what’s outlined in our Privacy Policy.
     
    Communication about this opportunity, or any open role at Filevine, will only come from representatives with email addresses using "filevine.com". Other addresses reaching out are not affiliated with Filevine and should not be responded to.
     
    Bereit, sich bei Filevine zu bewerben?
    Bei Filevine bewerben

    Ähnliche Jobs

    Verisign
    Site Reliability Engineer - IBM AIX
    Verisign
    ⚡ Früh bewerben Reston,Virginia,United States Hybrid $135,800–$183,800
    ● Neu 👁 Gesehen ✓ Beworben vor 5 Std.
    Verisign
    SRE - Linux
    Verisign
    ⚡ Früh bewerben Reston,Virginia,United States Hybrid $135,800–$183,800
    ● Neu 👁 Gesehen ✓ Beworben vor 5 Std.
    Verisign
    Site Reliability Engineer
    Verisign
    ⚡ Früh bewerben Reston,Virginia,United States Hybrid $135,800–$183,800
    ● Neu 👁 Gesehen ✓ Beworben vor 5 Std.
    Viant Technology
    Cloud Reliability Engineer (Irvine/Los Angeles Interview Required)
    Viant Technology
    ⚡ Früh bewerben Irvine, California, United Sta... Vor Ort $130,000–$150,000
    ● Neu 👁 Gesehen ✓ Beworben vor 7 Std.
    Viant Technology
    Staff Cloud Reliability Engineer
    Viant Technology
    ⚡ Früh bewerben Irvine, California, United Sta... Vor Ort $180,000–$200,000
    ● Neu 👁 Gesehen ✓ Beworben vor 7 Std.
    Viant Technology
    Senior Cloud Reliability Engineer
    Viant Technology
    ⚡ Früh bewerben Irvine, California, United Sta... Vor Ort $150,000–$180,000
    ● Neu 👁 Gesehen ✓ Beworben vor 7 Std.
    ClickHouse
    Senior Site Reliability Engineer- Remote
    ClickHouse
    ⚡ Früh bewerben Germany Vor Ort
    ● Neu 👁 Gesehen ✓ Beworben vor 8 Std.
    Roblox
    Senior Machine Learning Engineer, Reliability
    Roblox
    ⚡ Früh bewerben San Mateo, CA, United States Vor Ort $196,750–$243,290
    ● Neu 👁 Gesehen ✓ Beworben vor 8 Std.
    Roblox
    Senior Site Reliability Engineer, Compute
    Roblox
    ⚡ Früh bewerben San Mateo, CA, United States Vor Ort $243,290–$295,250
    ● Neu 👁 Gesehen ✓ Beworben vor 8 Std.

    Registrieren für Vorschläge, die auf die von Ihnen geöffneten Jobs und gespeicherten Suchen zugeschnitten sind.

    Mehr Jobs bei Filevine

    Alle Jobs bei Filevine ansehen →

    Jetzt bewerben
    🤖

    Moment — langsam

    JobsRadar wurde für echte Menschen gebaut, die eine schwere Zeit bei der Jobsuche haben — nicht für automatisierte Anfragen. Sie klicken viel zu schnell und sind jetzt vorübergehend blockiert.

    Kommen Sie später wieder. Wenn Sie wirklich auf Jobsuche sind, stehen wir hinter Ihnen — verhalten Sie sich einfach wie ein Mensch.

    Catch your next role the second it’s posted.

    Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

    Create free account

    Free forever · takes 30 seconds · already have one?

    Verschaffe dir einen Vorsprung bei der Jobsuche.

    Tritt unserem Telegram-Kanal bei für das, was dir hilft, die Stelle zu bekommen — Gehaltsbenchmarks, den wöchentlichen Marktpuls und neue Feature-Drops. Kein Spam, nur Signal.

    Dem Kanal beitreten — kostenlos