Jobs › Companies › Vultr › Senior Site Reliability Engineer, Databases

About this Senior Site Reliability Engineer, Databases role at Vultr

Vultr · Remote · Remote - United States

Who We Are

Vultr is on a mission to make high-performance cloud infrastructure easy to use, affordable, and locally accessible for enterprises and AI innovators around the world. With 33 global cloud data center locations, Vultr is trusted by hundreds of thousands of active customers across 185 countries for its flexible, scalable, global Cloud Compute, Cloud GPU, Bare Metal, and Cloud Storage solutions. In December 2024 Vultr announced an equity financing at a $3.5 billion valuation. Founded by David Aninowsky and self-funded for over a decade, Vultr has grown to become the world’s largest privately-held cloud infrastructure company.

Vultr Cares

  • 100% company-paid insurance premiums for employee medical, dental and vision plans.

  • 401(k) plan that matches 100% up to 4%, with immediate vesting

  • Professional Development Reimbursement of $2,500 each year

  • 11 Holidays + Paid Time Off Accrual + Rollover Plan

  • Commitment matters to Vultr! Increased PTO at 3 year and 10 year anniversary + 1 month paid sabbatical every 5 years + Anniversary Bonus each year

  • $500 stipend for remote office setup in first year + $400 each following year

  • Internet reimbursement up to $75 per month

  • Gym membership reimbursement up to $50 per month

  • Company paid Wellable subscription

Join Vultr

Vultr is seeking a highly skilled Senior Site Reliability Engineer, Databases to join our Platform Engineering team. You will be the reliability and operational backbone for Vultr's database infrastructure spanning MySQL InnoDB Clusters, PostgreSQL, and other database technologies. Working alongside our Senior Platform Engineer (Databases), you will own database monitoring, backup verification, disaster recovery testing, on-call incident response, replication health, access management, and compliance remediation. This is a peer-level role where you will apply SRE methodology — error budgets, runbooks, automation-first thinking, and toil reduction — specifically to database systems that power a global cloud platform serving millions of customers. You bring deep operational database expertise to complement our existing schema and DDL strengths, and you are comfortable writing PHP, Python, or Go to automate and instrument everything you build.



Key Responsibilities

  • Own and evolve comprehensive monitoring and alerting for MySQL InnoDB Clusters, and PostgreSQL infrastructure across multiple datacenters

  • Establish and execute a quarterly backup verification and disaster recovery testing program across all database systems

  • Serve as first-tier on-call responder for database incidents, executing documented runbooks and escalating to senior engineering when architecture-level decisions are required

  • Monitor and maintain replication health across databases

  • Own database user lifecycle management — provisioning, deprovisioning, access audits, and role-based access control across all database systems

  • Execute and track security compliance remediation including pen-test findings, GRC audit requirements, and encryption-at-rest verification

  • Manage operational health of the Debezium/Kafka Connect data pipeline in coordination with the Kafka infrastructure team

  • Build and maintain Puppet profiles for database infrastructure configuration management and write PHP, Python, or Go automation tooling to reduce operational toil

  • Develop and maintain runbooks, operational documentation, and disaster recovery procedures for all database systems

  • Partner with the Senior Platform Engineer (Databases) as a peer — reviewing each other's work, sharing on-call, and splitting ownership of database reliability across production systems

Qualifications

  • 7+ years of experience in Site Reliability Engineering, DevOps, or Database Operations roles in production environments at scale

  • Deep operational expertise with MySQL in production — InnoDB Cluster, Group Replication, MySQL Router, and ProxySQL — with strong troubleshooting skills for replication, performance, and reliability issues

  • Production experience with PostgreSQL — replication, high availability, performance tuning, and operational management

  • Strong proficiency with configuration management tools (Puppet preferred) and infrastructure-as-code practices

  • Experience with database backup tools (xtrabackup, pg_dump, mysqldump) and disaster recovery procedures

  • Proficiency in PHP and Python or Go for automation, tooling, and integration with existing codebases

  • Experience with database observability — Prometheus exporters, Grafana dashboards, alerting frameworks, and SLO/error-budget methodology

  • Familiarity with Kafka Connect, Debezium, or similar change-data-capture pipelines

  • Strong incident response skills with experience in on-call rotations, including post-incident review and remediation

  • Excellent communication skills and ability to collaborate across engineering teams as a senior peer

Compensation

$125,000 - $135,000

We are currently accepting applications from candidates residing in the following states: Alabama, Arizona, Colorado, Connecticut, Florida, Georgia, Idaho, Illinois, Indiana, Iowa, Kentucky, Louisiana, Maryland, Massachusetts, Michigan, Minnesota, Missouri, Montana, Nebraska, Nevada, New Jersey, New Mexico, New York, North Carolina, Ohio, Oklahoma, Pennsylvania, Rhode Island, South Carolina, Tennessee, Texas, Utah, Vermont, Virginia, Wisconsin.

Inclusion & Privacy

We are an equal opportunity employer and are committed to creating an inclusive environment for all employees. We welcome applications from individuals of all backgrounds and experiences, and we prohibit discrimination based on race, color, religion, sex, sexual orientation, gender identity, national origin, age, disability, veteran status, or any other protected status under applicable laws. Vultr will consider qualified applicants with arrest or conviction records in accordance with applicable laws and will not conduct a background check until after an offer of employment has been extended and accepted.

We also take your privacy seriously. We handle personal information responsibly and follow applicable laws, including U.S. privacy rules and India’s Digital Personal Data Protection Act, 2023. Your data is used only for legitimate business purposes and is protected with proper security measures.

Where allowed by law, applicants may request details about the data we collect, access or delete their information, withdraw consent for its use, and opt out of nonessential communications. For more details, please see our Privacy Policy.

Ready to apply to Vultr?
Apply to Vultr

How this SRE salary compares

This role pays $130,000/yr — in line with the typical range for SRE roles.

$94,040 median $157,295 $230,000

Typical range $119,000–$195,857/yr, from 825 comparable SRE listings on JobsRadar (pay annualized to USD). See SRE salary insights →

Similar jobs

Cloudlinux
Senior Database Reliability Engineer (DBRE) (remote work)
Cloudlinux
⚡ Apply early Warsaw, Masovian Voivodeship,... · location restricted
● New 👁 Seen ✓ Applied 2h ago
Rowan Digital Infrastructure
Reliability Engineer - Commissioning
Rowan Digital Infrastructure
⚡ Apply early Temple, TX Hybrid $130,000–$150,000
● New 👁 Seen ✓ Applied 3h ago
Camunda
Senior Site Reliability Engineer
Camunda
⚡ Apply early Remote · location restricted
● New 👁 Seen ✓ Applied 4h ago
Megaport
Senior Site Reliability Engineer
Megaport
⚡ Apply early Sao Paulo Hybrid
● New 👁 Seen ✓ Applied 4h ago
Vast
Thermal Fluids Design Reliability Engineer
Vast
⚡ Apply early Long Beach, California, United... Onsite $162,360–$265,392
● New 👁 Seen ✓ Applied 4h ago
Vast
Reliability Engineer
Vast
⚡ Apply early Long Beach, California, United... Onsite $112,340–$159,468
● New 👁 Seen ✓ Applied 4h ago
Vast
Structures Design Reliability Engineer
Vast
⚡ Apply early Long Beach, California, United... Onsite $162,360–$265,392
● New 👁 Seen ✓ Applied 4h ago
Rent the Runway
Site Reliability Engineer
Rent the Runway
⚡ Apply early Galway, Ireland Hybrid
● New 👁 Seen ✓ Applied 5h ago
NL
Senior Cloud Platform & Site Reliability Engineering Lead
National Life Insurance Company
⚡ Apply early Addison, TX; Montpelier, VT Onsite $136,875–$200,750
● New 👁 Seen ✓ Applied 6h ago

Sign up for suggestions tailored to the jobs you open and the searches you save.

More jobs at Vultr

See all jobs at Vultr →

Apply now
🤖

Whoa — hold up

JobsRadar was built for real people having a rough time in their job search — not for automated requests. You're clicking way too fast and you're now temporarily blocked.

Come back later. If you're genuinely job hunting, we've got your back — just act like a human.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Get an edge on your job hunt.

Join our Telegram channel for the stuff that helps you land the role — salary benchmarks, the weekly market pulse, and new-feature drops. No spam, just signal.

Join the channel — it's free