Jobs Companies Cloudlinux Senior Database Reliability Engineer (DBRE) & Architect (worldwide remote)

Sobre este puesto de Senior Database Reliability Engineer (DBRE) & Architect (worldwide remote) en Cloudlinux

Cloudlinux · Warsaw, Masovian Voivodeship, Poland

CloudLinux is transforming the Linux infrastructure market by ensuring security and stability for over 500,000 servers worldwide. Our products - CloudLinux OS, TuxCare, and Imunify360 - are the de facto standard in the hosting industry and Enterprise segment.

We are seeking a visionary engineer to lead the evolution of our data platform. In 2025, we are shifting from classic database administration to an Internal Database-as-a-Service (DBaaS) model. We need a specialist who doesn’t just "configure backups," but designs resilient distributed systems, writes code to automate infrastructure, and transforms databases into a reliable service for product teams.

If you are tired of endless tickets and want to build platforms capable of processing petabytes of data, this role is for you.

Your Challenges & Responsibilities:

  • DBaaS Architecture: Design and implement a self-service platform based on Terraform and Ansible, enabling the deployment of HA clusters (PostgreSQL and ClickHouse, MongoDB, Redis) in a heterogeneous environment (Bare Metal + OpenNebula + Kubernetes + Public Clouds). You will turn infrastructure into a product.
  • Scaling ClickHouse: Manage exponentially growing analytics clusters (12+ clusters, tens of terabytes of data). You will tackle sharding, table engine optimization (ReplicatedMergeTree), and building reliable S3 backup pipelines under high load.
  • Data Platform & Analytics Support: Maintain and scale the infrastructure for Apache Airflow and Redash. You will ensure the reliability of ETL pipelines and visualization tools, bridging the gap between raw infrastructure and the data analytics team.
  • Reliability as Code: Implement SRE practices in data management. Replace manual incident response with automated self-healing mechanisms. Define and implement SLO/SLI for all databases.
  • Stack Modernization: Lead the migration process from legacy solutions to modern cloud patterns. Participate in decision-making regarding the implementation of Kubernetes operators for stateful workloads.
  • Expertise & Mentorship: Serve as the technical authority for product teams, helping them optimize data schemas and SQL queries for high-load systems.

Our Tech Stack:

  • Databases: PostgreSQL 15+ (Patroni, PgBouncer), ClickHouse (Sharded/Replicated), MongoDB, Redis, Kafka
  • Data & Analytics: Apache Airflow, Redash (Infrastructure & Integration).
  • Infrastructure: Own 3+DC colocation (OpenNebula, Kubernetes, Bare Metal), AWS, Google Cloud, Azure, DO – Hybrid Cloud.
  • Automation & IaC: Terraform, Ansible, Python/Go, GitLab, Jenkins, Gerrit.
  • Observability: VictoriaMetrics, Grafana, Loki.

Why CloudLinux?

  • Culture: A Remote-first company with an "Employees First" principle. We value results, not hours in the office.
  • Impact: Your architectural decisions will determine the stability of services used by thousands of companies around the world.
  • Growth: We support professional development and pay for training and conferences.

Requirements

What We Expect From You:

  • AI-Augmented Engineering: You don't view AI as a replacement for deep technical fundamentals, but as a high-leverage tool. We actively use AI agents (Claude, Codex, Gemini, etc.) to automate boilerplate, analyze complex logs, and speed up research. We expect you to be open to modern workflows and integrate AI into your day-to-day operations, allowing you to focus your brainpower on the true architectural challenges.
  • Deep PostgreSQL Expertise (5+ years): You know MVCC internals, understand locking mechanics, can configure Patroni and PgBouncer "with your eyes closed," and have experience with seamless major version upgrades under load.
  • ClickHouse Mastery: Experience operating large clusters, understanding ZooKeeper/ClickHouse Keeper, sharding, replication internals, and the ability to diagnose performance issues at the data-part level.
  • Engineering Mindset (SRE/DevOps): You hate doing the same task twice by hand. Experience writing complex Terraform modules and Ansible roles is mandatory. Programming skills in Python or Go for automation are a huge plus.
  • Hybrid Environment Experience: You understand the differences between running DBs on Bare Metal vs. Kubernetes vs. Cloud and know how to optimize TCO and disk subsystem performance (NVMe, Network Storage).
  • Systems Approach: You see the big picture - from the network packet to the application business logic. You understand the importance of security (FIPS, Audit logs) and Disaster Recovery.
  • English: upper-intermediate or higher - to ensure clear communication of progress within the teams

Nice to Have:

  • Experience building an Internal Developer Platform (IDP).
  • Experience operating databases in Kubernetes (CloudNativePG, Altinity Operator).
  • Experience working in Cloud and Hosting providers on similar services.

Benefits

What's in it for you?

  • A focus on professional development.
  • Interesting and challenging projects.
  • Fully remote work with flexible working hours, which allows you to schedule your day and work from any location worldwide.
  • Paid 24 days of vacation per year, 10 days of national holidays, and unlimited sick leaves.
  • Compensation for private medical insurance.
  • Co-working and gym/sports reimbursement.
  • Budget for education.
  • The opportunity to receive a reward for the most innovative idea that the company can patent.

By applying for this position, you agree with CloudLinux Privacy Policy and give us your consent to maintain and process your personal data with this respect. Please read our Privacy Policy for more information.

¿Listo para postularte en Cloudlinux?
Postúlate en Cloudlinux

Sobre Cloudlinux

CloudLinux is on a mission to make Linux secure, stable, and profitable. We have spent more than 500 combined years working on Linux, and are changing how hosting companies and data centers use this technology we love by bringing it to millions of their customers. With more than 500,000 product installations and 4,000 customers, including Liquid Web, 1&1, and Dell, CloudLinux combines in-depth technical knowledge of hosting, kernel development, and open source with unique client care expertise.

CloudLinux team members are not tied to a physical office location, and everyone works remotely full-time. We provide flexible working hours and an open management style, to avoid unnecessary bureaucracy and excessive control while getting the best from our employees. This system allows each of us to fully realize our ideas and ambitions, while comfortably combining work with our usual lifestyles.

Ver todos los empleos en Cloudlinux →

Empleos similares

Tenstorrent
Principal Debug and SRE Lead
Tenstorrent
⚡ Postúlate pronto Gdańsk, Pomeranian Voivodeship... Híbrido
● Nuevo 👁 Visto ✓ Postulado hace 1sem
Cloudlinux
Senior Database Reliability Engineer (DBRE) (worldwide remote)
Cloudlinux
⚡ Postúlate pronto Warsaw, Masovian Voivodeship,... Remoto
● Nuevo 👁 Visto ✓ Postulado hace 1sem
Altoros
(803) Senior Python Developer - L3 Support (SRE/Python + Unity-integration)
Altoros
⚡ Postúlate pronto Warsaw, Masovian Voivodeship,... Remoto
● Nuevo 👁 Visto ✓ Postulado hace 2sem
Plaud
SRE Engineer - San Francisco
Plaud
⚡ Postúlate pronto San Francisco, CA Híbrido
● Nuevo 👁 Visto ✓ Postulado hace 4h
Asana
Senior Software Engineer, Site Reliability
Asana
⚡ Postúlate pronto Warsaw Híbrido
● Nuevo 👁 Visto ✓ Postulado hace 6h
Asana
Senior Software Engineer, Platform Reliability
Asana
⚡ Postúlate pronto Warsaw Híbrido
● Nuevo 👁 Visto ✓ Postulado hace 6h
Clover Health
Senior Site Reliability Engineer
Clover Health
⚡ Postúlate pronto Remote - USA · restringido por ubicación $160,000–$208,000
● Nuevo 👁 Visto ✓ Postulado hace 6h
Roblox
Senior Machine Learning Engineer, Reliability
Roblox
⚡ Postúlate pronto San Mateo, CA, United States Presencial $196,750–$243,290
● Nuevo 👁 Visto ✓ Postulado hace 6h
Roblox
Senior Site Reliability Engineer, Compute
Roblox
⚡ Postúlate pronto San Mateo, CA, United States Presencial $243,290–$295,250
● Nuevo 👁 Visto ✓ Postulado hace 6h

Regístrate para recibir sugerencias adaptadas a los empleos que abres y las búsquedas que guardas.

Más empleos en Cloudlinux

Ver todos los empleos en Cloudlinux →

Postúlate ahora
🤖

Un momento — para

JobsRadar se creó para personas reales que están pasando un mal momento en su búsqueda de empleo — no para solicitudes automatizadas. Estás haciendo clic demasiado rápido y ahora estás bloqueado temporalmente.

Vuelve más tarde. Si de verdad estás buscando empleo, cuentas con nosotros — solo compórtate como una persona.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Toma ventaja en tu búsqueda de empleo.

Únete a nuestro canal de Telegram para lo que te ayuda a conseguir el puesto — referencias salariales, el pulso semanal del mercado y avisos de nuevas funciones. Sin spam, solo señal.

Únete al canal — es gratis