Jobs Companies Capital Senior SRE Engineer (Observability Focus)

Sobre este puesto de Senior SRE Engineer (Observability Focus) en Capital

Capital · Remoto · Warsaw, Mazowieckie, Poland

We are a leading trading platform that is ambitiously expanding to the four corners of the globe. Our top-rated products have won prestigious industry awards for their cutting-edge technology and seamless client experience. We deliver only the best, so we are always in search of the best people to join our ever-growing talented team.

We're building out our observability practice and need a senior engineer who can own it end to end. This is a hands-on role. You'll design and operate the telemetry stack that gives our engineering teams real visibility into production — across a hybrid AWS and on-premise environment, at scale.

Responsibilities:

  • Own the full observability stack: metrics (VictoriaMetrics), logs (OpenSearch), and traces (OpenTelemetry) — from pipeline design to day-2 operations.
  • Architect and run VictoriaMetrics cluster topology (vmstorage/vminsert/vmselect), including vmagent scraping, remote write configuration, vmalert rules, and cardinality control.
  • Operate OpenSearch clusters: index lifecycle management (ISM), hot-warm-cold architecture, shard tuning, and ingest pipelines via Data Prepper.
  • Build and maintain OTEL Collector pipelines — receivers, processors, exporters — and instrument services across Java, Python, and JS/TS stacks (auto and manual).
  • Run Kafka as the telemetry transport layer (OTEL Collector → Kafka → backends), including topic design, partition strategy, consumer group lag monitoring, and throughput tuning for high-volume telemetry.
  • Manage log shipping infrastructure using Fluent Bit, Vector, or Fluentd; define structured logging standards and field normalization across services.
  • Build Grafana dashboards and alerting that engineers actually use — clear, actionable, with well-structured variables and thresholds.
  • Work with platform and application teams to improve sampling strategies (head/tail), batching, and context propagation across distributed services.
  • Contribute to incident response, post-mortems, and reliability improvements driven by observability signals.
  • Mentor engineers on observability practices, tooling, and structured logging standards.
  • Requirements:

  • 6+ years in a DevOps, SRE, or platform engineering role, with at least 2 years focused on observability tooling at production scale.
  • Deep hands-on experience with VictoriaMetrics (or Prometheus) — MetricsQL/PromQL, exporters, service discovery, remote write, downsampling, and retention management.
  • Solid OpenSearch or Elasticsearch skills: cluster operations, Query DSL, ISM policies, and ingest pipeline design.
  • Production experience with OpenTelemetry: Collector configuration, OTLP, context propagation, and instrumentation across multiple languages.
  • Strong Kafka skills — producer/consumer patterns, consumer group management, Kafka Connect, Schema Registry, and JMX-based monitoring. Strimzi experience a plus if you've run Kafka on Kubernetes.
  • Proficiency with log shippers (Fluent Bit, Vector, Fluentd) and structured log parsing/normalization.
  • Working knowledge of Kubernetes (operators, Helm), Argo CD/GitOps, and Terraform/Ansible.
  • Comfortable in a hybrid AWS + on-prem environment; solid understanding of networking as it applies to scraping and shipping pipelines.
  • Scripting ability in Bash or Python for automation and tooling.
  • Strong communication skills — you can explain observability tradeoffs clearly to engineers and non-engineers alike.
  • English proficiency.

  • What you will get in return:
     
    • Competitive Salary: We believe great work deserves great pay! Your skills and talents will be rewarded with a salary that makes you feel valued and motivated.
    • Work-Life Harmony: Join a company that genuinely cares about you - because your life outside of work matters just as much as your time on the clock. #LI-Hybrid
    • Generous Time Off: Need a breather? Our annual leave policy lets you recharge and enjoy life outside of work without a worry.
    • Employee Referral Program: Love working here? Share the love! Bring your talented friends on board and get rewarded for growing our awesome team.
    • Comprehensive Health & Pension Benefits: From medical insurance to pension plans, we’ve got your back. Plus, location-specific benefits and perks!
    • Workation Wonderland: Live your digital nomad dreams with 30 extra days to work remotely from anywhere in the world (some restrictions apply). Adventure awaits!
    • Volunteer Days: Make a difference! Take two additional paid days each year to support causes you care about and give back to the community.
     
    Be a key player at the forefront of the digital assets movement, propelling your career to new heights! Join a dynamic and rapidly expanding company that values and rewards talent, initiative, and creativity. Work alongside one of the most brilliant teams in the industry.
     
    Our company has an Internal Reporting Procedure. It is available from the Human Resources Department upon request hr@capital.com. You may report a violation referred to in the Procedure under the terms specified therein.
    ¿Listo para postularte en Capital?
    Postúlate en Capital

    Empleos similares

    Capital
    Senior Database Reliability Engineer (PostgreSQL, Terraform, AWS)
    Capital
    ⚡ Postúlate pronto Warsaw, Mazowieckie, Poland Remoto
    ● Nuevo 👁 Visto ✓ Postulado hace 2sem
    Capital
    Senior DevOps/SRE Engineer
    Capital
    ⚡ Postúlate pronto Warsaw, Mazowieckie, Poland Remoto
    ● Nuevo 👁 Visto ✓ Postulado hace 4sem
    Everpure
    Sr. IAM Site Reliability Engineer
    Everpure
    ⚡ Postúlate pronto Santa Clara, California Presencial $186,000–$279,000
    ● Nuevo 👁 Visto ✓ Postulado hace 1h
    Optiver Private Jobs
    Data - Site Reliability Engineer
    Optiver Private Jobs
    ⚡ Postúlate pronto Sydney, New South Wales, Austr... Presencial
    ● Nuevo 👁 Visto ✓ Postulado hace 3h
    Boomi
    Software Principal Engineer - SRE, Production Engineer​ing
    Boomi
    ⚡ Postúlate pronto India Presencial
    ● Nuevo 👁 Visto ✓ Postulado hace 3h
    Forward Networks
    Site Reliability Engineer
    Forward Networks
    ⚡ Postúlate pronto Santa Clara, CA Presencial $230,000–$250,000
    ● Nuevo 👁 Visto ✓ Postulado hace 4h
    Vast
    Thermal Fluids Design Reliability Engineer
    Vast
    ⚡ Postúlate pronto Long Beach, California, United... Presencial $162,360–$265,392
    ● Nuevo 👁 Visto ✓ Postulado hace 5h
    Vast
    Structures Design Reliability Engineer
    Vast
    ⚡ Postúlate pronto Long Beach, California, United... Presencial $162,360–$265,392
    ● Nuevo 👁 Visto ✓ Postulado hace 5h
    Vast
    Propulsion Design Reliability Engineer
    Vast
    ⚡ Postúlate pronto Long Beach, California, United... Presencial $162,360–$265,392
    ● Nuevo 👁 Visto ✓ Postulado hace 5h

    Regístrate para recibir sugerencias adaptadas a los empleos que abres y las búsquedas que guardas.

    Más empleos en Capital

    Ver todos los empleos en Capital →

    Postúlate ahora
    🤖

    Un momento — para

    JobsRadar se creó para personas reales que están pasando un mal momento en su búsqueda de empleo — no para solicitudes automatizadas. Estás haciendo clic demasiado rápido y ahora estás bloqueado temporalmente.

    Vuelve más tarde. Si de verdad estás buscando empleo, cuentas con nosotros — solo compórtate como una persona.

    Catch your next role the second it’s posted.

    Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

    Create free account

    Free forever · takes 30 seconds · already have one?

    Toma ventaja en tu búsqueda de empleo.

    Únete a nuestro canal de Telegram para lo que te ayuda a conseguir el puesto — referencias salariales, el pulso semanal del mercado y avisos de nuevas funciones. Sin spam, solo señal.

    Únete al canal — es gratis