Jobs Companies CXM Direct LLC Application Site Reliability Engineer (SRE)

Über diese Application Site Reliability Engineer (SRE) Stelle bei CXM Direct LLC

CXM Direct LLC · Remote · Argentina

Join our Platform & Production Reliability team and help ensure the reliability, performance, and availability of our mission-critical trading systems. As an Application Site Reliability Engineer (SRE), you will own the day-to-day reliability of our .NET/C# services running on Windows, starting with our in-house liquidity bridge that connects MetaTrader trading servers to external liquidity providers. Over time, you will expand your impact across related trading and back-office services.

This is a hands-on role for an engineer who enjoys solving production challenges, improving observability, automating operations, and building resilient systems where uptime directly impacts customer experience.

Position Details

TeamPlatform & Production Reliability

LocationRemote (Americas, LatAm preferred)

Working HoursAmericas time zones (UTC-3 to UTC-8)

On-callRotation aligned with the London trading day

Employment TypeFull-time, Permanent

Experience LevelMid-Level (3–5 years)

Technology Stack.NET/C#, Windows Server, AWS, Aurora PostgreSQL, Prometheus, Grafana, Terraform

About the Role

Our trading platform powers every customer interaction, making reliability a first-class product concern. You will be responsible for maintaining and improving the operational reliability of our .NET/C# services on Windows, ensuring they remain highly available, observable, and resilient.

You'll collaborate closely with software engineers to improve monitoring, deployment safety, automation, fault isolation, and incident response, while driving continuous improvements in platform reliability and operational excellence.

What You'll Do

  • Participate in the on-call rotation for production trading systems and lead incident response during service disruptions.
  • Investigate production incidents, perform root cause analysis, and implement preventive actions to eliminate recurring issues.
  • Build and maintain Grafana dashboards, Prometheus alerts, and operational health views across applications, infrastructure, and databases.
  • Instrument .NET services to improve telemetry, metrics, logging, and visibility into service health and customer impact.
  • Define, implement, and monitor Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets.
  • Troubleshoot issues across:
    • .NET/C# applications
    • Windows Server
    • Aurora PostgreSQL databases
    • AWS infrastructure
    • CI/CD pipelines and deployments
  • Improve deployment safety, release automation, and rollback strategies.
  • Partner with developers to improve application operability, resilience, and fault isolation.
  • Automate operational tasks through scripting and infrastructure automation.
  • Create and maintain runbooks, operational documentation, and incident response procedures.
  • Continuously improve monitoring, alert quality, automation, and platform reliability.

Requirements

Required Technical Skills

.NET & Windows

  • Strong experience debugging and supporting .NET/C# applications in production.
  • Hands-on experience with Windows Server environments.

Scripting & Automation

  • Strong PowerShell scripting skills.
  • Experience with Python or Bash.

Observability

  • Experience with Grafana, Prometheus, and Loki (or equivalent monitoring and observability tools).
  • Solid understanding of metrics, logging, tracing, and alerting best practices.

CI/CD & DevOps

  • Experience with modern CI/CD pipelines.
  • Knowledge of deployment strategies, release automation, and rollback mechanisms.

Cloud & Infrastructure

  • Experience working with AWS.
  • Hands-on experience with Terraform or other Infrastructure as Code (IaC) tools.

Databases

  • Experience troubleshooting and supporting Aurora PostgreSQL or other relational database platforms.

Reliability Engineering

  • Practical experience with:
    • SLIs & SLOs
    • Error Budgets
    • Incident Response
    • Root Cause Analysis (RCA)
    • Alert Design
    • Production Operations

Preferred Qualifications

  • Experience supporting high-availability or low-latency financial or trading systems.
  • Familiarity with MetaTrader environments or financial technology platforms.
  • Experience with distributed systems and microservices.
  • Knowledge of OpenTelemetry or similar observability frameworks.
  • Exposure to Docker, Kubernetes, or containerized environments.

Benefits

Why Join Us?

  • Work on mission-critical trading infrastructure that directly impacts customers.
  • Solve challenging reliability and scalability problems in a real-time environment.
  • Build world-class observability, automation, and deployment practices.
  • Collaborate with experienced engineers in a modern engineering culture.
  • Influence reliability strategy and engineering best practices across the platform.

If you're passionate about production engineering, automation, and building reliable systems at scale, we'd love to hear from you.

Bereit, sich bei CXM Direct LLC zu bewerben?
Bei CXM Direct LLC bewerben

Über CXM Direct LLC

CXM Group is an international group of companies established in 2015 to provide reliable trading solutions within the global forex industry. Our team of experts has decades of experience across the Asia Pacific, US, and European markets, and has delivered multi-award-winning services for traders of all levels. With a focus on B2B and institutional customers, we have expanded to provide retail traders with cutting-edge technology tools and account types. Our offering includes SWAP-free trading for all clients, unlimited leverage, competitive spreads, deep liquidity pools, and strong relationships with tier-one banks and prime brokers. Our goal is to provide our customers with the optimal trading experience, regardless of their experience level.

Alle Jobs bei CXM Direct LLC ansehen →

Ähnliche Jobs

Registrieren für Vorschläge, die auf die von Ihnen geöffneten Jobs und gespeicherten Suchen zugeschnitten sind.

Mehr Jobs bei CXM Direct LLC

Alle Jobs bei CXM Direct LLC ansehen →

Jetzt bewerben
🤖

Moment — langsam

JobsRadar wurde für echte Menschen gebaut, die eine schwere Zeit bei der Jobsuche haben — nicht für automatisierte Anfragen. Sie klicken viel zu schnell und sind jetzt vorübergehend blockiert.

Kommen Sie später wieder. Wenn Sie wirklich auf Jobsuche sind, stehen wir hinter Ihnen — verhalten Sie sich einfach wie ein Mensch.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Verschaffe dir einen Vorsprung bei der Jobsuche.

Tritt unserem Telegram-Kanal bei für das, was dir hilft, die Stelle zu bekommen — Gehaltsbenchmarks, den wöchentlichen Marktpuls und neue Feature-Drops. Kein Spam, nur Signal.

Dem Kanal beitreten — kostenlos