Jobs › Companies › 2K › Senior Site Reliability Engineer

À propos de ce poste Senior Site Reliability Engineer chez 2K

2K · Sur site · Burnaby, British Columbia, Canada

At 2K, we create some of the most iconic and culture-shaping video games in entertainment, including NBA® 2K, one of the top-selling franchises in the world, and legendary titles like BioShock®, Borderlands®, Mafia, Sid Meier’s Civilization®, and XCOM®, as well as fan favorites WWE® 2K, TopSpin®, and PGA TOUR® 2K. We build unforgettable experiences by pushing the boundaries of creativity, authenticity and innovation across every genre.

Our portfolio is brought to life by some of the most influential game development studios in the world. Visual Concepts, Firaxis Games, Hangar 13, Cat Daddy Games, 31st Union, Cloud Chamber, Gearbox, HB Studios, and 2K SportsLab create world-class experiences across platforms.

But what truly powers 2K is our people.

We believe the best ideas come from teams that feel empowered, supported, and inspired. As an equal opportunity employer, we are committed to fostering a diverse, inclusive workplace where people are encouraged to come as they are and do their best work.

What We Need

The 2K SRE team owns the infrastructure behind every player connection — All 2K game services, account platforms, CI/CD pipelines, and developer tooling spanning AWS, GCP, and on-premises data centers across multiple global regions. Global launch windows and live-service events push systems to their limits, and this team is expected to hold the line.

Post-mortems here focus on systems, not people. Automation is the default answer to repetitive work. The infrastructure keeps millions of players connected — and the team takes that seriously!

The Senior SRE at 2K is a hands-on technical leader — shaping production infrastructure across multiple clouds and regions while partnering with network engineers, systems architects, and game studio developers. This is an ownership role: driving technical direction, influencing reliability from architecture review through production operation, and closing the gap between what engineering ships and what players experience.

What You'll Do

Platform & Infrastructure

Design, build, and operate scalable multi-cloud and hybrid infrastructure using Terraform, Pulumi, and GitOps workflows (ArgoCD, Flux). Own Kubernetes platforms (EKS, GKE) end-to-end — cluster lifecycle, multi-tenancy, networking (Istio, Cilium), and autoscaling — and push progressive delivery patterns (blue/green, canary) across game service deployments.

Observability & Reliability

  • Build and run the full observability stack: Prometheus + Grafana + Datadog

  • Define SLI/SLO/error budget policies and build alerting that cuts through the noise

  • Lead chaos engineering exercises to surface failure modes before players encounter them

  • Drive incident response and post-mortems with a focus on systemic fixes and real follow-through

Automation, Security & Developer Experience

Eliminate toil through self-service provisioning, automated remediation, and intelligent scaling. Harden CI/CD pipelines (GitHub Actions, Jenkins, ArgoCD) . Embed security at the platform layer through secrets management (PasswordState, 1Password, and AWS Secrets Manager), policy-as-code (OPA/Gatekeeper).

Leadership

  • Promote SRE practices across 2K studios through reliability reviews, runbooks, and embedded collaboration

  • Shape architectural decisions and author engineering RFCs that move the platform forward

What Will Make You A Great Fit

Required Qualifications

  • 5+ years in SRE, Platform Engineering, or equivalent infrastructure work at production scale

  • Deep Kubernetes experience in cloud environments (EKS or GKE preferred) — networking, storage, multi-cluster patterns

  • Strong IaC proficiency with Terraform and/or Pulumi; hands-on with Helm, Terragrunt, and GitOps tooling (ArgoCD or GitHub Actions)

  • Modern and Legacy Tech: AWS, GCP, VMware, and Bare metal servers

  • Server Configuration using Ansible, Puppet, and AWS Systems Manager

  • Observability stack experience: Datadog, Prometheus + Grafana, and OpenTelemetry,

  • SLI/SLO/error budget fluency — including how to operationalize them inside engineering teams

  • Production-quality code in Go, Python, or TypeScript: tools, automation, and internal libraries

  • Linux internals, TCP/IP networking, DNS, and TLS — proven enough to debug at the system level

  • Incident response and post-mortem leadership with a track record of systemic follow-through

Preferred Qualifications

  • Live-service game or large-scale consumer internet experience at millions of concurrent users

  • Service mesh depth (Istio, Cilium) and advanced Kubernetes networking

  • FinOps and managing resources at cloud scale

  • Experience with AI and Agentic Development

  • Cloud certifications (AWS Solutions Architect, GCP Professional Cloud Architect, CKA/CKS, or equivalent)

  • Experience mentoring SREs or leading reliability working groups

The pay range for this position in Burnaby, Canada at the start of employment is expected to be between $96,400 and $142,660 per Year. However, base pay offered is based on market location, and may vary further depending on individualized factors for job candidates, such as job-related knowledge, skills, experience, and other objective business considerations.

Subject to those same considerations, the total compensation package for employees in regular roles may also include other elements, including a bonus and/or equity awards, in addition to a full range of medical, financial, and/or other benefits, provided that temporary or intern roles will not be eligible for many of these payments or benefits. Details of participation in compensation and benefit plans (if applicable) will be provided if an employee receives an offer of employment. If hired the company reserves the right to modify base salary (as well as any other discretionary payment or compensation or benefit program) at any time, including for reasons related to individual performance, company or individual department/team performance, and market factors.

As an equal opportunity employer, we are committed to ensuring that qualified individuals with disabilities are provided reasonable accommodation to participate in the job application or interview process, to perform their essential job functions, and to receive other benefits and privileges of employment. Please contact us if you need reasonable accommodation.

Please note that 2K Games and its studios never uses instant messaging apps or personal email accounts to contact prospective employees or conduct interviews and when emailing, only use 2K.com accounts.

Please note that 2K Publishing is unable to provide visa sponsorship or assistance for this position. All candidates must be legally authorized to work in Canada without requiring current or future employer sponsorship.

 

 

#LI-Hybrid 

Prêt à postuler chez 2K ?
Postuler chez 2K

Comment se compare ce salaire pour SRE

Ce poste paie $119,530/yr — dans la fourchette habituelle pour les postes SRE.

$93,009 la médiane $155,600 $230,000

Fourchette typique $118,247–$196,125/yr, à partir de 826 annonces SRE comparables sur JobsRadar (rémunération annualisée en USD). Voir les aperçus de salaire pour SRE →

Emplois similaires

MeridianLink
Senior Site Reliability Engineer (AWS)
MeridianLink
⚡ Postuler tôt US Remote · lieu restreint $104,148–$162,800
● Nouveau 👁 Vu ✓ Postulé il y a 7 s
Pointclickcare
Senior Site Reliability Engineer, AI Infrastructure
Pointclickcare
⚡ Postuler tôt Mississauga, Ontario Hybride
● Nouveau 👁 Vu ✓ Postulé il y a 19 min
Palantir
Site Reliability Operations Analyst - UK Government
Palantir
⚡ Postuler tôt London, United Kingdom Hybride
● Nouveau 👁 Vu ✓ Postulé il y a 19 min
Jump Trading
Site Reliability Engineer
Jump Trading
⚡ Postuler tôt Chicago, New York Sur site $200,000–$225,000
● Nouveau 👁 Vu ✓ Postulé il y a 1 h
PlayStation Global
Service Reliability Engineer
PlayStation Global
⚡ Postuler tôt Australia, Adelaide Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 1 h
2K
Senior Site Reliability Engineer
2K
⚡ Postuler tôt Bangalore, Karnataka, India Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 1 h
Roblox
Senior Site Reliability Engineer, Compute
Roblox
⚡ Postuler tôt San Mateo, CA, United States Sur site $243,290–$295,250
● Nouveau 👁 Vu ✓ Postulé il y a 1 h
Accenture Federal Services
Site Reliability Engineer (DevOps)
Accenture Federal Services
⚡ Postuler tôt Reston, VA Sur site $111,800–$221,800
● Nouveau 👁 Vu ✓ Postulé il y a 1 h
Accenture Federal Services
Senior Site Reliability Engineer
Accenture Federal Services
⚡ Postuler tôt Arlington, VA Sur site $106,300–$221,100
● Nouveau 👁 Vu ✓ Postulé il y a 1 h

Inscrivez-vous pour des suggestions adaptées aux emplois que vous ouvrez et aux recherches que vous enregistrez.

Plus d’emplois chez 2K

Voir tous les emplois chez 2K →

Postuler maintenant
🤖

Doucement — un instant

JobsRadar a été conçu pour de vraies personnes qui traversent une période difficile dans leur recherche d’emploi — pas pour des requêtes automatisées. Vous cliquez beaucoup trop vite et vous êtes maintenant temporairement bloqué.

Revenez plus tard. Si vous cherchez réellement un emploi, nous sommes de votre côté — agissez simplement comme un être humain.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Prenez une longueur d’avance dans votre recherche d’emploi.

Rejoignez notre canal Telegram pour ce qui vous aide à décrocher le poste — références salariales, le pouls hebdomadaire du marché et les annonces de nouveautés. Pas de spam, que du signal.

Rejoindre le canal — c’est gratuit