×
Register Here to Apply for Jobs or Post Jobs. X

Senior Site Reliability Engineer British Columbia

Job in Burnaby, BC, Canada
Listing for: 2K Games, Inc.
Full Time position
Listed on 2026-09-03
Job specializations:
  • IT/Tech
    SRE/Site Reliability, Cloud Computing: Infrastructure & Operations
Salary/Wage Range or Industry Benchmark: 94000 - 139000 CAD Yearly CAD 94000.00 139000.00 YEAR
Job Description & How to Apply Below
Position: Senior Site Reliability Engineer New British Columbia

At 2K, we create some of the most iconic and culture-shaping video games in entertainment, including NBA® 2K, one of the top-selling franchises in the world, and legendary titles like Bio Shock®, Borderlands®, Mafia, Sid Meier’s Civilization®, and XCOM®, as well as fan favorites WWE® 2K, Top Spin®, and PGA TOUR® 2K. We build unforgettable experiences by pushing the boundaries of creativity, authenticity and innovation across every genre.

Our portfolio is brought to life by some of the most influential game development studios in the world. Visual Concepts, Firaxis Games, Hangar 13, Cat Daddy Games, 31st Union, Cloud Chamber, Gearbox, HB Studios, and 2K Sports Lab create world-class experiences across platforms.

But what truly powers 2K is our people.

We believe the best ideas come from teams that feel empowered, supported, and inspired. As an equal opportunity employer, we are committed to fostering a diverse, inclusive workplace where people are encouraged to come as they are and do their best work.

What We Need

The 2K SRE team owns the infrastructure behind every player connection — All 2K game services, account platforms, CI/CD pipelines, and developer tooling spanning AWS, GCP, and on-premises data centers across multiple global regions. Global launch windows and live-service events push systems to their limits, and this team is expected to hold the line.

Post-mortems here focus on systems, not people. Automation is the default answer to repetitive work. The infrastructure keeps millions of players connected — and the team takes that seriously!

The Senior SRE at 2K is a hands‑on technical leader — shaping production infrastructure across multiple clouds and regions while partnering with network engineers, systems architects, and game studio developers. This is an ownership role: driving technical direction, influencing reliability from architecture review through production operation, and closing the gap between what engineering ships and what players experience.

What You'll Do

Platform & Infrastructure

Design, build, and operate scalable multi‑cloud and hybrid infrastructure using Terraform, Pulumi, and Git Ops workflows (ArgoCD, Flux). Own Kubernetes platforms (EKS, GKE) end‑to‑end — cluster lifecycle, multi‑tenancy, networking (Istio, Cilium), and autoscaling — and push progressive delivery patterns (blue/green, canary) across game service deployments.

Observability & Reliability

Build and run the full observability stack:
Prometheus + Grafana + Datadog

Define SLI/SLO/error budget policies and build alerting that cuts through the noise

Lead chaos engineering exercises to surface failure modes before players encounter them

Drive incident response and post‑mortems with a focus on systemic fixes and real follow‑through

Automation, Security & Developer Experience

Eliminate toil through self‑service provisioning, automated remediation, and intelligent scaling. Harden CI/CD pipelines (Git Hub Actions, Jenkins, ArgoCD). Embed security at the platform layer through secrets management (Password State, 1

Password, and AWS Secrets Manager), policy‑as‑code (OPA/Gatekeeper).

Leadership

Promote SRE practices across 2K studios through reliability reviews, runbooks, and embedded collaboration

Shape architectural decisions and author engineering RFCs that move the platform forward

What Will Make You A Great Fit

Required Qualifications

5+ years in SRE, Platform Engineering, or equivalent infrastructure work at production scale

Deep Kubernetes experience in cloud environments (EKS or GKE preferred) — networking, storage, multi‑cluster patterns

Strong IaC proficiency with Terraform and/or Pulumi; hands‑on with Helm, Terragrunt, and Git Ops tooling (ArgoCD or Git Hub Actions)

Modern and Legacy Tech: AWS, GCP, VMware, and Bare metal servers

Server Configuration using Ansible, Puppet, and AWS Systems Manager

Observability stack experience:
Datadog, Prometheus + Grafana, and Open Telemetry,

SLI/SLO/error budget fluency — including how to operationalize them inside engineering teams

Production‑quality code in Go, Python, or Type Script: tools, automation, and internal libraries

Linux internals, TCP/IP networking, DNS, and TLS — proven…

Position Requirements
10+ Years work experience
Note that applications are not being accepted from your jurisdiction for this job currently via this jobsite. Candidate preferences are the decision of the Employer or Recruiting Agent, and are controlled by them alone.
To Search, View & Apply for jobs on this site that accept applications from your location or country, tap here to make a Search:
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary