Senior Site Reliability Engineer
2k · Austin, Texas, United States · 2026-07-01
About this role
<h3 data-path-to-node="2">Who We Are</h3> <p id="p-rc_ed02c07fb1fa8ac6-19" data-path-to-node="3"><span data-path-to-node="3,1"><span class="citation-39">At 2K, we create some of the most iconic and culture-shaping video games in entertainment, including NBA® 2K, one of the top-selling franchises in the world, and legendary titles like BioShock®, Borderlands®, Mafia, Sid Meier’s Civilization®, and XCOM®, as well as fan favorites WWE® 2K, TopSpin®, and PGA TOUR® 2K. </span></span><span data-path-to-node="3,4"><span class="citation-38">We build unforgettable experiences by pushing the boundaries of creativity, authenticity and innovation across every genre. </span></span></p> <p id="p-rc_ed02c07fb1fa8ac6-20" data-path-to-node="4"><span data-path-to-node="4,1"><span class="citation-37">Our portfolio is brought to life by some of the most influential game development studios in the world. </span></span><span data-path-to-node="4,4"><span class="citation-36">Visual Concepts, Firaxis Games, Hangar 13, Cat Daddy Games, 31st Union, Cloud Chamber, Gearbox, HB Studios, and 2K SportsLab create world-class experiences across platforms. </span></span><span data-path-to-node="4,7"><span class="citation-35">But what truly powers 2K is our people. </span></span><span data-path-to-node="4,10"><span class="citation-34">We believe the best ideas come from teams that feel empowered, supported, and inspired. </span></span><span data-path-to-node="4,13"><span class="citation-33">As an equal opportunity employer, we are committed to fostering a diverse, inclusive workplace where people are encouraged to come as they are and do their best work. </span></span></p> <h3 data-path-to-node="5">The Team</h3> <p data-path-to-node="6">The 2K SRE team owns the infrastructure behind every player connection—All 2K game services, account platforms, CI/CD pipelines, and developer tooling spanning AWS, GCP, and on-premises data centers across multiple global regions. Global launch windows and live-service events push systems to their limits, and this team is expected to hold the line.</p> <p data-path-to-node="7">Post-mortems here focus on systems, not people. Automation is the default answer to repetitive work. The infrastructure keeps millions of players connected, and the team takes that seriously!</p> <h3 data-path-to-node="8">The Role</h3> <p data-path-to-node="9">The Senior SRE at 2K is a hands-on technical leader—shaping production infrastructure across multiple clouds and regions while partnering with network engineers, systems architects, and game studio developers. This is an ownership role: driving technical direction, influencing reliability from architecture review through production operation, and closing the gap between what engineering ships and what players experience.</p> <h3 data-path-to-node="10">What You'll Do</h3> <p data-path-to-node="11"><strong data-path-to-node="11" data-index-in-node="0">Platform &amp; Infrastructure</strong></p> <ul data-path-to-node="12"> <li> <p data-path-to-node="12,0,0">Design, build, and operate scalable multi-cloud and hybrid infrastructure using Terraform, Pulumi, and GitOps workflows (ArgoCD, Flux).</p> </li> <li> <p data-path-to-node="12,1,0">Own Kubernetes platforms (EKS, GKE) end-to-end cluster lifecycle, multi-tenancy, networking (Istio, Cilium), and autoscaling.</p> </li> <li> <p data-path-to-node="12,2,0">Push progressive delivery patterns (blue/green, canary) across game service deployments.</p> </li> </ul> <p data-path-to-node="13"><strong data-path-to-node="13" data-index-in-node="0">Observability &amp; Reliability</strong></p> <ul data-path-to-node="14"> <li> <p data-path-to-node="14,0,0">Build and run the full observability stack: Prometheus + Grafana + Datadog.</p> </li> <li> <p data-path-to-node="14,1,0">Define SLI/SLO/error budget policies and build alerting that cuts through the noise.</p> </li> <li> <p data-path-to-node="14,2,0">Lead chaos engineering exercises to surface failure modes before players encounter them.</p> </li> <li> <p data-path-to-node="14,3,0">Drive incident response and post-mortems with a focus on systemic fixes and real follow-through.</p> </li> </ul> <p data-path-to-node="15"><strong data-path-to-node="15" data-index-in-node="0">Automation, Security &amp; Developer Experience</strong></p> <ul data-path-to-node="16"> <li> <p data-path-to-node="16,0,0">Eliminate toil through self-service provisioning, automated remediation, and intelligent scaling.</p> </li> <li> <p data-path-to-node="16,1,0">Harden CI/CD pipelines (GitHub Actions, Jenkins, ArgoCD).</p> </li> <li> <p data-path-to-node="16,2,0">Embed security at the platform layer through secrets management (PasswordState, 1Password, and AWS Secrets Manager) and policy-as-code (OPA/Gatekeeper).</p> </li> </ul> <p data-path-to-node="17"><strong data-path-to-node="17"…
Skills asked for
- sre
- ci/cd
- aws
- gcp
- terraform
- pulumi
- argocd
- kubernetes
Your next role is already in here.
Search live openings from thousands of employers, save the ones worth a second look, and let JobBob keep watch for the rest.