PENN Entertainment, Inc. is North America’s leading provider of integrated entertainment, sports content, and casino gaming experiences. From casinos and racetracks to online gaming, sports betting and entertainment content, we deliver the experiences people want, how and where they want them.
We’re always on the lookout for those who are passionate about creating and delivering cutting‑edge online gaming and sports media products. Whether it’s through Hollywood Casino, the Score Bet Sports book, or the Score media app, we’re excited to push the boundaries of what’s possible. These state‑of‑the‑art platforms are powered by proprietary in‑house technology, a key component of PENN’s omnichannel gaming and entertainment strategy.
When you join PENN Entertainment’s digital team, you’ll not only work on these cutting‑edge platforms through the Score and PENN Interactive, but you’ll also be part of a company that truly cares about your career growth. We’re committed to supporting you as you expand your skills and explore new opportunities.
With locations throughout North America, you can build a future at PENN Entertainment wherever you are. If you want to challenge conventions in gaming, media and entertainment, we want to talk to you.
About the Role& TeamThe SRE team at PENN Entertainment is looking for a Senior Site Reliability Engineer to help build and operate the infrastructure behind a large-scale sports betting and media platform. You'll own critical infrastructure across compute, networking, storage, and cloud services (GCP/AWS) — driving complex migrations, building platform tooling and automation (ArgoCD, Helm, Git Hub Actions), and improving observability and incident response across hundreds of production services spanning multiple regulated jurisdictions.
We're looking for someone with strong Kubernetes and distributed systems experience, proficiency in Go, Python or Bash, and a track record of leading cross‑team infrastructure projects with high autonomy. You'll solve ambiguous problems, reduce operational toil through automation, mentor teammates, and bring a production‑first perspective to architecture decisions — all on a team that values pragmatic engineering, real ownership, and continuous improvement of how we work.
the Work
- Drive complex infrastructure migrations and projects — scoping, planning, execution, and validation across multiple production environments and jurisdictions
- Build and maintain platform tooling and automation — ArgoCD, Helm, Git Hub Actions, release pipelines, and service onboarding workflows that reduce toil for SRE and development teams
- Support development teams — consult on infrastructure needs, unblock cross‑team dependencies, review architecture proposals, and help teams adopt platform tooling and best practices
- Design and improve observability and alerting — Datadog monitors, dashboards, and runbooks that surface meaningful signals and make systems operable by the whole team
- Provide operational support and incident response — investigate and resolve production issues through structured debugging and root cause analysis, and contribute to on‑call rotations to maintain platform reliability
- 5+ Years of Experience in a similar role (Dev Ops, Site Relatability Engineer)
- Strong experience operating and troubleshooting Kubernetes in a production Linux environment (cluster lifecycle, networking, storage, scheduling)
- Experience working with AWS, GCP, and/or on‑premise environments
- Proficiency in at least two of:
Go, Python, Bash/Shell — for building tooling, automation, and debugging production systems - Deep understanding of distributed systems — failure modes, networking fundamentals, capacity planning, and performance analysis
- Experience with Git…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).