Site Reliability Engineer II
Listed on 2026-07-22
-
Software Development
Unix/Linux
Service Reliability Engineer
Do you want to use transformative technologies to achieve greater scalability and efficiency? If so, join us as part of Sony Computer Entertainment’s Gaming, Developer, and Future Technology group. Our SREs focus on overall ownership of production, production code quality, and deployments.
Responsibilities- Own overall production and be self‑directed in the operational lifecycle.
- Provide critical feedback during different phases of the operational lifecycle.
- Participate in decision‑making at different levels and be engaged throughout the S/W development lifecycle, ensuring operational readiness and stability.
- Minimum of 5+ years working experience in Software Development and/or Linux Systems Administration.
- Strong written and verbal communication skills; able to work effectively with engineers and stakeholders.
- Willingness to participate in an on‑call rotation and lead/assist in incident response.
- 5+ years of experience in software development and/or Linux systems administration in production environments.
- BS in Computer Science or equivalent experience required.
- Hands‑on experience with Kong API Gateway or similar enterprise API gateway technologies; operate and troubleshoot the gateway in production: routing/services, plugins, TLS/cert management, authn/authz, and performance/latency tuning.
- Own lifecycle and reliability practices: upgrades and configuration changes (IaC/Git Ops), safe rollouts/rollback, monitoring, etc.
- Hands‑on experience with Service Mesh technologies such as Istio, Linkerd or Kuma.
- Strong understanding of traffic management concepts, including routing, retries, timeouts, circuit breaking and traffic shifting (canary, blue/green).
- Familiarity with mTLS, service‑to‑service communication, and zero‑trust networking principles.
- Experience troubleshooting service‑mesh‑related issues, including latency, connectivity, and policy misconfigurations.
- Development experience in one or more of the following programming languages:
Python (preferred), Bash, Go, Java, C++, or Rust. - Distributed data storage at scale (Ceph, Rook).
- No
SQL at scale (MongoDB clusters, sharded Redis, Cassandra). - Data aggregation technologies (Elastic Search, Kafka).
- Scaling and running traditional RDBMS (PostgreSQL, MySQL) with high availability.
- Observability and incident management:
Prometheus, Grafana, alerting/on‑call workflows, post‑incident review. - Container orchestration (Kubernetes, Rancher) at scale.
- Release engineering (package management and distribution at scale).
- S/W performance analysis and load testing (QA or SDET experience: a plus).
Base pay range: $145,700—$218,500 USD. Eligible for top‑tier benefits including medical, dental, vision, 401(k) matching, paid time off, wellness program, and employee discounts on Sony products. This role may also be eligible for a bonus package.
We conduct background checks at the offer stage for all new employees. Please refer to our Candidate Privacy Notice for more information.
Equal Opportunity Statement:
Sony is an Equal Opportunity Employer. All persons will receive consideration for employment without regard to gender (including gender identity, gender expression and gender reassignment), race (including colour, nationality, ethnic or national origin), religion or belief, marital or civil partnership status, disability, age, sexual orientation, pregnancy, maternity or parental status, trade union membership or membership in any other legally protected category.
Sony Interactive Entertainment is a Fair Chance employer and qualified applicants with arrest and conviction records will be considered for employment.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).