×
Register Here to Apply for Jobs or Post Jobs. X

Senior Software Engineer (Platform Data Reliability & Automation

Job in San Diego, San Diego County, California, 92189, USA
Listing for: Sony Playstation
Full Time position
Listed on 2026-08-22
Job specializations:
  • Software Development
    Cloud Engineer - Software, DevOps, Backend Developer
Salary/Wage Range or Industry Benchmark: 177300 - 265900 USD Yearly USD 177300.00 265900.00 YEAR
Job Description & How to Apply Below
Position: Senior Software Engineer (Platform Data Reliability & Automation)

Role Overview

We are seeking a Senior Software Engineer (Platform Data Reliability & Automation) to play a critical role in building, automating, and operating scalable data platforms with a strong emphasis on Infrastructure as Code (IaC) and cloud technologies.

This role focuses on the reliability and automation of No

SQL, Streaming, and Caching services across AWS and GCP environments. You’ll design robust automation frameworks, ensure high availability, and partner with product and platform teams to deliver resilient infrastructure supporting billions of transactions and millions of players globally.

By embracing Development & DBRE principles, driving automation‑first practices, and applying AI/ML where applicable, you’ll enhance system uptime and reduce manual toil, enabling velocity for engineering teams across Play Station.

You’ll work closely with platform and product teams to ensure seamless integration and delivery of high-performance, scalable solutions across Play Station’s global ecosystem. Your contributions will directly support the reliability, scalability, and operational excellence of our data platform powering millions of players worldwide.

Responsibilities
  • Design and implement IaC and automate the provisioning, monitoring, scaling, and lifecycle management of No

    SQL, Streaming, and Caching platforms (e.g., Cassandra, Aerospike, Kafka, Redis).
  • Drive end-to-end automation to enable repeatable, reliable, and self-service deployment of data services across cloud and hybrid environments.
  • Ensure high availability, scalability, and resiliency of the platform data solutions.
  • Define and enforce SLIs, SLOs, and error margins for data platforms to drive reliability engineering practices.
  • Build highly performant, self‑healing systems, automated failover, and auto‑scaling solutions for databases and streaming platforms.
  • Develop observability solutions (metrics, logging, tracing) for Cassandra, Aerospike, Redis, and Kafka/MSK to ensure proactive issue detection.
  • Partner with engineering and platform teams to provide reliable, scalable, and performant data services.
  • Lead incident response for critical database/caching/streaming issues and drive root cause analysis with permanent automated fixes.
  • Explore and apply AI‑driven approaches to automation (e.g., anomaly detection, predictive scaling, automated remediation) to enhance operational efficiency.
  • Drive and implement best practices, procedures, operational playbooks to facilitate knowledge sharing and support continuous improvement across global teams.
  • Mentor junior engineers and influence best practices in automation, distributed systems, and database reliability.
Skills and Qualifications
  • Bachelor’s or Master’s degree in Computer Science or a related field.
  • 6+ years of software development and DBRE experience, with at least 3+ years focused on Go and Infrastructure as Code with an emphasis on automation.
  • Deep proficiency in Go (Golang), writing performant, idiomatic, and maintainable code for production‑scale systems.
  • Proven experience designing modular, domain‑driven architectures in Go, supporting large and complex backend services.
  • Expertise with infrastructure‑as‑code tools such as Terraform, Ansible.
  • Deep expertise operating large‑scale No

    SQL, caching, and streaming platforms (Apache Kafka, Redis, AWS MSK, etc.) including tuning, compaction strategies, repair operations, backup/recovery, and performance optimization.
  • Solid understanding of Linux internals, networking, and storage systems.
  • Experience building, deploying and operating stateful workloads on Kubernetes, including automation and lifecycle management of database and streaming platforms.
  • Hands‑on experience with AWS and/or GCP, including managed services such as MSK, DynamoDB, Elasti Cache, or equivalent technologies.
  • Strong problem‑solving and analytical skills, with a passion for automation and distributed systems reliability.
  • Excellent communication and collaboration skills, with experience mentoring and influencing peers across diverse teams.
  • Experience building internal developer platforms, self‑service infrastructure, or platform engineering solutions that improve developer productivity…
Position Requirements
10+ Years work experience
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary