×
Register Here to Apply for Jobs or Post Jobs. X

Platform Reliability, Availability, Serviceability

Job in Santa Clara, Santa Clara County, California, 95053, USA
Listing for: Qualcomm
Full Time position
Listed on 2026-09-03
Job specializations:
  • IT/Tech
    SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, Systems Engineer, Network Engineer
Salary/Wage Range or Industry Benchmark: 135000 - 180000 USD Yearly USD 135000.00 180000.00 YEAR
Job Description & How to Apply Below

Qualcomm seeks a Platform Reliability, Availability, Serviceability engineer to ensure always on performance of large scale telecom and media platforms. You will design and implement highly available, fault tolerant services, build monitoring and alerting, and drive incident response for 5G and connectivity solutions. Partner with software, hardware, and SRE teams to improve resiliency, automate recovery, and optimize SLAs. Analyze production issues, perform root cause analysis, and implement long term fixes.

This role suits engineers who thrive in fast paced, innovative environments and enjoy solving complex distributed systems challenges.

Responsibilities

  • Design and maintain highly available, fault tolerant telecom and media platforms
  • Implement monitoring, logging, and alerting for large scale distributed systems
  • Lead and participate in incident response, troubleshooting, and post mortems
  • Perform root cause analysis and implement long term reliability fixes
  • Develop automation for deployment, recovery, and scaling of services
  • Collaborate with software, hardware, SRE, and network teams on resiliency improvements
  • Define and track SLAs, SLOs, and SLIs for critical services
  • Optimize performance, capacity, and reliability of 5
  • G and connectivity platforms
  • Contribute to reliability architecture, tooling, and best practices
  • Document systems, runbooks, and reliability standards

Required Skills

  • Site Reliability Engineering (SRE)
  • Distributed systems design
  • Linux administration
  • Cloud platforms (AWS, GCP, or Azure)
  • Kubernetes and containers
  • Monitoring and observability
  • Incident response and troubleshooting
  • Automation and scripting (Python, Bash)
  • Networking and TCP/IPCI/CD pipelines
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary