×
Register Here to Apply for Jobs or Post Jobs. X

Site Reliability Engineer

Job in Vancouver, BC, Canada
Listing for: Apple
Full Time position
Listed on 2026-01-29
Job specializations:
  • Software Development
Salary/Wage Range or Industry Benchmark: 100000 - 125000 CAD Yearly CAD 100000.00 125000.00 YEAR
Job Description & How to Apply Below

Summary

The Apple Service Engineering - SRE team is looking for Site Reliability Engineers with experience in developing processes, tools, and automation for managing distributed systems in production environments. Our SRE team combines software and systems engineering and system administration practices to build and run large-scale, massively distributed, fault-tolerant systems. Our software ensures that Apple’s services are reliable, scalable and secure, and we leverage both open source and home-grown technologies to provide managed data infrastructure services.

You will help building next generation search infrastructure and platform services, collaborating cross-functionally with various ASE teams, from store and commerce to search and recommendations. You’ll create platforms that can rapidly scale to serve personalized and non-personalized data with very low latencies. You should be someone who is not afraid to question assumptions, are a good standout colleague under tight deadlines, and can take on problems with elegant technical solutions.

Description

The ASE SRE team develops applications and tooling that are safe, reliable, scalable, and fast. Our Data Reliability Engineering team is responsible for all aspects of managing Voldemort key-value distributed database infrastructure deployment on on-premise bare metal and public cloud platforms, including maintenance, deployment automation, backup, observability and telemetry, with focus on reliability, performance, and scaling to deliver continuous data store availability to ASE Media Applications.

Success in this role requires expertise in several of the following:

  • Understanding of core SRE concepts - Monitoring, Alerting, Incident management
  • Performance engineering (design concepts, profile-guided optimization)
  • Service management across bare metal, and virtualized (EC2) platforms
  • Prepare alert handling procedures, run-books, and collaborate with other SRE team members.
  • Excellent communication and a high degree of customer focus when engaging with internal platform customers
  • As a distributed team, ability to work optimally with colleagues based in other locations is also essential; experience in this area is a plus
  • Prior experience with development or maintenance of distributed databases, and operating systems is recommended

Come join us at Apple Services Engineering and help us deliver services and applications that are fluid and responsive. You will collaborate with engineers from across Apple to define the metrics, set targets, uncover optimization opportunities, and ship a service that will delight our customers. This role is for engineers who enjoy deep technical engineering that spans large cross-organizational projects. Your openness to learning and implementing new technologies will contribute to the continuous evolution of our organization.

Good ideas are valued and rewarded.

Minimum Qualifications
  • Success in this role requires expertise in several of the following:
  • BS/MS in Computer Science or Equivalent
  • At least 2-5 years in a Reliability Engineering, Dev Ops or infrastructure focused role
  • Support of internet-facing production services and distributed systems via deployments, onCall and Incident Management.
  • Understanding of distributed database concepts (consistency models, isolation levels, crash and recovery semantics).
  • Performance engineering (design concepts, profile-guided optimization).
  • Datacenter architecture (networking topologies, host placement strategies, and failure modes); design of multi-datacenter systems; failure domains; and wide-area networking.
  • Automation advocate - prior history of removing operational toil via software.
  • Self motivated, inquisitive and always looking to learn more.
Preferred Qualifications
  • Demonstrated expertise developing distributed systems, storage engines, distributed systems, or performance engineering.
  • Experience developing critical internet services and/or platform infrastructure.
  • Proficient in one or more of the following programming languages:
    Java, Go (golang), Python
  • Optional experience managing services on Kubernetes
  • Optional experience with EC2, EBS, and Terraform
Pay & Benefits

At Apple, base pay…

Note that applications are not being accepted from your jurisdiction for this job currently via this jobsite. Candidate preferences are the decision of the Employer or Recruiting Agent, and are controlled by them alone.
To Search, View & Apply for jobs on this site that accept applications from your location or country, tap here to make a Search:
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)

Job Posting Language
Employment Category
Education (minimum level)
Filters
Education Level
Experience Level (years)
Posted in last:
Salary