Site Reliability Engineer - Backup Infrastructure
Listed on 2026-10-10
-
IT/Tech
Disaster Recovery IT, Cloud Computing: Infrastructure & Operations
Location Plano, Texas, 75024 Category Accounting & Finance, Fintech, & Treasury Job Posted Date 10/07/2026
Collaborative. Respectful. A place to dream and do. These are just a few words that describe what life is like one of the world’s most admired brands, Toyota is growing and leading the future of mobility through innovative, high-quality solutions designed to enhance lives and delight those we serve. We’re looking for talented team members who want to Dream. Do. Grow.
with us.
An important part of the Toyota family is Toyota Financial Services (TFS), the finance and insurance brand for Toyota and Lexus in North America. While TFS is a separate business entity, it is an essential part of this world-changing company- delivering on Toyota's vision to move people beyond what's possible. At TFS, you will help create best-in-class customer experience in an innovative, collaborative environment.
Toyota does not offer support or sponsorship of job applicants for employment-based visas or any other work authorization for this role now or in the future. You must have the right to work in the United States and not require
Toyota support or sponsorship for immigration-related employment (e.g., H-1B, O-1, E-3, H-1B1, TN, F-1 OPT, F-1 STEMOPT, F-1 CPT, ‘job flexibility benefits’ [also known as I-140 or Adjustment of Status portability], etc.) now or in thefuture. You should not apply for this role if you will require Toyota to assist with immigration support or sponsorship now or in the future.
The Toyota Financial Services Technology Operations Center is looking for a passionate and highly motivated Senior Site Reliability Engineer (SRE) - Backup Infrastructure. In this role, you will apply software engineering principles to ensure the reliability, availability, and performance of enterprise backup infrastructure. You will play a key role in maintaining, modernizing, and automating backup ecosystems, ensuring that daily backups and restores meet defined Recovery Time Objectives (RTO) and Recovery Point Objectives (RPO).
What you’ll be doing- Manage and maintain enterprise backup infrastructure including AWS Backups, Net Backup, Cohesity, and native database backup solutions.
- Ensure backup infrastructure is reliable and performant, with direct ownership of the reliability of daily backups and restores.
- Build and maintain automation to improve backup and restore workflows, increase efficiency, and improve developer experience.
- Design and implement observability and alerting in tools such as Dynatrace to proactively detect backup issues and failures.
- Generate automated reports for auditing, compliance, and operational visibility.
- Troubleshoot complex production issues related to backup failures and implement durable fixes to improve reliability.
- Participate in capacity planning, restore testing, disaster recovery, and business continuity exercises.
- Define and manage SLIs/SLOs, health checks, and automated remediation processes for backup services.
- Collaborate across infrastructure, engineering, and operations teams to ensure service reliability and operational readiness.
- Support incident response, postmortems, and follow-up action items to reduce recurring issues.
- Participate in on-call rotations and major incident restoration as needed.
- Bachelor’s degree in information technology, computer science, or related field.
- 5+ years of hands-on experience managing enterprise backup infrastructure in production environments, including AWS Backups, Net Backup, Cohesity, and native database backup solutions.
- 5+ years of experience ensuring reliability, availability, and operational health of daily backups and restores.
- 4+ years of…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).