Production Engineer - Java Full Stack
Listed on 2026-08-02
-
IT/Tech
IT Support, Systems Engineer, Cloud Computing: Infrastructure & Operations
Overview
Collaborative. Respectful. A place to dream and do. These are just a few words that describe what life is like one of the world’s most admired brands, Toyota is growing and leading the future of mobility through innovative, high-quality solutions designed to enhance lives and delight those we serve. We’re looking for talented team members who want to Dream. Do. Grow.
with us.
An important part of the Toyota family is Toyota Financial Services (TFS), the finance and insurance brand for Toyota and Lexus in North America. While TFS is a separate business entity, it is an essential part of this world-changing company- delivering on Toyota's vision to move people beyond what's possible. At TFS, you will help create best‑class customer experience in an innovative, collaborative environment.
Toyota does not offer support or sponsorship of job applicants for employment-based visas or any other work authorization for this role now or in the future. You must have the right to work in the United States and not require Toyota support or sponsorship for immigration-related employment (e.g., H-1B, O-1, E-3, H-1B1, TN, F-1 OPT, F-1 STEM OPT, F-1 CPT, TN, ‘job flexibility benefits’ (also known as I-140 or Adjustment of Status portability), etc.
now or in the future. You should not apply for this role if you will require Toyota to assist with immigration support or sponsorship now or in the future.
The Toyota Financial Services Site Reliability Engineering is looking for a passionate and highly motivated Production Engineer. Reporting to the Manager of Production Engineering, the individual in this role will play a pivotal role in troubleshooting and resolving complex technical issues for financial software and systems, acts as an escalation point for frontline support, and collaborates with product and engineering teams to improve the product.
Additionally, the engineer will collaborate closely with the Major Incident Management team to address and resolve production environmental issues.
- Act as the primary point of contact for production support, managing and resolving tickets across various business units.
- Triage and prioritize support tickets based on urgency and impact.
- Conduct root cause analysis and apply critical thinking to solve complex problems.
- Analyze trends in support tickets to proactively identify recurring issues and implement long-term solutions.
- Document processes, resolutions, and recurring issues for knowledge management and training purposes.
- Suggest and implement improvements to processes to reduce ticket volume and enhance efficiency.
- Monitor system performance and proactively address potential issues.
- Collaborate with the Major Incident Management team to address and resolve critical issues in the production environment.
- Bachelor’s degree in information technology or related field.
- Excellent troubleshooting and problem-solving skills.
- Strong communication and collaboration skills.
- 4+ years of experience to Build, deploy, and maintain cloud-native microservices using Java, Spring Boot, and JavaScript frameworks, ensuring high availability and scalability.
- 4+ years of experience to Design and implement RESTful APIs and event-driven architectures using AWS services such as Lambda, ECS/EKS, SQS, and SNS.
- 3+ years of experience to Develop and maintain CI/CD pipelines with Jenkins, Git Lab CI, or AWS Code Pipeline for automated testing and deployment.
- 3+ years of experience to Monitor application and infrastructure health using Dynatrace, AWS Cloud Watch, Prometheus, Grafana, and distributed tracing tools like Jaeger or AWS X‑Ray.
- 4+ years of experience to Troubleshoot production issues, perform root cause analysis, and implement fixes to improve system reliability.
- 3+ years of experience to Implement security controls including IAM roles, OAuth2, JWT, and encryption for data in transit and at rest.
- 4+ years of experience to Collaborate with cross-functional teams to design fault-tolerant, resilient systems with automated failover and recovery.
- 3+ years of experience to Optimize cloud resource usage and cost through rightsizing and autoscaling…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).