More jobs:
Site Reliability Engineer at Hitachi
Job in
Toronto, Ontario, C6A, Canada
Listed on 2026-07-26
Listing for:
Socket.dev
Full Time
position Listed on 2026-07-26
Job specializations:
-
IT/Tech
SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, Unix/Linux, IT Infrastructure
Job Description & How to Apply Below
Join Hitachi's SRE Operations team to reinforce application reliability across cloud-native and hybrid platforms. This position requires 2-5 years of experience in IT Operations or SRE, a solid understanding of Kubernetes, and proficiency in Linux troubleshooting. Key responsibilities include monitoring infrastructure, performing incident triage, and collaborating with engineering teams to ensure service reliability.
Key Responsibilities:
• Monitor applications and infrastructure across environments
• Perform incident triage and execute operational runbooks
• Troubleshoot application issues using Linux utilities
• Support Kubernetes deployment and validation of pod health
• Maintain incident documentation and knowledge base updates
Requirements:
• 2–5 years in IT Operations, SRE, or Dev Ops
• Strong Linux administration knowledge
• Familiarity with AWS, Azure, or GCP
• Experience with monitoring tools like Prometheus
• Basic scripting skills in Python or Bash
Bring your passion for automation and operational excellence to Hitachi's innovative SRE team.
#J-18808-Ljbffr
Note that applications are not being accepted from your jurisdiction for this job currently via this jobsite. Candidate preferences are the decision of the Employer or Recruiting Agent, and are controlled by them alone.
To Search, View & Apply for jobs on this site that accept applications from your location or country, tap here to make a Search:
To Search, View & Apply for jobs on this site that accept applications from your location or country, tap here to make a Search:
Search for further Jobs Here:
×