Senior Product Reliability Engineer
Position Title and Compensation
Position Title: Senior Product Reliability Engineer
Compensation: $90,000 - $130,000 + annual bonus (paid in local currency; range varies by location)
Reports To: Software Development Director
Location: Kitchener, ON (Onsite)
About the RoleWe are seeking a skilled and passionate Senior Product Reliability Engineer to join our dynamic team and contribute to the development and on-site success of utility-grade systems for clean energy. This is a deliberately hybrid role: approximately 70% of the work is customer‑facing — partnering with our Delivery and Deployment teams through project bring‑up and acting as the first point of contact from the development team for site‑reported issues after handover — with the remaining 30% contributing to R&D quality assurance, reliability engineering, and release readiness.
As Senior Product Reliability Engineer for EQ‑S, you will play a pivotal role in bridging R&D and the field for our hardware‑driven software products. You will work closely with the Quality Assurance function to bring real customer‑site context, failure patterns, and field‑derived test cases back into the product, ensuring that what works in the lab also works on energized sites.
The ideal candidate will possess strong skills in field reliability engineering, root‑cause investigation, and customer engineering for hardware software products, and be passionate about advancing clean energy initiatives in a dynamic startup‑like environment.
- Partner with the Delivery team on Factory Acceptance Testing (FAT) and project bring‑up, providing development‑side validation of test protocols, configurations, and shipped releases prior to site dispatch.
- Support the Deployment team on site‑specific configuration, system setup, and integration troubleshooting from commissioning through Commercial Operation Date (COD).
- Tune and validate product performance to meet specifications across diverse customer site configurations - including varying project size, networking topology, and hardware composition - and confirm the product behaves to spec under each customer’s deployed architecture.
- Act as the first point of contact from the development team for on‑going site issues after handover; lead structured Root Cause Analysis (RCA) and own issues end‑to‑end through reproduction, corrective action, and verification on the next release.
- Operate a closed‑loop corrective‑action process that feeds field insights back into design reviews, failure‑mode analyses, and the regression test suite.
- Author and execute commissioning protocols, performance test plans, and customer‑facing acceptance documentation; maintain reliability and availability metrics per site and across the fleet.
- Deliver customer training and produce runbooks, SOPs, and field‑facing technical documentation that translate R&D system behavior into actionable on‑site procedures.
- Advocate for customers inside R&D - bringing field constraints, deployability concerns, and observed defect patterns into sprint planning, design reviews, and release readiness - and collaborate with development, product, and project teams to drive defect resolution into product releases.
- Partner with the Quality Assurance function to extend test coverage with field‑derived scenarios; contribute to hardware‑in‑the‑loop (HIL) testing, regression and soak coverage, and CI integration for software and firmware releases.
- Contribute to Failure Mode and Effects Analyses (FMEA / FMECA) for new product features and hardware revisions, and define the telemetry, logs, and metrics needed to diagnose field issues from the office.
- Participate in a shared on‑call rotation for post‑COD critical site incidents; lead blameless postmortems and ensure…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).