Site Reliability Engineer
Listed on 2026-08-22
-
IT/Tech
SRE/Site Reliability
Select how often (in days) to receive an alert:
Create Alert
Acuity Inc. (NYSE: AYI) is a market-leading industrial technology company. We use technology to solve problems in spaces, light and more things to come. Through our two business segments, Acuity Brands Lighting (ABL) and Acuity Intelligent Spaces (AIS), we design, manufacture, and bring to market products and services that make a valuable difference in people’s lives.
We achieve growth through the development of innovative new products and services, including lighting, lighting controls, building management solutions, and an audio, video and control platform. We focus on customer outcomes and drive growth and productivity to increase market share and deliver superior returns. We look to aggressively deploy capital to grow the business and to enter attractive new verticals.
Acuity Inc. is based in Atlanta, Georgia, with operations across North America, Europe and Asia. The Company is powered by approximately 13,000 dedicated and talented associates. Visit us at .
Job SummaryAt Acuity, you will join an Agile team focused on building and supporting advanced platforms and applications that drive our business forward.
We are seeking a Site Reliability Engineer (SRE) to help define and raise the reliability bar for our Commerce Platform. As an SRE on the AI Commerce team, you will be a hands‑on contributor, operating at the intersection of software engineering, cloud infrastructure, and production operations. This role is focused on engineering‑led reliability, building resilient systems, automating operational workflows, and embedding reliability practices directly into how Commerce is designed and delivered.
You will have meaningful ownership and influence in shaping observability standards, incident response, SLOs, runbooks, and platform reliability patterns for the Commerce platform. You will collaborate closely with software engineers, Dev Ops, and product teams, ensuring the platform can scale safely while enabling teams to move fast and with confidence.
This role offers the chance to shape the future of AI commerce by ensuring our platforms are secure, scalable, and efficient. You will contribute to platform improvements, reliability, support key initiatives, and make a meaningful impact within a global leader in industrial technology.
Key Tasks & Responsibilities (Essential Functions) Reliability & Production Ownership- Own the availability, reliability, and performance of critical production environments
- Define, track, and report on service health metrics including uptime, availability, and reliability indicators.
- Drive root cause analysis (RCA), analyze system logs and ensure corrective and preventative actions are implemented.
- Part of a cross-geo team providing operational & escalation coverage, leading incident response and recovery for critical services.
- Automate operational workflows to reduce manual toil and improve consistency.
- Support and improve deployment processes for features, patches, and hotfixes while maintaining a strong security posture.
- Create, maintain, and continuously improve runbooks and standard operating procedures (SOPs).
- Design and evolve monitoring, alerting, and observability standards across the platform.
- Build and maintain dashboards and alerts that provide clear, actionable insight into system health.
- Ensure monitoring supports SLOs and operational decision‑making, not just data collection.
- Work closely with software engineers to embed reliability best practices into system design and delivery.
- Partner with product and platform teams to translate business requirements into reliable, scalable technical solutions.
- Contribute to a culture of shared production ownership and continuous improvement.
- Stay up to date with the latest technologies and industry trends to drive innovation.
- Bachelor’s degree in computer science, Engineering, or a related field, or equivalent practical experience.
- 3+ years of professional experience in software engineering, SRE, or a related role.
- Strong hands‑on…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).