SRE - Remote
New York, New York County, New York, 10261, USA
Listed on 2026-07-29
-
IT/Tech
SRE/Site Reliability, AWS, Cloud Computing: Infrastructure & Operations, Unix/Linux
About Glassbox
Glassbox is a leading force in shaping digital experiences, empowering enterprises to uncover digital issues, boost conversion rates, enhance accessibility, prevent fraud, and more. Leveraging AI‑driven customer intelligence, Glassbox enables secure, proactive digital experiences for clients that include some of the world’s largest banks, hotels, healthcare providers and telecommunications companies.
PositionGlassbox is looking for an SRE to join our global Cloud team.
What will you do?- Provide technical and operational support for customers according to defined SLAs.
- Work in cloud‑based environments (primarily AWS) and operate/support Kubernetes‑based systems (EKS).
- Design, build, and maintain advanced automation systems to enhance reliability, monitoring, and operational efficiency across production environments.
- Develop scalable monitoring and alerting solutions to proactively detect issues before they impact customers.
- Build and maintain runbooks for NOC/SOC teams.
- Serve as Tier‑2 escalation for production incidents, including collaboration with Dev Ops and participation in a 24×7 on‑call rotation.
- Implement automation‑driven improvements using scripts and configuration management tools to streamline system operations.
- Leverage modern technologies and tooling to optimize system performance, observability, and resilience.
- Work closely with the Cloud Dev Ops team to transition products from development to the production environment via continuous integration and deployment processes.
- At least 3 years of experience as an SRE/Dev Ops or in a similar cloud/monitoring role.
- Hands‑on experience with AWS.
- Strong knowledge and practical experience working with Kubernetes and EKS.
- Scripting experience with Bash and hands‑on experience working with Linux systems.
- Experience working with cloud monitoring, management, and alerting tools.
- Strong troubleshooting skills in production environments.
- Willingness to participate in a 24×7 on‑call rotation.
- Ability to work effectively as part of a collaborative team, with strong interpersonal skills and a positive, team‑oriented mindset.
- Assertive, confident, fast learner, and comfortable working in a fast‑paced environment.
- Experience with Azure and AKS.
- Knowledge of additional programming languages.
- Experience with Prometheus, Grafana, or similar monitoring/observability tools.
- Bachelor’s degree in Computer Information Systems, Management Information Systems, Computer Science, or a related field.
- AWS or Azure certifications.
At Glassbox, we value curiosity, ownership, and a constant drive to learn and improve, no matter the gender, nationality, religion, or background. We believe diverse perspectives make us better and encourage people of all shapes and sizes to apply.
If this role excites you and you’re motivated to make an impact – we’d love to hear from you, even if you don’t meet every listed qualification.
#J-18808-Ljbffr(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).