Senior Site Reliability Engineer
Listed on 2026-07-08
-
IT/Tech
AI Engineer (Applied/Software), SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, Systems Engineer
Overview Our Culture and Impact
Cvent is a leading meetings, events, and hospitality technology provider with more than 5,500+ employees and ~30,000 customers worldwide, including 60% of the Fortune 500. Founded in 1999, Cvent delivers a comprehensive event marketing and management platform for marketers and event professionals and offers software solutions to hotels, special event venues and destinations to help them grow their group/MICE and corporate travel business.
Our technology brings millions of people together at events around the world. In short, we’re transforming the meetings and events industry through innovative technology that powers the human connection.
Cvent's strength lies in its people, fostering a culture where everyone is encouraged to think like entrepreneurs, taking risks and making decisions confidently. We value diverse perspectives and celebrate differences, working together with colleagues and clients to build strong connections.
AI at Cvent:Leading the Future
Are you ready to shape the future of work at the intersection of human expertise and AI innovation? At Cvent, we’re committed to continuous learning and adaptation—AI isn’t just a tool for us, it’s part of our DNA. We’re looking for candidates who are eager to evolve alongside technology. If you love to experiment boldly, share your discoveries, and help define best practices for AI-augmented work, you’ll thrive here.
Our team values professionals who thoughtfully integrate AI into their daily work, delivering exceptional results while relying on the human judgment and creativity that drive real innovation.
Throughout our interview process, you’ll have the chance to demonstrate how you use AI to learn, iterate, and amplify your impact. If you’re excited to be part of a team that’s leading the way in AI-powered collaboration, we’d love to meet you.
As a Senior Site Reliability Engineer, you will use your advanced development and operations knowledge to run Infrastructure as Code applications, build pipelines, and enable development teams. You will guide development teams through infrastructure decisions, conduct incident retrospectives, and implement and maintain SLIs/SLOs. You will help teams evaluate their reliability posture and prioritize work to solve reliability issues, identify and prioritize issues, find universal solutions to common problems, and mentor and support junior staff.
You will put AI at the center of how we work – using AI tools and agents to automate toil, accelerate incident response, and build smarter, self‑healing systems. As a Cvent SRE you will be a force for positive change and drive continuous improvement.
- Enlighten, enable and empower a fast‑growing set of multi‑disciplinary teams, across multiple applications and locations.
- Guide development teams through infrastructure decisions and help them evaluate their reliability posture and prioritize work to solve reliability issues.
- Run Infrastructure as Code applications, build pipelines, and enable development teams.
- Tackle complex development, automation and business process problems.
- Champion Cvent standards and best practices.
- Ensure the scalability, performance, and resilience of our suite of products.
- Work with the development and product team of a new application to establish the right monitoring and alerting strategy.
- Implement and maintain SLIs/SLOs and conduct incident retrospectives.
- Develop build, test and deployment automation that seamlessly targets multiple on‑premises and AWS regions.
- Help a dev team working on a legacy code base to realize zero‑down‑time deployments.
- Focus on automation of tasks and AI‑driven operational efficiency – automate all the things!
- Leverage AI tools and agents (e.g., Claude) to accelerate development, reduce toil, and speed up incident response.
- Build and integrate AI‑driven automation into monitoring, alerting, and remediation workflows – such as anomaly detection, intelligent runbooks, and auto‑remediation.
- Apply AI/ML techniques to observability, log analysis, and capacity planning to surface issues earlier and resolve them faster.
- Evaluate, pilot, and champion emerging AI tooling,…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).