Senior Manager, Site Reliability Engineering
Our Culture and Impact
Cvent is a leading meetings, events, and hospitality technology provider with more than 5,500+ employees and ~30,000 customers worldwide, including 60% of the Fortune 500. Founded in 1999, Cvent delivers a comprehensive event marketing and management platform for marketers and event professionals and offers software solutions to hotels, special event venues and destinations to help them grow their group/MICE and corporate travel business. Our technology brings millions of people together at events around the world. In short, we’re transforming the meetings and events industry through innovative technology that powers the human connection.
Cvent's strength lies in its people, fostering a culture where everyone is encouraged to think like entrepreneurs, taking risks and making decisions confidently. We value diverse perspectives and celebrate differences, working together with colleagues and clients to build strong connections.
AI at Cvent: Leading the Future
Are you ready to shape the future of work at the intersection of human expertise and AI innovation? At Cvent, we’re committed to continuous learning and adaptation—AI isn’t just a tool for us, it’s part of our DNA. We’re looking for candidates who are eager to evolve alongside technology. If you love to experiment boldly, share your discoveries, and help define best practices for AI-augmented work, you’ll thrive here. Our team values professionals who thoughtfully integrate AI into their daily work, delivering exceptional results while relying on the human judgment and creativity that drive real innovation.
Throughout our interview process, you’ll have the chance to demonstrate how you use AI to learn, iterate, and amplify your impact. If you’re excited to be part of a team that’s leading the way in AI-powered collaboration, we’d love to meet you.
Cvent's site reliability organization is seeking a hands-on technical manager to lead a globally distributed "Product SRE" team. The platform runs more than 8 million events per year, including live gatherings with over 50,000 attendees. Your team will operate cloud-native infrastructure, build reliability tooling and AI agents, accelerate software delivery for product development teams, and lead the technical onboarding of newly acquired companies. You will run towards the fire, write code alongside your team, raise the bar, and build a team that does the same.
In This Role, You Will:
- Set the strategy for your team and help shape the overall SRE program for the company
- Build partnerships with engineering and platform teams around reliability standards and shared practices
- Champion a culture of excellence in site reliability and automation
- Lead and technically mentor a team of SREs, writing code, reviewing pull requests, and pairing on hard problems alongside them
- Directly manage a team of 5–10 SREs: set clear expectations, run 1:1s, build individualized growth plans, deliver candid feedback, and own hiring and performance decisions
- Own the reliability, security posture, and cloud cost efficiency of the infrastructure your team operates
- Be on call as the team's escalation point and Incident Commander
- Enable partner engineering teams to continuously improve software delivery and developer experience
- Partner with product teams to evaluate and strengthen SLI/SLO definitions, translating each product's reliability needs into objectives that are achievable and meaningful
- Lead the technical onboarding of acquired companies, integrating their infrastructure, teams, and systems onto Cvent's platform and reliability standards
- Foster a culture of learning and technical ownership across the team and its engineering partners
- Design, build, and own AI agents and reliability tooling with dedicated budget
- Apply AI to your own management work: use it to accelerate QBR prep, KTLO analysis, and routine reporting
- Drive AI adoption across the team's full SDLC, from planning and code review to incident response and postmortems
Here's What You Need:
- 5+ years of experience managing technology professionals
- 10+ years of overall technology experience
- Proven experience running large-scale distributed systems in AWS, including container orchestration (ECS, EKS, Kubernetes); familiarity with Azure or multi-cloud architecture a plus
- Deep command of SRE principles and modern platform engineering practices
- Excellence in technical team leadership, with a history of developing engineers through pairing, code review, and hands-on mentorship
- Hands-on experience with Infrastructure as Code at production scale (AWS CDK preferred; Terraform or equivalent welcome)
- Strong hands-on background in automation and CI/CD, with experience building and maintaining pipelines from the ground up
- Fluency with modern observability platforms (Datadog or equivalent), including hands-on use of logs, traces, and dashboards and the ability to set team-wide standards for their use
- Demonstrated AI fluency: active use of AI coding assistants, LLM-based tooling, or AIOps platforms, with a track record of bringing your team along
- Ability to write and review production-quality code (TypeScript or Java preferred; Python, Go, or equivalent welcome); this is a working technical role, not a purely managerial one
- Experience managing roadmaps and resource allocations
This job posting is intended to comply with all applicable laws. If we learn during the course of our recruitment process that, due to an applicant’s location, further information about the position is required, including certain salary information, this information in this posting will be supplemented accordingly.
The estimated base salary range for new hires into this role is $200,000 - $260,000 annually + bonus depending on factors such as job-related knowledge, relevant experience, and location. We also offer a competitive benefits package, details of which can be found here.
Hybrid: 2 days in office
We are not able to offer sponsorship for this position
- Set the strategy for your team and help shape the overall SRE program for the company
- Build partnerships with engineering and platform teams around reliability standards and shared practices
- Champion a culture of excellence in site reliability and automation
- Lead and technically mentor a team of SREs, writing code, reviewing pull requests, and pairing on hard problems alongside them
- Directly manage a team of 5–10 SREs: set clear expectations, run 1:1s, build individualized growth plans, deliver candid feedback, and own hiring and performance decisions
- Own the reliability, security posture, and cloud cost efficiency of the infrastructure your team operates
- Be on call as the team's escalation point and Incident Commander
- Enable partner engineering teams to continuously improve software delivery and developer experience
- Partner with product teams to evaluate and strengthen SLI/SLO definitions, translating each product's reliability needs into objectives that are achievable and meaningful
- Lead the technical onboarding of acquired companies, integrating their infrastructure, teams, and systems onto Cvent's platform and reliability standards
- Foster a culture of learning and technical ownership across the team and its engineering partners
- Design, build, and own AI agents and reliability tooling with dedicated budget
- Apply AI to your own management work: use it to accelerate QBR prep, KTLO analysis, and routine reporting
- Drive AI adoption across the team's full SDLC, from planning and code review to incident response and postmortems
- 5+ years of experience managing technology professionals
- 10+ years of overall technology experience
- Proven experience running large-scale distributed systems in AWS, including container orchestration (ECS, EKS, Kubernetes); familiarity with Azure or multi-cloud architecture a plus
- Deep command of SRE principles and modern platform engineering practices
- Excellence in technical team leadership, with a history of developing engineers through pairing, code review, and hands-on mentorship
- Hands-on experience with Infrastructure as Code at production scale (AWS CDK preferred; Terraform or equivalent welcome)
- Strong hands-on background in automation and CI/CD, with experience building and maintaining pipelines from the ground up
- Fluency with modern observability platforms (Datadog or equivalent), including hands-on use of logs, traces, and dashboards and the ability to set team-wide standards for their use
- Demonstrated AI fluency: active use of AI coding assistants, LLM-based tooling, or AIOps platforms, with a track record of bringing your team along
- Ability to write and review production-quality code (TypeScript or Java preferred; Python, Go, or equivalent welcome); this is a working technical role, not a purely managerial one
- Experience managing roadmaps and resource allocations
This job posting is intended to comply with all applicable laws. If we learn during the course of our recruitment process that, due to an applicant’s location, further information about the position is required, including certain salary information, this information in this posting will be supplemented accordingly.
The estimated base salary range for new hires into this role is $200,000 - $260,000 annually + bonus depending on factors such as job-related knowledge, relevant experience, and location. We also offer a competitive benefits package, details of which can be found here.
Hybrid: 2 days in office
We are not able to offer sponsorship for this position
Required Skills
Required Languages
🇬🇧 English