Lead Site Reliability Engineer (Security)
Cvent · Gurgaon
- Experience9–14 yrs
- SalaryNot disclosed
- Work modeonsite
- Levelexecutive
- Posted11 Sept 2026
About Cvent
Cvent is hiring in Gurgaon in hospitality travel. This role looks for around 9+ years of experience.
Skills
- AWS WAF
- AWS Shield Advanced
- bot mitigation
- AWS
- CloudFormation
- AWS CDK
- CI/CD
- Jenkins
- TypeScript
- JavaScript
- Python
- Ruby
- Bash
- Agile
- prompt engineering
- Retrieval-Augmented Generation
The role
A site reliability and platform engineer at an event technology company designs secure production systems using AWS WAF, AWS Shield Advanced, and bot mitigation, and improves DevSecOps operations through automation and AI-assisted workflows. Expertise also covers CI/CD pipelines and cloud infrastructure as code.
Full job description
Cvent is a leading meetings, events, and hospitality technology provider with more than 6,000 employees and 30,000 customers worldwide, including 60% of the Fortune 500. Founded in 1999, Cvent delivers a comprehensive event marketing and management platform for marketers and event professionals and offers software solutions to hotels, special event venues and destinations to help them grow their group/MICE and corporate travel business. Our technology brings millions of people together at events around the world. In short, we’re transforming the meetings and events industry through innovative technology that powers the human connection.
Cvent's strength lies in its people, fostering a culture where everyone is encouraged to think like entrepreneurs, taking risks and making decisions confidently. We value diverse perspectives and celebrate differences, working together with colleagues and clients to build strong connections.
AI at Cvent: Leading the FutureAre you ready to shape the future of work at the intersection of human expertise and AI innovation? At Cvent, we’re committed to continuous learning and adaptation—AI isn’t just a tool for us, it’s part of our DNA. We’re looking for candidates who are eager to evolve alongside technology. If you love to experiment boldly, share your discoveries, and help define best practices for AI-augmented work, you’ll thrive here. Our team values professionals who thoughtfully integrate AI into their daily work, delivering exceptional results while relying on the human judgment and creativity that drive real innovation.
Throughout our interview process, you’ll have the chance to demonstrate how you use AI to learn, iterate, and amplify your impact. If you’re excited to be part of a team that’s leading the way in AI-powered collaboration, we’d love to meet you.
Recruitment Scam Notice: Beware of recruitment scams. Legitimate Cvent recruiting communications will come from accounts formatted as “name@cvent.com” email address. Cvent will never request payment or ask for sensitive personal or financial information through chat or social media platforms. Cvent is not responsible for losses or damages arising from communications, payments, or disclosures made to third parties impersonating Cvent. If you believe a communication may be fraudulent, we strongly recommend that you don't respond and report the incident at recruitingindia@cvent.com. For more information, visit Cvent’s Recruitment Fraud Notice. https://www.cvent.com/en/notice-recruitment-fraud.
About the role: As a Lead SRE on the SRE Security team, you will be responsible for mentoring others and helping Cvent to both envision and achieve our DevSecOps goals. We are looking for someone with the drive, ownership and ability to take on challenging problems, both technical and process related, in a dynamic, collaborative and highly distributed, multi-disciplinary team environment. You will use your background as a generalist to work closely with product development teams, Information Security, Cloud Infrastructure and other SRE teams to ensure the effective and efficient maintenance of our platforms' security. You must be able to see the big picture and work collaboratively with teams to solve hard multi-disciplinary problems.
Technical expertise in topics such as cloud operations, the software development lifecycle, and security vulnerability management will be of great help to you. However, excellent soft skills in mentorship, communication and the ability to drive alignment are must haves. We use SRE principles such as blameless postmortems and a focus on automation to ensure we're constantly improving our knowledge and maintaining a good quality of life.
Overall, we're passionate about continuous improvement, learning and participating in dynamic day to day work where success is rewarded with recognition and upward mobility.
In this Role, You Will:Enlighten, Enable and Empower a fast-growing set of multi-disciplinary teams, across multiple applications and locations.Tackle complex development, automation and business process problems. Champion Cvent standards and best practices.Ensure the scalability, performance, and resilience of security related systems and processes.Work with product development teams, Information Security, Cloud Automation and other SRE teams to ensure a holistic understanding of security concerns and their effective and efficient identification and resolution.Identify recurring problems and anti-patterns in development, operational and security processes.Develop build, test and deployment automation that seamlessly targets multiple on-premises and AWS regions.Give back by working on and contributing to Open-Source projects.
Preferred candidate profile
Must Have Skills:9-14 years of hands-on experience in Site Reliability Engineering with a demonstrated track record of owning reliability, security, and operational excellence at scale in production environments.Excellent communication skills and a track record of driving alignment across multi-disciplinary teams.A passion for and track record in making things better for your peers.Hands-on experience with AWS WAF including rule authoring, rate-based rules, bot control integration, WAF rule group management, and multi-product WAF sharing strategies (e.g., managing WAF rule limits across applications sharing the same WebACL).Experience designing and implementing DDoS protection using AWS Shield Advanced including transitioning endpoints from count to block mode, building observability solutions (Lambda + CloudWatch alarms), and selfservice enablement for product teams.Experience with bot mitigation strategies including AWS Bot Control, silent challenge / token-based traffic classification (verified humans, verified bots, unknown traffic), JA4+ASN fingerprinting, and evaluation of thirdparty bot mitigation vendors (e.g., Datadome).Experience managing AWS services and operational knowledge of running applications in AWS ideally via automation and Infrastructure as Code (IaC) using CloudFormation or CDK.Strong understanding of CI/CD pipelines experience with Jenkins or equivalent, PR-based deployment workflows, build/test/deploy automation, and troubleshooting pipeline failures in distributed environments.Incident management experience — able to act as IC, write clear incident summaries, drive RCA, and coordinate resolution across teams under pressure.Change management discipline — ability to communicate changes proactively to stakeholders, document rollout strategies, and manage phased production deployments with rollback plans.Fluent in at least one scripting language such as TypeScript, JavaScript, Python, Ruby, or Bash.Experience with SDLC methodologies (preferably Agile).AI-assisted Workflow & Process Automation — experience using or building AI-powered automations in operational contexts, such as automated incident summarization, alert enrichment, change risk assessment, or post-mortem drafting using LLM integrations (e.g., via MCP tools, Slack bots, or custom pipelines).AI & Automation Literacy (Must Have): Practical understanding and hands-on exposure to AI fundamentals as applied to SRE and operational workflows:Prompt Engineering — ability to design effective prompts for LLMs to assist with incident analysis, RCA generation, runbook creation, and on-call triage.Retrieval-Augmented Generation (RAG) — basic understanding of RAG patterns; ability to leverage or contribute to RAG-based internal tools that surface relevant runbooks, past incidents, and knowledge base articles during operational events.
Good to Have Skills:Disaster recovery planning and execution — experience with multi-region failover, DR runbooks, and recovery time / recovery point objective (RTO/RPO) management.Experience managing CloudFront distributions, API Gateways, and ALBs as part of a layered security posture.Experience with APM, monitoring and logging tools (Datadog, New Relic, Splunk).Familiarity with security assessment tools and methodologies:Cloud Security Posture Management (CSPM)Infrastructure Vulnerability ScanningStatic Code AnalysisSoftware Composition Analysis (SCA)Static, Interactive and Dynamic Application Security Testing (SAST, IAST and DAST)Runtime Application Self Protection (RASP)Good understanding of containerization concepts — Docker, ECS, EKS, Kubernetes.Experience managing 3-tier application stacks.Understanding of basic networking concepts.Familiarity with risk assessment and management concepts and practices.