About the Role
As the Enterprise Systems Operations Manager at Cboe, you lead the team responsible for the real-time health of enterprise infrastructure. Your team watches systems, responds to alerts, owns incident response, and drives operational workflows. A core expectation is leading the adoption and scaling of AI-assisted operations, building Claude skills and agentic workflows to accelerate monitoring, triaging, incident response, and operational artefact generation.
Responsibilities
- Own the monitoring and observability programme, covering alert coverage, log aggregation, dashboard health, and the full path from detection to resolution.
- Keep alerts accurate, actionable, and correctly routed; cut noise, false positives, and alert fatigue.
- Oversee log management across enterprise platforms, ensuring logs are retained, searchable, and used in investigations.
- Build and scale Claude skills and agentic workflows for the team, spanning alert triage, log summarisation, incident classification, runbook lookup, status updates, drift detection, and Jira ticket creation, and document them for the broader Enterprise Systems team.
- Replace repetitive operational tasks with autonomous agents, and measure the impact on response times.
- Own the on-call programme: scheduling, rotation coverage, escalation paths, and after-hours standards.
- Lead major incident response across Engineering, Security, and application teams to restore service fast, then run post-incident reviews to find root causes.
- Automate routine work with PowerShell, Graph API, and infrastructure-as-code; maintain runbooks and self-healing scripts for known failures.
- Govern change management: review change requests and coordinate maintenance windows to minimise disruption.
- Own documentation standardisation and clean-up for Enterprise Systems: build an effective documentation standard, then drive your team to bring existing runbooks, procedures, and operational artefacts in line with it.
- Manage and develop the operations engineering team, coaching them to use Claude and other AI tools effectively.
- Own the Enterprise Systems support queue operations programme: Vulnerability Management, SLA tracking, ticket routing/escalation, and leadership reporting.
- Manage vendor relationships for operational tooling, including ITSM, monitoring, log management, and backup/recovery.
- Ensure audit readiness and represent operations in compliance reviews and risk assessments.
Requirements
- 5+ years in IT or infrastructure operations, including 2+ years leading a team.
- Hands-on experience with monitoring and observability platforms (Grafana, Loki, Dynatrace, Azure Monitor, or equivalent).
- Working knowledge of log management tooling and using logs actively in incident investigation.
- Solid grasp of ITIL or an equivalent service management framework.
- Experience with ITSM tooling (Jira Service Management, ServiceNow, or equivalent).
- Familiarity with Windows Server, Microsoft 365, VMware, and Azure.
- Experience running on-call programmes and leading major incident response.
- Track record of improving operational metrics such as MTTR, SLA compliance, and alert noise.
- Hands-on experience building Claude skills, LLM-based workflows, or agentic pipelines (preferred).
- PowerShell or other operational automation experience (preferred).
- Background in financial services or another regulated industry (preferred).
Skills
- Monitoring and observability platforms
- Log management tooling
- ITIL
- ITSM tooling
- Windows Server
- Microsoft 365
- VMware
- Azure
- On-call programmes
- Incident response
- PowerShell
- Graph API
- Infrastructure-as-code
- Claude skills
- LLM-based workflows
- Agentic pipelines
Location
- London
Work Type
- Four day in office work model
Experience Level
- 2+ years leading a team
- 5+ years in IT or infrastructure operations
Benefits
- Private Medical Insurance
- Life Insurance
- GP Service
- Fitness Corporate Membership
- Employee assistance Program (EAP)
- Eye Care
- Short Term Incentive (STI)
- Pension
- Income Protection
- Accident Insurance
- Business Travel Insurance
- Associate Referral Program
- Perks at Work
- Employee Stock Purchase Plan (ESPP)
- Commuting Allowance
- LinkedIn Learning courses
- Service Awards
- Subsided Lunch
- Corporate Events
- Education Assistance
- Holiday (Annual Leave)
- Enhanced Leave (Maternity, Paternity, Adoption, Compassionate Leave, Community Service / Volunteering Day & School Visit)
- Working Abroad Days Allowance
About the Company
- Cboe Global Markets is a leading provider of market infrastructure and tradable products, delivering cutting-edge trading, clearing and investment solutions to market participants around the world.
- We are reimagining the future of the workplace by focusing on what matters most, our people.
- Our journey is an inclusive one, investing deeply in leadership programs and career development initiatives that ensure everyone has an equal chance to succeed.
- We work with purpose, solving problems with ingenuity, collaboration, and a lot of passion.
- We’re an engaged and excited team connecting markets across borders and embracing growth in all its forms to achieve incredible outcomes.
Equal Opportunity
- Cboe Global Markets is proud to be an equal opportunity employer and does not discriminate against any employee or applicant for employment based on any legally protected characteristic, including race, color, religion, sex, sexual orientation, gender identity, national origin, age, disability, genetic information, or Veteran status.
- We are committed to fostering a workplace where all individuals are valued and respected.
