Building trusted markets — powered by our people
At Cboe Global Markets, we inspire our people to solve complex challenges together because what we do matters. We provide the financial infrastructure that powers the global economy. As a leading provider of market infrastructure and tradable products, Cboe delivers cutting-edge trading, clearing and investment solutions to market participants around the world.
We're building meaningful ways to support professional and personal development while strengthening the trust we've earned as a global market leader. Our teams are empowered to share ideas, actively pursue them and bring on a challenge. As champions of internal mobility and access to opportunity, we encourage our people to "go for it" and equip our managers with the training to coach their teams to the next level. We strive to provide employees a safe space to network, share ideas and create opportunities.
Sound like the place for you? Join us!
The Site Reliability Engineer (London) is a role served by experienced technologists with a diverse set of skills ranging from software development to systems, network, application, and/or database management. The Cboe Site Reliability Engineering team is a highly skilled unit responsible for platform engineering, configuration management, implementation, capacity planning, performance tuning, analysis, troubleshooting, reporting, and process automation.
This position is instrumental in support of both Cboe’s European markets and Cboe's follow-the-sun support model for its US Global Trading Hours (GTH) markets, providing critical overnight and early-session coverage from London that ensures continuous, high-availability operations across Cboe's real-time low-latency trading platforms. The London-based SRE provides technical support to Cboe Trade Desk and Operations Support Center staff across time zones, and works closely with Software Engineering, Systems Engineering, and Network Engineering teams to troubleshoot complex issues and coordinate platform configuration updates. A Site Reliability Engineer must be able to work independently with little to no direct supervision in performing their duties.
Platform Configuration Management: Provide configuration management of new and existing trading platforms and support implementation of new features and functionality based on new business requirements. Monitor development activities, change management tickets, and evaluate their impact on Cboe Operations. Execute daily change tickets assigned to Site Reliability Engineering in support of updates to production, disaster recovery, and certification systems. While the primary focus of this role involves support of bare-metal on-premises infrastructure, experience with cloud platforms (e.g., AWS, Azure, GCP) and containerization technologies (e.g., Docker, Kubernetes) is desirable.
Incident Response & Technical Troubleshooting: Serve as a technical responder for production incidents occurring during US GTH market hours covered from the London time zone. Participate in incident triage, root cause analysis, and resolution in coordination with globally distributed engineering and operations teams. Provide timely, precise communication to stakeholders during active incidents and contribute to post-incident reviews and remediation tracking to drive long-term platform stability.
System Availability & Technical Support: Provide technical support and operational oversight to sustain resiliency and high availability of critical business operations. Monitor production, disaster recovery, and certification systems for issues. Analyze and optimize performance of real-time trading platforms. Operate and maintain low-latency bare-metal infrastructure, including hardware health, Linux OS tuning, and kernel-bypass networking stacks such as Solarflare/Onload. Investigate software defects. Assist the build team to resolve build/deployment issues.
Reporting & Data Analysis: Create and improve upon existing reports related to Operations management. Analyze technical data sets (e.g., order entry, market data, matching engine logs) to troubleshoot or explain perceived issues. Execute SQL queries against a database to perform data analysis for customers and associates. Service and support historical data product requests.
On-Call & Weekend Testing: Participate in weekend testing (e.g., capacity testing, fail-over, etc.) and provide follow-the-sun on-call technical support as part of Cboe's global Operations team.
The ideal candidate has Bachelor's Degree: Computer Science, Computer Engineering, Software Engineering, or a related discipline.