What you’ll do
In his role, within VodafoneThree's Performance and Chaos Engineering (PaCE) team, you will play a key role in enabling engineering teams to deliver scalable, resilient, and high-performing digital services. PaCE is responsible for embedding performance, resilience, and reliability practices throughout the software development lifecycle, helping teams identify and address risks early and build confidence in their solutions before they reach production.
Working closely with product, platform, development, and Site Reliability Engineering teams, you will champion a shift-left approach to performance and resilience engineering. Your focus will be on providing the frameworks, tooling, standards, and automation that enable teams to validate scalability, reliability, and operational readiness from the earliest stages of design and development.
Your expertise in software engineering, performance testing, observability, and automation will help drive the adoption of engineering best practices across the organisation. You will lead initiatives to integrate performance and resilience validation into CI/CD pipelines, establish meaningful service level objectives (SLOs), and develop self-service capabilities that empower teams to identify bottlenecks, validate system behaviour under load, and assess failure scenarios independently.
Through collaboration, coaching, and continuous improvement, you will help cultivate a culture where performance, resilience, and operational excellence are shared responsibilities. By leveraging data-driven insights, experimentation, and automation, you will enable teams to deliver services that are reliable, scalable, and capable of supporting VodafoneThree's evolving customer and business demands.
Main responsibilities:
-
You will drive shift-left performance and resilience engineering by embedding non-functional testing, reliability, and scalability validation throughout the software development lifecycle.
-
You will integrate automated performance and resilience validation into CI/CD pipelines, ensuring risks are identified and addressed before production deployment.
-
Be part of enabling engineering teams through self-service capabilities, providing tooling, frameworks, standards, and guidance for performance, resilience, and chaos testing.
-
Support in defining and championing engineering best practices, including SLOs, observability, capacity planning, and operational readiness to improve service reliability and customer experience.
-
You will collaborate closely Product, Engineering, Platform, Architecture, and SRE teams to design resilient, scalable solutions and proactively identify performance and reliability risks.
-
Drive continuous improvement through data-driven insights, experimentation, and chaos engineering, helping teams build confidence in system behaviour under load and failure conditions.
-
Alongside with the team, you will shape the development of an in-house framework, used daily by hundreds of engineers across the organization