Job Title: Infrastructure Software Engineer
We're seeking a talented Infrastructure Software Engineer to join our innovative startup, The Innovation Game (TIG). Our mission is to revolutionise scientific research by creating a market-based framework that accelerates the development of computational methods crucial to data-driven sciences, creating a sustainable ecosystem for open collaboration in scientific research.
About the Role:
You'll play a key role in building, maintaining, and improving the infrastructure that powers TIG. You'll take ownership of deployment systems, monitoring, reliability, and operational resilience, ensuring our infrastructure remains secure, performant, and scalable as we grow.
This role is ideal for someone who enjoys solving practical engineering challenges, automating repetitive work, improving system reliability, and responding calmly and effectively when things go wrong. You should be comfortable operating production systems and willing to work across a broad range of infrastructure and operational responsibilities in a startup environment.
Key Responsibilities: - Develop and improve monitoring, alerting, and observability across TIG's infrastructure.
- Act as a first responder to production incidents, attacks, outages, and unexpected infrastructure failures.
- Manage and maintain deployment infrastructure, remote servers, databases, networking configuration, and cloud services.
- Maintain and improve operational documentation, runbooks, and internal infrastructure processes.
- Automate operational workflows and develop scripts to improve reliability and engineering productivity.
- Contribute to infrastructure architecture decisions and help establish best practices around security, deployment, and operational resilience.
- Identify and implement infrastructure optimisations, including caching strategies, deployment improvements, and performance tuning.
Qualifications:
- Strong programming ability, with the judgement to use software engineering to solve infrastructure and operational problems.
- Strong Linux administration skills, particularly Ubuntu and command-line environments.
- Experience operating and maintaining production systems.
- Experience managing deployment infrastructure, remote servers, databases, reverse proxies, DNS, and infrastructure services such as Cloudflare.
- Experience working with Docker and docker-compose.
- Ability to troubleshoot infrastructure issues
- Ability to communicate clearly on technical matters.
- Verbal and written English language fluency.
- Comfortable working independently and taking ownership in a fast-moving startup environment.
Nice to Have:
- Experience with Rust.
- Experience with monitoring and observability tooling (e.g. Prometheus, Grafana, OpenTelemetry).
- Experience with CI/CD systems and deployment automation.
- Familiarity with infrastructure-as-code tooling.
- Security-focused mindset and experience hardening production systems.
We're looking for a self-motivated individual who can thrive in a startup environment, adapt to changing priorities, and contribute across various areas as needed.
We value strong problem-solvers above specialists. The ideal candidate enjoys using programming as a tool to solve real-world engineering problems - automating repetitive work, building internal tools, and improving systems rather than relying on manual processes.
We're interested in people who are curious, learn quickly, and can reason through unfamiliar technical challenges, even if they haven't worked with every technology in our stack.
Location: Cambridge, UK.
To Apply: Send your CV and cover letter by email to [email protected]
Pay: £40,000.00-£60,000.00 per year
Work Location: In person