This position requires local presence. Please view similar jobs below.
Who We Are:
IMG is a leading global sports marketing agency, specializing in media rights management and sales, multi-channel content production and distribution, brand partnerships, strategic consulting, digital services, and event management. It powers growth of revenues, fanbases and IP for more than 250 federations, associations, events, and teams, including the National Football League, English Premier League, International Olympic Committee, National Hockey League, Major League Soccer, ATP and WTA Tours, the AELTC (Wimbledon), Euroleague Basketball, CONMEBOL, World Rugby, DP World Tour, and The R&A, as well as UFC, WWE, and PBR. IMG is a subsidiary of TKO Group Holdings, Inc. (NYSE: TKO), a premium sports and entertainment company.
TKO Group Holdings, Inc. (NYSE: TKO) is a premium sports and entertainment company. TKO owns iconic properties including UFC, the world’s premier mixed martial arts organization; WWE, the global leader in sports entertainment; and PBR, the world’s premier bull riding organization. Together, these properties reach 1 billion households across 210 countries and territories and organize more than 500 live events year-round, attracting more than three million fans. TKO also services and partners with major sports rights holders through IMG, an industry-leading global sports marketing agency; and On Location, a global leader in premium experiential hospitality.
Working Conditions
Permanent Position, Mon-Fri, 9am-5pm
This role is based at our facilities in Stockley Park, Uxbridge, with hybrid working options where applicable.
You may be required to work unsociable hours, including occasional weekends or on-call rotations, to support live operations and critical systems.
Occasional travel may be required depending on project and client needs.
IMG is looking for a Site Reliability Engineer to help design, build, operate, and continuously improve resilient, secure, and highly available platforms that underpin our digital, cloud, and broadcast-adjacent services. This role is suited to someone who combines strong infrastructure and software engineering capability with an operational mindset, and who can help embed reliability engineering practices across systems that support live, business-critical environments.
The successful candidate will play a key role in improving service reliability, observability, incident response, automation, and disaster recovery readiness across IMG platforms, while working closely with engineering, operations, and project stakeholders.
Key Responsibilities and Accountabilities
Design, build, and maintain reliable, scalable infrastructure and platform services across on-premises and cloud environments.
Improve service availability, latency, performance, and operational efficiency through engineering-led reliability practices.
Build and enhance observability across services and infrastructure, including monitoring, logging, alerting, dashboards, and service health indicators.
Define and maintain SLIs, SLOs, alerting standards, and operational runbooks for critical services.
Automate infrastructure provisioning, configuration, deployment, and recovery processes using Infrastructure as Code and scripting.
Partner with software, platform, broadcast engineering, and operational teams to improve release quality, resilience, and supportability.
Act as an escalation point for production incidents, leading or supporting rapid diagnosis, mitigation, communication, and post-incident follow-up.
Drive root cause analysis and corrective actions following incidents, with a focus on prevention and continuous improvement.
Support the design, testing, and documentation of high availability, backup, failover, and disaster recovery arrangements.
Help enforce security, access control, patching, and operational best practices across infrastructure and services.
Optimise system capacity, cost, and performance across environments.
Produce and maintain clear technical documentation, operational procedures, and support handover materials.
Support live event and critical operational workflows where reliability, rapid response, and stakeholder communication are essential.
Contribute to technical planning for new services, migrations, and platform enhancements, ensuring resilience is designed in from the start.
Improve reliability, stability, and recoverability of IMG’s platform services.
Reduced mean time to detect and resolve incidents through better observability and response processes.
Higher levels of automation across provisioning, deployment, remediation, and operational support.
Clearer operational ownership, documentation, and service standards across critical environments.
Stronger resilience for live and client-facing workflows through tested failover and recovery approaches.
Knowledge and Experience
Mandatory
Proven experience in a Site Reliability Engineer, DevOps Engineer, Platform Engineer, or similar role.
Strong knowledge of Linux and operating system fundamentals.
Strong hands-on experience with cloud platforms such as AWS, Azure, or Google Cloud.
Experience with containerisation and orchestration technologies such as Docker and Kubernetes.
Strong experience with CI/CD tooling and modern software delivery practices.
Hands-on experience with Infrastructure as Code tools such as Terraform or CloudFormation.
Experience with monitoring, logging, and alerting tooling, and with designing actionable observability solutions.
Solid understanding of networking, security, system architecture, and distributed systems principles.
Strong scripting or programming capability in Python, Bash, or similar languages.
Experience working in high-availability, live production, or other business-critical operational environments.
Strong troubleshooting skills, calm decision-making under pressure, and a continuous improvement mindset.
Excellent communication and collaboration skills, including the ability to work effectively with technical and non-technical stakeholders.
Desirable
Experience supporting media, broadcast, streaming, or live event platforms.
Familiarity with incident management, postmortem practice, and error-budget based operational models.
Experience with resilience engineering, multi-site failover, and disaster recovery testing.
Exposure to event-driven or low-latency systems, media transport, or hybrid on-prem/cloud architectures.
Understanding of compliance, operational risk management, and support processes in client-facing environments.
Personal Attributes
Proactive and ownership driven.
Methodical, analytical, and detail oriented.
Comfortable operating in fast-moving, high-pressure environments.
Pragmatic in balancing engineering excellence with operational needs.
Collaborative, service oriented, and committed to raising reliability standards across teams.
In addition, success at IMG is driven by four core competencies that apply to all employees:
Business Acumen – Understanding financial drivers, interpreting business data, aligning decisions to strategic outcomes
Operational Excellence – Driving efficiency, governance, and continuous improvement in delivery & operations
Innovation Mindset – Cultivating curiosity, experimentation, and forward-looking capability development
Leadership & Collaboration – Inspiring others, building trust, and enabling collaboration across teams
TKO EEO Statement
TKO is an Equal Opportunity Employer and complies with all applicable federal, state, and local laws regarding non-discrimination in employment. TKO makes employment decisions based on merit and qualifications, without considering an employee’s or applicant’s race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, disability, marital status, veteran status, or any other basis prohibited under federal or local laws governing non-discrimination in employment in every location in which the Company has facilities. TKO also provides reasonable accommodations for qualified individuals with disabilities in accordance with the Americans with Disabilities Act (ADA) and applicable state or local laws. For information about Privacy and Information Security for TKO employment candidates, please review our Privacy Policy. For information regarding Terms of Use for this and other TKO websites, please review our Terms of Use.
TKO EEO Statement:
TKO is an Equal Opportunity Employer and complies with all applicable federal, state, and local laws regarding non-discrimination in employment. TKO makes employment decisions based on merit and qualifications, without considering an employee’s or applicant’s race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, disability, marital status, veteran status, or any other basis prohibited under federal, state or local laws governing non-discrimination in employment in every location in which the Company has facilities. TKO also provides reasonable accommodations for qualified individuals with disabilities in accordance with the Americans with Disabilities Act (ADA) and applicable state or local laws. For information about Privacy and Information Security for TKO employment candidates, please review our Privacy Policy . For information regarding Terms of Use for this and other TKO websites, please review our Terms of Use.
- £121k - £164k per annumEstimated...Site Reliability Engineer – Fintech Quant Capital is urgently looking for a Site Reliability Engineer to join or well known Fintech50 client who produces software disrupting the wealth management market. My client is a market leading SAAS provider to financial advisory...Suggested
£90k - £120k per annum
...your decisions Ability to focus on what matters most, manage your time, and get things done To be a team player, ready to help engineers investigate issues or teach them new things Love to automate manual work and try new modern technology/approaches What we...SuggestedRelocation packageVisa sponsorshipOn-siteRemoteFlexible hours1 day/week- £52k - £68k per annumEstimatedSite Reliability Engineer Position Description We are seeking an experienced and proactive Site Reliability Engineer (SRE) to join a team supporting multiple data product and platform groups. This role is focused on improving the reliability, scalability, observability...Suggested5 days/week
- £42k - £54k per annumEstimated...obsessed about achieving the high quality and reliability our customers demand. You will work... ...not only with your peers, but also the RTO engineering teams, allowing your technical... ...and securely. Exemplify cloud-native site reliability best practices. Write code...SuggestedLong-term contractFlexible hours
- £97k - £127k per annumEstimated...leading technology-driven trading firm where engineering, automation, and high-performance... ...engineering practices, enabling teams to build reliable, scalable, and highly automated... ...environments. This opportunity is ideal for a Site Reliability Engineer looking to work at the...SuggestedPermanentFlexible hours
- £65k - £86k per annumEstimated...holders and more than 40 million verified users, facilitating over $1 trillion in crypto transactions. We are looking for a Site Reliability Engineer to join our Core team to encourage infrastructure best practices across our organization that would allow to securely...Full-timeApprenticeshipOn-siteRemoteFlexible hours
- £59k - £78k per annumEstimated...Keep our global cloud platform reliable, scalable, and resilient. Do you enjoy solving complex... ...turn into incidents? Are you the kind of engineer who automates repetitive tasks, improves... ...management? We're looking for a Senior Site Reliability Engineer to join the Global Platform...Long-term contractPermanentFull-timeHybrid workingOn-site
- £63k - £84k per annumEstimated...managers representing $40 trillion in AUM . For more information, please visit . The Role CMG is looking for a Site Reliability Engineer (SRE) with a strong focus on monitoring, observability, and alerting to ensure the reliability, performance, and...Full-time
£60k - £70k per annum
Company: SPECTRUM IT RECRUITMENT Job Type: Permanent, Full Time Salary: £60000 - £70000/annum bonus, medical carePermanentFull-time- £60k - £80k per annumEstimated...Senior Site Reliability Engineer Are you passionate about building resilient, scalable systems that power mission-critical applications? Do you thrive on automating operations, improving reliability, and ensuring exceptional system performance? About the team Embedded...Long-term contractFull-timeImmediate startRemoteFlexible hoursRotating shifts
- £73k - £96k per annumEstimated...of both work styles in a workplace that is intentional about belonging, collaboration, and accomplishment. Being a Senior Site Reliability Engineer at iManage Means… You are an engineer, a builder, and a systems thinker. You’ll create middleware and platform guardrails...Full-timeHybrid workingOn-siteRemoteMonday to FridayFlexible hours
- £38k - £51k per annumEstimated...nation states, armed forces and commercial businesses can unlock digital advantage in the most demanding environments. Site Reliability Engineering is a rapidly growing concept in industry, with a remit to drive the quality, reliability and performance of essential systems...Hybrid workingOn-siteRemoteRotating shifts
£250k - £300k per annum
...mutual respect and a flat structure. This role sits in the Shared Engineering team that focuses on designing, developing, and maintaining infrastructure and tools. The team requires a Network Site Reliability Engineer (SRE) with strong network fundamentals, problem-solving...On-site- £64k - £84k per annumEstimated...to manage platform infrastructure and applications Improve reliability, quality, and time-to-market of our suite of software solutions... ...improve Provide primary operational support and engineering for multiple large distributed software applications How...Hybrid workingOn-siteRemote
£120k per annum
~ Role: Site Reliability Engineer – Trading Client: Elite FinTech Compensation: £70,000 - £120,000 + Bonus Location: London Overview My client are seeking an SRE / Linux Platform Engineer to work on their low latency Linux estate. The role is a cross between...PermanentImmediate start- £60k - £79k per annumEstimated...We're looking for a Senior Site Reliability Engineer – Observability to join our team in London, United Kingdom in a hybrid working mode. You will be part of the Production Engineering – Observability team, driving the strategic initiative to implement and expand a modern...Long-term contractHybrid working
£90k per annum
...technical excellence, the organisation continues to push the boundaries of low-latency infrastructure and reliable system design. The team is hiring a Site Reliability Engineer (London) to build, monitor, and optimise mission-critical trading systems. The role will focus on...Permanent- £69k - £93k per annumEstimated...Capital On Tap, we run a hybrid embedded SRE model - We aim to work closely with the teams to provide them the best support. As a Site Reliability Engineer (SRE) you will help ensure our platforms are fast, reliable, and scalable. You’ll design, build, and monitor systems,...Hybrid workingOn-site
- £53k - £68k per annumEstimated...complex, mission-critical systems which help our clients keep us all safe and secure. We are currently looking for an experienced site reliability engineers to join our cross-functional team who, in partnership with our clients, will help define, guide and assure the delivery of...Full-time5 days/week
£62.4k - £93.6k per annum
...LinkedIn Mollie is the leading payments and financial services partner for business, rooted in Europe, with global reach. Platform Engineering at GoCardless The Platform Engineering team is a unified, globally distributed team. We are currently located in London, Riga,...£150k per annum
...Site Reliability Engineer Job Opportunity Role: Site Reliability Engineer Client: Most Elite FinTech Firm in London Compensation: Up to £150k + Bonus + Package Location: Montreal Overview An Elite FinTech Firm is looking for a highly talented DevOps...PermanentOn-siteFlexible hours£90k per annum
...Site Reliability Engineer Trading £90,000 Plus Bonus Quant Capital is urgently looking for a Site Reliability Engineer to join our high profile client. Our client has recently been voted in the fintech 100 (best financial technology businesses). They provide...Immediate startFlexible hours- £48k - £64k per annumEstimated...is to gain as much visibility into our platform's health while offering scalable solutions. What You’ll be doing: As a Site Reliability Engineer (SRE) you will help ensure our platforms are fast, reliable, and scalable. You’ll design, build, and monitor systems,...Hybrid workingOn-site
£80k - £100k per annum
Site Reliability Engineer with Python Our Client looking to bring on a site reliability engineer to help deploy, manage, troubleshoot, and enhance our complex cloud-based set of internal tools and externally managed services for a variety of users across our wide-ranging...Full-time- £60k - £81k per annumEstimated...a globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within the Commercial and Investment Banking, Markets Technology – Athena Core team, you hold...Long-term contract
- £44k - £58k per annumEstimated...As a critical and trusted member of the Systems Engineering team, you'll be working side-by-side with software engineers to design and deliver mission critical services and systems. You'll be working with infrastructure and services at scale, utilizing wide variety of cutting...FreelanceOn-site
£107k per annum
...becoming the fabric of everyday healthcare through the lifetime - from birth to old age. Job Description Principal Site Reliability Engineer As Principal Site Reliability Engineer, you will help establish and grow Genomics England's organisation-wide SRE...Long-term contractPermanentFull-timePart-timeFixed-term contractHybrid workingOn-siteImmediate startRemoteFlexible hours- £87k - £114k per annumEstimated...We're looking for a Director of Site Reliability Engineering to join our team in London, United Kingdom in a hybrid working mode. This role is responsible for driving reliability engineering and operational excellence across global technology platforms while leading the...Hybrid working
£90k per annum
...Site Reliability Engineer – Fintech / Linux £90,000 – bonus and benefits Quant Capital is urgently looking for a Site Reliability Engineer to join our high profile client. Our client is a major global financial exchange, driven by technology. They are at the...- £73k - £97k per annumEstimated...Overview Goldman Sachs has embarked on one of its most ambitious engineering programs: the Consolidated Trade Ledger (CTL), a ground-up... ...significant operational efficiencies. We are seeking a Site Reliability Engineer to help build, run and continuously improve this...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer, Studios. Be the first to apply!
- senior site reliability engineer London
- site reliability engineer London
- cloud site reliability engineer London
- lead site reliability engineer London
- yoga studio London
- studio London
- photographic studio London
- part time receptionist required for busy yoga studio London
- studio internship London
- studio coordinator London
