Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Infrastructure and MLOps Engineer

£45.3 - £52.7 per hourEstimated

About Graphcore 

At Graphcore, we’re building the future of AI compute.

We’re a team of semiconductor, software and AI experts, with deep experience in creating the complete AI compute stack - from silicon and software to infrastructure at datacenter scale.

As part of the SoftBank Group, backed by significant long-term investment, we are delivering key technology into the fast-growing SoftBank AI ecosystem.To meet the vast and exciting AI opportunity, Graphcore is expanding its teams around the world.We are bringing together the brightest minds to solve the toughest problems, in a place where everyone has the opportunity to make an impact on the company, our products and the future of artificial intelligence .

Job Summary 

Join our dynamic Software Infrastructure team and take a pivotal role in scaling and managing our infrastructure. You will develop essential tools and services that empower our broader software team. Your contributions will enhance the build, test, deployment, and productisation processes of our Machine Learning Software components. Work with our High-Performance Computing (HPC) AI platforms and gain invaluable experience in distributed systems

The Team

The Software Infrastructure team provides critical platforms and services for software development teams across the business. Our responsibilities include managing the CI platform and services, build engineering, component integration, and packaging and release systems. We operate in squads, fostering a culture of service ownership and empowerment for our engineers. We focus on long-term engineering solutions and strive to eliminate toil wherever possible.  

Responsibilities and Duties  

  • Develop, own, and maintain tools and services to support AI research and engineering teams  
  • Deploy and maintain services with Kubernetes and Docker  
  • Manage our Cloud Infrastructure using tools such as Terraform  

Candidate Profile  

Essential:  

  • Knowledge of Python  
  • Familiarity with cloud services (e.g. AWS) 
  • Experience managing or developing in Linux environments  
  • Understanding of CI/CD principles  
  • Experience using Kubernetes (k8s) 
  • Experience of one of the following:

  • maintaining machine learning applications.

  • deploying ML orchestration tools (e.g. NV Ray, KFP, SkyPilot).

  • managing ML accelerator hardware (e.g. DCGM).

Desirable  

  • Experience with Infrastructure as Code (IaC) tools (e.g. Terraform/OpenTofu) 
  • Experience with GitHub Actions  
  • Experience with modern observability tooling (e.g. Prometheus) 
  • Experience with Grafana  
  • Knowledge of Go/Java/C++ (or similar language)

    Benefits

    In addition to a competitive salary, Graphcore offers flexible working, a generous annual leave policy, private medical insurance and health cash plan, a dental plan, pension (matched up to 5%), life assurance and income protection. We have a generous parental leave policy and an employee assistance programme (which includes health, mental wellbeing, and bereavement support). We offer a range of healthy food and snacks at our central Bristol office and have our own barista bar! We welcome people of different backgrounds and experiences; we’re committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments.

Vacancy posted a month ago
Similar jobs that could be interesting for youBased on the Infrastructure and MLOps Engineer in Cambridge, Cambridgeshire vacancy
  • £47k - £63k per annumEstimated
     ...and yield optimization, and real-time, low-latency ad serving. This is a senior, hands-on ML engineering role spanning modelling, large-scale training, and production infrastructure — you’ll take models from idea through to serving at scale, working in Spark, Python, and Java... 
    Suggested
    Hybrid working
    On-site
    Remote
    Monday to Thursday
    5 days/week
    Flexible hours

    Roku

    Cambridge, Cambridgeshire
    a month ago
  • £48k - £65k per annumEstimated
     ...join our multidisciplinary team of clinicians, scientists, and engineers. What unites us is our open culture, continuous learning...  ...and scalable engineering components. Support cloud-based ML infrastructure such as MLFlow. Create and maintain CI pipelines for model... 
    Suggested
    Full-time
    Immediate start

    Qureight Ltd

    Cambridge, Cambridgeshire
    a month ago
  • £61k - £79k per annumEstimated
     ...compute stack - from silicon and software to infrastructure at datacenter scale.As part of the...  ...Summary As a Senior Machine Learning Engineer in the Applied AI team at Graphcore, you...  ...Desirable: Experience in one or more of: MLOps for Kubernetes-based clusters... 
    Suggested
    Long-term contract
    Visa support
    On-site
    Flexible hours

    Graphcore

    Cambridge, Cambridgeshire
    a month ago
  • £55k - £80k per annum

     ...Machine Learning Engineer A fantastic opportunity for a Machine Learning Engineer to join a world-leading AI technology company developing cutting-edge software solutions. This is a specialist Python-focused engineering role with a strong Machine Learning and Large Language... 
    Suggested
    Full-time
    Hybrid working
    On-site

    RedTech Recruitment Ltd.

    Cambridge, Cambridgeshire
    5 days ago
  • £68k - £90k per annumEstimated
     ...We are hiring a Principal Machine Learning Engineer to work on cutting-edge R&D and translating research into customer-centric solutions....  ..., and experience with cloud platforms Strong familiarity with MLOps principles, CI/CD pipelines, and containerization (Docker, Kubernetes... 
    Suggested
    Long-term contract
    Internship
    Hybrid working
    On-site
    Remote
    Work from home
    Flexible hours
    3 days/week

    Speechmatics

    Cambridge, Cambridgeshire
    more than 2 months ago
  • $120k - $192.5k per annum

     ...over 100 groundbreaking companies since 2000, including Moderna.     The Role   FL105 seeks a talented  Machine Learning Research Engineer . The successful candidate will innovate, develop and apply machine learning (ML) methods to create foundational psychological tools... 

    Flagship Pioneering, Inc.

    Cambridge, Cambridgeshire
    a month ago
  • £48k - £63k per annumEstimated
     ...We're looking for an ML Data & Platform Engineer to own the infrastructure that powers our speech AI models: the pipelines that source and prepare training...  ..., and inference Continuously improving our data and MLOps practices, and helping shape the roadmap for how our... 
    Internship
    Hybrid working
    On-site
    Remote
    Work from home
    Flexible hours
    3 days/week

    Speechmatics

    Cambridge, Cambridgeshire
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Infrastructure and MLOps Engineer. Be the first to apply!