Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff Storage Platform Engineer (AI Storage) - Radian Arc

£30.6 - £35.9 per hourEstimated
United Kingdom
  • Remote job

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Staff Storage Platform Engineer (AI Storage) - Radian Arc based in United Kingdom.

This is a Staff-level opportunity to shape the storage architecture powering large-scale GPU and AI infrastructure across edge and core environments.
You will design, build, and operate high-performance storage systems supporting training, fine-tuning, and distributed inference workloads.
The role covers hyperconverged, local NVMe, and disaggregated storage architectures, with a strong focus on throughput, latency, resilience, and cost efficiency.
You will work at the intersection of storage, GPUs, networking, Kubernetes, and AI platform engineering, ensuring storage never becomes a bottleneck for compute.
As the primary storage specialist, you will combine architectural ownership with hands-on engineering, troubleshooting, performance optimization, and deployment.
You will also define reusable standards, influence long-term platform direction, and mentor engineers across adjacent infrastructure domains.
This is an ideal environment for a senior storage expert who wants significant technical ownership and direct impact on next-generation AI infrastructure.

Accountabilities
  • Storage architecture: Design scalable storage architectures for edge and core GPU deployments, covering hyperconverged platforms such as StorPool, local NVMe, and disaggregated systems such as VAST Data and Weka. Define reference architectures, reusable design patterns, fault domains, lifecycle strategies, and scaling approaches while balancing throughput, latency, resilience, data locality, operability, and cost.
  • AI workload optimization: Optimize storage for distributed training, fine-tuning, and inference workloads, including large dataset ingestion, model artifact distribution, checkpointing, and high-concurrency access. Establish realistic performance baselines and ensure storage architecture aligns with actual GPU workload behavior.
  • Distributed inference: Design storage architectures supporting inference platforms such as NVIDIA Dynamo, llm-d, or similar systems. Optimize model distribution, token-generation data paths, and KV-cache persistence and retrieval so infrastructure can scale efficiently across large GPU clusters without storage becoming a throughput or latency bottleneck.
  • High-performance data paths: Engineer efficient storage-to-GPU data paths using technologies such as GPU Direct Storage, RDMA/RoCE, NVMe-oF, and SPDK. Investigate and tune performance across hardware, networking, operating systems, filesystems, storage layers, and distributed workloads.
  • Platform integration: Integrate block, object, and shared file storage into Kubernetes and platform orchestration systems. Implement and maintain CSI integrations, support multi-tenant storage architectures, and define standards for storage integration across different deployment models.
  • Distributed storage: Contribute to large-scale storage platforms, including S3-compatible object storage, distributed file systems, and block storage. Design systems with clear operational boundaries, resilience models, scaling paths, and reusable operating patterns across multi-cluster and multi-site environments.
  • Performance and reliability: Lead storage benchmarking, capacity planning, performance investigations, incident response, and root-cause analysis. Establish measurable standards for throughput, latency consistency, recovery behavior, reliability, and operational maturity.
  • Engineering delivery: Own storage initiatives end to end, from architecture and validation through production rollout. Validate BOMs, topology decisions, node profiles, and deployment assumptions while ensuring changes are introduced safely with minimal customer impact.
  • Operational excellence: Improve storage observability, automation, runbooks, lifecycle management, and day-2 operations. Turn recurring incidents and operational pain points into durable engineering improvements and standardized practices.
  • Technical leadership: Act as the primary storage design authority, influencing platform architecture and roadmap decisions across compute, networking, DevOps, infrastructure, and operations. Communicate technical trade-offs clearly, mentor adjacent engineers, and raise the organization’s expertise in AI storage.
  • Requirements

    • Distributed storage expertise: Strong hands-on experience designing and operating distributed storage systems in high-performance computing, AI, GPU, or similarly demanding environments.
    • AI infrastructure experience: Proven experience designing storage architectures for large-scale AI training, fine-tuning, or inference, including dataset distribution, model artifacts, checkpointing, and high-concurrency data access.
    • AI storage knowledge: Deep understanding of how AI workload characteristics affect storage throughput, latency, concurrency, data locality, checkpoint recovery, and serving performance.
    • Storage technologies: Hands-on experience with technologies such as Weka, VAST Data, StorPool, local NVMe, distributed filesystems, S3-compatible object storage, block storage, and/or comparable enterprise storage platforms.
    • Linux and systems expertise: Strong knowledge of the Linux storage and I/O stack, storage hardware, NVMe devices, storage fabrics, and high-performance data paths.
    • Kubernetes: Familiarity with Kubernetes storage integrations, particularly CSI, and experience integrating storage into containerized or orchestrated platforms.
    • AI data paths: Practical knowledge of GPU Direct Storage, RDMA/RoCE, NVMe-oF, SPDK, and techniques for minimizing unnecessary data movement between storage and GPU compute.
    • Distributed inference: Experience with storage requirements for inference orchestration and model-serving environments, including model distribution and KV-cache persistence or retrieval, is highly valuable.
    • Troubleshooting: Ability to diagnose complex cross-layer issues involving storage hardware, networking, Linux kernels and I/O paths, filesystems, object/block storage, Kubernetes, and distributed workloads.
    • Automation: Strong Python and/or Bash skills, with experience applying software engineering practices to infrastructure automation, validation, lifecycle management, and operational tooling.
    • Observability: Experience designing or operating storage observability systems and using metrics and telemetry to identify performance, reliability, and capacity issues.
    • Technical leadership: Demonstrated ability to lead complex infrastructure initiatives, establish architectural standards, and influence multiple teams without relying on formal management authority.
    • Systems thinking: Ability to balance performance, scalability, reliability, operability, deployment complexity, and cost when making architecture decisions.
    • Communication and collaboration: Comfortable working with compute, networking, platform, DevOps, operations, deployment teams, vendors, and other technical stakeholders.
    • Ownership: Able to combine Staff-level strategic thinking with hands-on execution, particularly in a lean or fast-scaling environment where processes and standards are still being established.
    • Mentoring: Strong ability to share knowledge, guide engineers in adjacent domains, and raise the technical bar across the broader infrastructure organization.
    • Benefits

      • Attractive compensation package aligned with your expertise and experience.
      • Opportunity to play a foundational role in shaping a next-generation AI storage platform.
      • Significant architectural ownership and direct influence over long-term infrastructure strategy.
      • Exposure to cutting-edge GPU, AI inference, distributed storage, and high-performance data technologies.
      • International and diverse working environment with strong flexibility.
      • Remote-friendly work model across Europe.
      • Opportunity to join a fast-growing scale-up with an ambitious technology mission.
      • Broad cross-functional exposure across infrastructure, compute, networking, platform engineering, and operations.
      • Strong career growth potential as the infrastructure organization expands.
      • Inclusive environment committed to equal opportunity and professional development.
Vacancy posted 14 hours ago
Similar jobs that could be interesting for youBased on the Staff Storage Platform Engineer (AI Storage) - Radian Arc in United Kingdom vacancy
  • £79k - £107k per annumEstimated
     ...behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Staff Network Engineer (AI Fabric, Datacenter and Edge Networking) - Radian Arc based in United Kingdom. This is a high-impact Staff-level networking role responsible for... 
    Suggested
    Remote job
    Long-term contract
    Full-time
    Flexible hours
    United Kingdom
    14 hours ago
  • £46k - £61k per annumEstimated
     ...About Nscale Nscale is the GPU cloud engineered for AI. We provide cost-effective, high-performance infrastructure for AI start-ups...  ...that powers the future. About the Role We’re hiring a Staff Engineer, Storage Services to provide senior technical leadership across... 
    Suggested
    Long-term contract
    Flexible hours

    Nscale

    United Kingdom
    more than 2 months ago
  • £61k - £82k per annumEstimated
     ...Runware is building the API layer for the next generation of AI products. Our platform gives teams fast, reliable access to real-time...  ...Performance matters at every layer. We are looking for a Staff/Senior DevOps Engineer to help build, operate, and scale the infrastructure... 
    Suggested
    Remote
    Work from home
    Flexible hours

    Software Engineering

    United Kingdom
    more than 2 months ago
  • £106k - £141k per annumEstimated
     ...system for mobile services—a platform that lets tech companies embed...  ...bringing together early-stage engineers, product builders, and business...  ...senior IC track, hiring multiple staff and principal engineers,...  ...ships software comes from weaving AI into the platform itself. The... 
    Suggested
    On-site
    Remote
    Work from home

    Software Engineering, Data Science

    London
    a month ago
  • £151k - £246k per annum

     ...We’re looking for a curious, rigorous, problem-hungry platform engineer (who codes!) to carry the ball as we bring Ashby to the big leagues. Ashby...  ...component cascades throughout our app (short video below). AI-powered tooling. We think of AI as a way to automate the... 
    Suggested
    Full-time
    On-site

    Ashby

    Edinburgh
    7 days ago
  •  ...- we do this by bringing together our AI Augmented MSP platform, hybrid sovereign cloud solutions, deep...  ..., we have a requirement to engage an Engineer to support the operational running of...  ...virtualisation platforms, enterprise storage services, and backup and disaster recovery... 
    Hybrid working
    On-site

    Node4

    United Kingdom
    27 days ago
  • £90k - £120k per annum

     ...Senior Platform Engineer AWS | Platform Engineering | Cloud Infrastructure | Distributed Systems...  ...service ownership Working across compute, storage, databases, networking, messaging and...  ...Machine learning infrastructure LLM or AI-related workloads Building internal platforms... 
    Long-term contract
    Permanent
    Temporary
    Hybrid working

    SoCode Recruitment

    London
    14 days ago
  • £50k - £70k per annum

     ...Payment Facilitator (PayFac)/Acquiring platform on AWS and is building a dedicated platform engineering capability from the ground up. We...  ...— networking, compute, and storage — using infrastructure as code...  ...processes • Support deployment of AI generated applications and... 
    Live-in
    Hybrid working
    On-site
    Remote

    CreatePay

    Milton Keynes, Buckinghamshire
    3 days ago
  • £150k per annum

     ...the financial backbone of the AI and digital infrastructure revolution...  ...data, pricing, and execution platforms that enable enterprises and...  ...around cloud infrastructure, storage, CI/CD, and monitoring...  ...systems at scale Strong backend engineering experience with Python and/or... 
    Permanent
    Remote
    London
    more than 2 months ago
  • £60k - £80k per annum

     ...Software Engineer (Application platforms/AWS/Python) London (On-site, London Farringdon office) Full-time...  ...that powers the next generation of AI products? Do you thrive on creating reusable...  ...scalable systems for data ingestion, storage, and processing, supporting multi-... 
    Full-time
    On-site

    Travtus

    London
    more than 2 months ago
  • £58k - £77k per annumEstimated
     ...every aspect of the firm's hybrid platform. This ranges from our EUC/...  ...Role Overview:  The Platform Engineer, Network Services focus will...  ...core platforms, including Linux, Storage, Windows and Hypervisors. As a...  ...Infrastructure as Code, AI-enabled operations, automation... 
    Hybrid working
    On-site
    Flexible hours

    BlueCrest Capital Management

    London
    2 days ago
  • £43k - £58k per annumEstimated
     ...Senior Platform Engineer ~202604318 ~Reigate, England, United Kingdom ~Full time View...  ...services (e.g., Container Apps, AKS, Azure Storage, Azure SQL/Cosmos DB, Key Vault, Azure...  .... Strong awareness of emerging cloud, AI, DevOps and platform technologies, with an... 
    Full-time
    Hybrid working

    WTW

    Reigate, Surrey
    2 days ago
  • £64k - £86k per annumEstimated
     ...Helical is the AI-native lab for biology. We turn biological foundation models into...  .... The Role We’re hiring a Platform Engineer to build and scale the infrastructure behind...  ...observability, and deployments Handle databases, storage, and migrations Build automation (... 
    Full-time

    Helical

    London
    a month ago
  •  ...The AI-powered OS for beauty, wellness and self-care About...  ...professionals use an all-in-one platform to manage their entire...  ...sustainable. As a Senior Platform Engineer, you will take end-to-end...  ...PostgreSQL (RDS) for persistent storage GitHub Actions (primary CI)... 
    On-site
    Work from home

    Fresha

    London
    19 hours ago
  • £89k - £118k per annumEstimated
     ...operational problems that slow this work down. Our AI platform captures the context behind complex...  .... Overview We’re hiring a Platform Engineer to help design, build, and maintain the...  ...networks, operating systems, compilers, storage systems, or distributed services.... 
    Full-time
    Hybrid working
    On-site
    Flexible hours

    Cogna

    London
    a month ago
  • £70k - £92k per annumEstimated
     ...leading developer of Embodied AI technology.  Our advanced AI software...  ...opportunity to be a founding Staff SRE shaping the reliability of...  ...a Staff Cloud Site Reliability Engineer at Wayve, you will build and...  ...reliability foundations of our AI cloud platform. This includes our Model... 
    Hourly pay
    Full-time
    Hybrid working
    On-site
    Work from home

    Wayve

    London
    more than 2 months ago
  • £64k - £86k per annumEstimated
     ...science, machine learning, and AI are core components of its...  ...improvements. This is an engineering role, not an analytical one. You...  ...up the internal infrastructure platform. Working with the engineering...  ...and scaling the compute and storage the platform provisions across... 

    Intellias

    United Kingdom
    2 days ago
  •  ...The AI-powered OS for beauty, wellness and self-care About...  ...professionals use an all-in-one platform to manage their entire...  ...sustainable. As a Platform Engineering Team Lead, you will be accountable...  ...PostgreSQL (RDS) for persistent storage GitHub Actions (primary CI)... 

    Fresha

    London
    a month ago
  • £48k - £63k per annumEstimated
    AI Platform Engineer We are Lightsource bp – and we’re on a mission to become a global leader in onshore renewables, anchored by our proven...  ..., reliable, large-scale onshore renewable and energy storage solutions to help the world decarbonise.   Our growing business... 
    Long-term contract
    Permanent
    On-site

    Lightsource BP Renewable Energy Investments Limited

    London
    9 days ago
  • £98k - £127k per annumEstimated
     ...for We are seeking an experienced Cloud Platform Engineering Lead with deep AWS expertise to join...  ...on building, operating and securing an AI-first, agentic-native AWS cloud platform...  ...services including VPC networking, compute, storage, containers (ECS/EKS), and serverless (Lambda... 
    Full-time

    Schroders

    London
    10 days ago
  • £75k - £95k per annum

     ...Who We Are Lightning AI is the company behind PyTorch Lightning...  ...2019, we build an end-to-end platform for developing, training, and...  ...hire AI Platform Support Engineers to join our EMEA Customer...  ...latency, networking bottlenecks, storage performance, and platform reliability... 
    Long-term contract
    Hybrid working
    On-site
    Work from home
    Flexible hours
    Shift work

    Lightning AI

    London
    10 days ago
  • £98k - £128k per annumEstimated
     ...About the job Build the cloud platform that turns Graphcore AI systems into usable, scalable services. Graphcore needs a Bristol-based Staff Cloud Engineer to develop and deploy private and public cloud services for AI compute. You will join the Cloud Platform team within... 
    Visa support
    On-site
    Flexible hours

    Graphcore

    London
    22 days ago
  • £80k - £103k per annumEstimated
     ...We’re looking for a Platform Engineering Manager/Consultant to join our team in London, United Kingdom...  ...scalable, client-centric data and AI platform solutions. You will serve as a...  ...products Guide teams in leveraging AWS storage and compute services for platform optimization... 
    Hybrid working

    EPAM Systems

    London
    a month ago
  • £5k per annum

     .... We're looking for a passionate Data Platform Engineer to join the Alto Team. You'll help us scale...  ...data lakehouse that powers analytics and AI across the business, build the...  ...consistently, from ingestion through to storage and access. Play a key role in modernising... 
    Full-time
    Hybrid working

    Alto

    London
    a month ago
  • £51k - £68k per annumEstimated
     ...what’s on the market.    The Role:  Platform Engineering Manager (Cloud Foundations) Location:...  ...healthy on‑call culture.   ~ Automation (AI‑driven or traditional) delivers...  ...services including networking, compute, storage, containerisation, security and IAM.... 
    Full-time
    Hybrid working
    On-site
    Work from home

    Rightmove Careers

    London
    more than 2 months ago
  • £56k - £75k per annumEstimated
     ...applications and next steps. Our partner is looking for a Staff Network Engineer (AI Fabric, Datacenter and Edge Networking) based in United Kingdom...  ...function. You’ll collaborate closely with compute, platform, storage, SRE, observability, operations, and datacenter teams... 
    Remote job
    Long-term contract
    Hybrid working
    Immediate start
    Flexible hours
    United Kingdom
    14 hours ago
  • £81k - £107k per annumEstimated
     ...foundations that support the firm’s AI and analytics capabilities. This role sits within the engineering effort to develop a modern Lakehouse and AI data platform that enables reliable, well-...  ...JSON, Avro, Parquet Platforms and storage: Snowflake, Apache Iceberg, Databricks... 
    Long-term contract

    Goldman Sachs

    London
    more than 2 months ago
  • £105k - £141k per annumEstimated
     ...Role Overview: As a Senior / Staff Data Scientist, you will help...  ...closely with product, engineering, analytics and business teams...  ...statistical, machine learning and AI techniques to solve complex real...  ...PlayStation's experimentation platform and measurement capabilities,... 
    Long-term contract
    Hybrid working

    PlayStation Global

    London
    a month ago
  • £65k - £83k per annumEstimated
     ...Job title: Senior   Engineering Manager (Data Platform & Analytics Engineering) Location: London, Bristol...  ...electrified future. Leveraging Data, AI, and real-time decision-making we turn...  ...data ingestion, data cataloguing to data storage and exposing data externally. What... 
    Hybrid working
    On-site
    Flexible hours
    Shift work

    Kaluza

    Edinburgh
    13 hours ago
  • £25 per hour

     ...commitments. This role is known as Senior Power Platform Developer internally. What part will you play?...  ...Copilot - Improving productivity, collaboration and AI-assisted ways of working. SQL / Azure SQL - Data storage, reporting datasets, integrations and structured... 
    Permanent
    Full-time
    On-site
    Monday to Friday

    ICE - Industrial Cleaning Equipment

    Totton, Hampshire
    21 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff Storage Platform Engineer (AI Storage) - Radian Arc. Be the first to apply!