Ellison Institute of Technology logo

HPC Engineer - Generative Biology Institute

Ellison Institute of TechnologyOxford, United KingdomPosted May 28hybrid
via Workable

At the Ellison Institute of Technology (EIT), we’re on a mission to translate scientific discovery into real world impact. We bring together visionary scientists, technologists, engineers, researchers, educators and innovators to tackle humanity’s greatest challenges in four transformative areas:

  • Health, Medical Science & Generative Biology
  • Food Security & Sustainable Agriculture
  • Climate Change & Managing CO₂
  • Artificial Intelligence & Robotics

This is ambitious work - work that demands curiosity, courage, and a relentless drive to make a difference. At EIT, you’ll join a community built on excellence, innovation, tenacity, trust, and collaboration, where bold ideas become real-world breakthroughs. Together, we push boundaries, embrace complexity, and create solutions to scale ideas from lab to society. Explore more at www.eit.org.

Welcome to the Generative Biology Institute:

Led by Founding Director Jason Chin, the Generative Biology Institute (GBI) at the Ellison Institute of Technology is tackling the key challenges in making biology engineerable, and thereby unlocking the unrivalled power of biology for the benefit of humanity.

The vision of the GBI is to lay the foundations for engineering biology, and unlock its potential for good. To achieve this, we must overcome two key challenges. First, we need the ability to write in the natural language of biology, enabling the rapid and scalable synthesis of entire genomes with precision. Second, we must understand what to write - determining which DNA sequences will generate biological systems that perform the desired functions. Addressing these challenges will allow us to harness the full power of biology to create transformative solutions across health, agriculture, clean energy and more.   

GBI will have sustained and substantial funding to support the unique scale and ambition of its ground-breaking vision for engineering biology. GBI researchers will also be supported by cutting-edge technology hubs including mass spectrometry, flow cytometry, sequencing, automation, imaging, and bioprocessing. GBI will also have access to substantial compute resources that can be leveraged to further accelerate progress, including scientific compute, bioinformatics, and machine learning. The environment at GBI will allow researchers to undertake ambitious, long-term, collaborative research, and we will actively support the translation of research to commercial applications, where appropriate.  

The Generative Biology Institute commenced operations in 2025, occupying newly renovated bespoke space in the Oxford Science Park. The team will later move to a purpose-made facility in the Oxford Science Park, currently under construction. Once complete, this state-of-the-art facility will include more than 40,000 m² of research laboratory and office space.  It will house over 30 groups and up to 600 employees at scale, focused on solving the two critical challenges in making biology engineerable and applying the solutions to addressing the global challenges encapsulated in EIT’s Humane Endeavors.  

Your Role:

Working as part of a new Scientific Computing team within GBI, the HPC Engineer will help operate, improve, and scale the data and computing platform that will enable cutting-edge research in engineering biology. This is a broad, hands-on role at the interface of Linux systems, high-performance computing, cloud infrastructure, Kubernetes, Slurm, storage, monitoring, and researcher support. They will help turn emerging researcher needs and operational lessons into robust platform improvements, reusable tooling, and clear runbooks.

This role is particularly suited to someone who enjoys practical systems work, learning new technologies, and collaborating closely with scientists and engineers. We do not expect candidates to have deep experience in every technology listed in this description. Instead, we are looking for a strong, scientifically minded systems engineer: someone who can troubleshoot complex environments, communicate clearly with multidisciplinary teams, learn unfamiliar tools quickly, and help build reliable, scalable services that advance GBI’s scientific mission.

Key Responsibilities

  • Operate, maintain, and improve GBI’s hybrid HPC platform, including Linux-based compute environments, Slurm/Slinky workloads, Kubernetes/OKE services, Open OnDemand, GPU and CPU partitions, and shared storage.
  • Help provision, configure, scale, and validate compute, storage, networking, and platform services using infrastructure as code, configuration management, and automation tools such as Terraform, Helm and Ansible.
  • Monitor platform health, capacity, job scheduling, GPU utilisation, storage behaviour, and network performance; investigate issues using tools such as Prometheus and Grafana.
  • Support researchers in using our Scientific Computing Platform, including triaging user issues and translating common pain points into platform improvements.
  • Build and maintain reproducible runtime environments, container images, and workflow-supporting services for scientific computing workloads, including bioinformatics, AI/ML, data processing, and simulation workflows.
  • Contribute to safe rollout and maintenance processes for Slurm images, worker node pools, scheduler configuration, container runtime changes, security updates, and monitoring improvements.
  • Create and maintain clear technical documentation, runbooks, validation checks, and issue/PR notes so the platform can be operated consistently and improved safely by the wider team.

Essential Knowledge, Skills and Experience:

  • Bachelor’s or Master’s degree in Computer Science, Computational Biology, Engineering, Physics, Mathematics, or a related discipline, or equivalent practical experience.
  • Hands-on experience supporting or administering Linux-based systems in an HPC, cloud, research, academic, or production environment.
  • Working knowledge of HPC or batch-computing concepts, including schedulers, resource requests, queues/partitions, shared filesystems, and multi-user compute environments; Slurm experience is preferred.
  • Ability to troubleshoot issues across systems, networking, storage, identity, containers, schedulers, and user workloads, and to follow problems through to a reliable operational fix.
  • Experience with scripting, automation, and version-controlled operational changes using tools such as Git, CI/CD, Terraform, Ansible, Helm, or similar.
  • Ability to work closely with multidisciplinary research teams, understand scientific computing needs, and deliver practical services that advance scientific goals.
  • Strong communication and documentation skills, with the ability to explain technical concepts clearly to scientists, engineers, and non-specialist audiences.
  • A proactive, learning-oriented approach suited to a new team building and improving a platform while also operating it day to day.

Desirable Knowledge, Skills and Experience:

  • Experience operating Slurm clusters, Slinky/slurm-operator, Open OnDemand, JupyterLab services, or other researcher-facing HPC portals and access patterns.
  • Experience with Kubernetes or managed Kubernetes platforms such as OCI OKE, EKS, GKE, or AKS, including Helm, Argo CD, operators, services, storage classes, and workload troubleshooting.
  • Experience with cloud infrastructure, particularly OCI, and with infrastructure as code and remote execution models such as Terraform Cloud.
  • Experience with shared and high-performance storage such as Lustre, BeeGFS, GPFS, NFS, OCI File Storage, object storage, or data movement workflows for large scientific datasets.
  • Experience supporting GPU-accelerated workloads, NVIDIA tooling, CUDA-aware environments, DCGM metrics, GPU health monitoring, and/or AI/ML and bioinformatics workloads on shared compute platforms.
  • Experience with containerised HPC and scientific workflow tooling, such as Apptainer/Singularity, Docker/Podman, Pyxis/Enroot, Nextflow, Snakemake, CWL, or WDL.
  • Experience building monitoring and operational dashboards using Prometheus, Grafana, exporter metrics, alerting rules, or capacity and reliability reporting.
  • Familiarity with identity, access, and security controls in Linux or research environments, such as OIDC, Okta ASA/PAM, least-privilege access, and security patching.
  • Experience working in a scientific, academic, life-science, or research computing environment where requirements evolve through close collaboration with researchers.

Our Benefits:

  • Salary: Competitive + travel allowance + bonus
  • Enhanced holiday pay
  • Pension
  • Life Assurance
  • Income Protection
  • Private Medical Insurance
  • Hospital Cash Plan
  • Therapy Services
  • Perk Box
  • Electric Car Scheme

Working Together – What It Involves:

  • You must have the right to work permanently in the UK with a willingness to travel as necessary. In certain cases, we can consider sponsorship, and this will be assessed on a case-by-case basis.
  • You will live in, or within easy commuting distance of, Oxford (or be willing to relocate).
  • Hybrid working

Similar roles

  • Botsync logo

    Senior Robotics Engineer (QC)

    Botsync·Bengaluru, India

    Asia’s leading robotics solutions provider, automating enterprise supply chains since 2017. We build syncOS — the industry-first no-code platform that lets manufacturers and warehouse operators integrate, reprogram, and scale automation systems without halting production or calling in specialists. Our autonomous mobile robots (AMRs) run 24/7 for global brands — transforming production lines in hours, not weeks. As Botsync scales AMR…

    • Full-time
  • Relex logo

    Senior Product Engineer

    Relex·London, Canada

    <div class="content-intro"><p><img src="https://www.relexsolutions.com/wp-content/uploads/2024/01/relex-greenhouse-banner-top.png" alt="" width="1280" style="max-width: 100%;"></p></div><p><strong><span data-contrast="auto">Who we are</span></strong>&nbsp;<br><span data-contrast="auto">We're bold thinkers and kind teammates, growing fast but staying grounded. Our Nordic roots and global outlook shape a culture of friendliness, creativity, and collaboration, all powered by openness and shared wins. We care deeply about doing the right thing, and doing it together. We aim to improve the flow…

    • On-site
    • Full-time
    • Express EntryPR day one, no employer
  • Founders Factory logo

    Founding Lead Engineer — AI & Workflow Automation

    Founders Factory·Berlin, Germany

    FOUNDING LEAD ENGINEER — AI & WORKFLOW AUTOMATION Founders Factory builds and funds startups together with exceptional entrepreneurs and leading companies. Founded by experienced entrepreneurs, we combine early-stage capital with a hands-on Venture Studio team across product, engineering, growth, talent and fundraising. We’re now building a new venture tackling one of Germany’s biggest labour-market challenges: how to find, engage and…

    • On-site
    • Full-time
    • Blue Cardtied; settle 21–33mo
  • Docker logo

    Account Executive, Strategic (EMEA)

    Docker·United Kingdom

    ABOUT DOCKER Docker has been one of the most loved brands in developer tooling, trusted by more than 20 million monthly users and over 20 billion container image pulls. From solo founders to the world's largest companies, developers rely on Docker to build, share, and run their applications across our suite of products including Docker Desktop, Docker Hub, and Docker…

    • On-site
    • Full-time
  • EverAI logo

    Senior Product Designer (Full Remote - UK)

    EverAI·United Kingdom

    OUR VISION & PRODUCTS 🚀 EverAI — Building the Future of AI Companionship One of the Top 15 Largest & Fastest-Growing AI Companies in the World 50 Million Users in 2 years — Help Us Reach 100M first, 500M next At EverAI, we’re shaping what it means to connect with AI. With 50 million users and counting, we're not just…

    • On-site
    • Full-time
  • BAE Systems logo

    Manufacturing Engineer – Estimator

    BAE Systems·Rochester, United Kingdom

    Job Title: Manufacturing Engineer – Estimator Location: Rochester; Kent: Onsite Salary: Circa £40,000 depending on experience Who we are: Join BAE Systems and you’ll be part of something bigger. As a valued member of our global colleague network, you’ll bring your unique skills and perspectives to help pioneer progress and protect what matters most. You’ll be trusted to play your…

    • Full-time
    • Skilled Workertied to sponsor
  • Duets Network logo

    Founding Engineer

    Duets Network·Worldwide

    Duets Network is an AI-powered platform connecting startups and investors across the sports and entertainment industries. We are building a dynamic ecosystem that bridges founders, capital, and strategic partners to drive meaningful growth and collaboration at the intersection of technology and culture. The initial product is built out, and we're now looking for a fractional Founding Engineer to guide the…

    • On-site
    • Full-time
  • KM Education Recruitment Ltd logo

    Health and Social Care Assessor

    KM Education Recruitment Ltd·St. Helens, United Kingdom

    KM Recruitment is a specialist UK wide recruiter for the Skills & Employability sectors. Job Title: Health and Social Care Assessor Location: North West - Home/Field based Salary: 27,000 - 31,000 Type: Full Time, Permanent Essential Criteria: Must hold a minimum of 1 years' experience of delivering Health and Social Care Apprenticeships. Hold a recognised Assessor award: D32/33, A1, CAVA…

    • On-site
    • Full-time
    • £31k–£31k / yr