Help shape the compute platform behind one of the world’s most demanding technology environments.
This is an opportunity to lead the architecture of a large-scale Linux and HPC platform supporting quantitative research, AI/ML and high-performance distributed computing. You’ll define the future direction of compute, storage, networking and GPU infrastructure while working alongside world-class infrastructure and software engineers.
If you enjoy solving complex systems challenges at scale and influencing technical strategy, this role offers the opportunity to work with cutting-edge technologies in an environment where performance matters.
What You’ll Be Working On
- Designing and evolving large-scale Linux compute platforms supporting research, simulation and AI workloads.
- Architecting highly available infrastructure across compute, storage, networking and GPU environments.
- Defining platform standards, reference architectures and long-term infrastructure strategy.
- Evaluating next-generation server platforms, accelerators, storage technologies and high-speed interconnects.
- Optimising system performance across CPU, memory, storage, networking and GPU resources.
- Designing scalable monitoring, telemetry and capacity planning solutions.
- Driving improvements around workload scheduling, cluster utilisation and infrastructure efficiency.
- Leading complex infrastructure initiatives from design through implementation.
- Partnering with engineering and research teams to deliver scalable, high-performance platform solutions.
What We’re Looking For
- Extensive experience designing large-scale Linux infrastructure or HPC environments.
- Strong understanding of distributed compute platforms and systems architecture.
- Experience with GPU infrastructure and accelerator technologies.
- Knowledge of high-performance networking including InfiniBand, RoCE and high-speed Ethernet.
- Experience with parallel or enterprise storage technologies such as VAST, DDN, Lustre, BeeGFS, GPFS or similar.
- Hands-on experience with Kubernetes, Docker and workload schedulers such as Slurm.
- Strong Linux systems expertise with automation experience using Python, Go or similar.
- Understanding of AI/ML infrastructure and modern data centre technologies.
- Excellent communication skills with the ability to influence technical direction across multiple engineering teams.
What Can You Expect?
- Architect infrastructure that powers large-scale quantitative research and AI.
- Work with modern Linux platforms, GPU clusters and high-performance networking technologies.
- Influence technology decisions spanning compute, storage, networking and data centre infrastructure.
- Collaborate with highly skilled engineers solving complex distributed systems challenges.
- Competitive compensation, performance-based rewards, ongoing professional development and relocation support where applicable.
Does this sound like the next step in your career? We’d love to hear from you.