Infrastructure and MLOps Engineer
Posted 22 days ago
Job Description
About Graphcore
At Graphcore, we’re building the future of AI compute.
We’re a team of semiconductor, software and AI experts, with deep experience in creating the complete AI compute stack - from silicon and software to infrastructure at datacenter scale.
As part of the SoftBank Group, backed by significant long-term investment, we are delivering key technology into the fast-growing SoftBank AI ecosystem.To meet the vast and exciting AI opportunity, Graphcore is expanding its teams around the world.We are bringing together the brightest minds to solve the toughest problems, in a place where everyone has the opportunity to make an impact on the company, our products and the future of artificial intelligence.
Job Summary
Join our dynamic Software Infrastructure team and take a pivotal role in scaling and managing our infrastructure. You will develop essential tools and services that empower our broader software team. Your contributions will enhance the build, test, deployment, and productisation processes of our Machine Learning Software components. Work with our High-Performance Computing (HPC) AI platforms and gain invaluable experience in distributed systems
The Team
The Software Infrastructure team provides critical platforms and services for software development teams across the business. Our responsibilities include managing the CI platform and services, build engineering, component integration, and packaging and release systems. We operate in squads, fostering a culture of service ownership and empowerment for our engineers. We focus on long-term engineering solutions and strive to eliminate toil wherever possible.
Responsibilities and Duties
- Develop, own, and maintain tools and services to support AI research and engineering teams
- Deploy and maintain services with Kubernetes and Docker
- Manage our Cloud Infrastructure using tools such as Terraform
Candidate Profile
Essential:
- Knowledge of Python
- Familiarity with cloud services (e.g. AWS)
- Experience managing or developing in Linux environments
- Understanding of CI/CD principles
- Experience using Kubernetes (k8s)
- Experience of one of the following:
- maintaining machine learning applications.
- deploying ML orchestration tools (e.g. NV Ray, KFP, SkyPilot).
- managing ML accelerator hardware (e.g. DCGM).
Desirable
- Experience with Infrastructure as Code (IaC) tools (e.g. Terraform/OpenTofu)
- Experience with GitHub Actions
- Experience with modern observability tooling (e.g. Prometheus)
- Experience with Grafana
- Knowledge of Go/Java/C++ (or similar language)
Benefits
In addition to a competitive salary, Graphcore offers flexible working, a generous annual leave policy, private medical insurance and health cash plan, a dental plan, pension (matched up to 5%), life assurance and income protection. We have a generous parental leave policy and an employee assistance programme (which includes health, mental wellbeing, and bereavement support). We offer a range of healthy food and snacks at our central Bristol office and have our own barista bar! We welcome people of different backgrounds and experiences; we’re committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments.
Applicants for this position must hold the right to work in the UK. Unfortunately at this time, we are unable to provide visa sponsorship or support for visa applications
Please mention that you found this job on MoAIJobs, this helps us grow. Thank you!
Graphcore
16 jobs posted
Apr 23, 2026
May 23, 2026
Similar Jobs
- 22d
Infrastructure and MLOps Engineer
Graphcore
Bristol, UK; Cambridge, United KingdomCambridge, UKInfrastructure and MLOps Engineer
Graphcore
Bristol, UK; Cambridge, United KingdomCambridge, UK22d - 10d
Staff Backend Engineer (Streaming and AI Infrastructure)
Coupang
$152K - $282KMountain View, USA; SeattleSeattle, USAStaff Backend Engineer (Streaming and AI Infrastructure)
Coupang
$152K - $282KMountain View, USA; SeattleSeattle, USA10d - 22d
Research Engineer, RL Infrastructure and Reliability (Knowledge Work)
Anthropic
$350K - $850KSan Francisco, CAResearch Engineer, RL Infrastructure and Reliability (Knowledge Work)
Anthropic
$350K - $850KSan Francisco, CA22d - 2d
Principal Engineer, AI Infrastructure (R4941)
Shield AI
San Francisco, CAPrincipal Engineer, AI Infrastructure (R4941)
Shield AI
San Francisco, CA2d - 28d
AI Infrastructure Engineer
Intercom
London, United KingdomAI Infrastructure Engineer
Intercom
London, United Kingdom28d - 28d
AI Infrastructure Engineer
Intercom
Berlin, GermanyAI Infrastructure Engineer
Intercom
Berlin, Germany28d - 28d
AI Infrastructure Engineer
Intercom
Dublin, IrelandAI Infrastructure Engineer
Intercom
Dublin, Ireland28d - 25d
AI Engineer
Toshiba
Guadalajara OfficeAI Engineer
Toshiba
Guadalajara Office25d - 23d
Principal Engineer, LLM (Platform and Tooling)
Upstart
Remote$238K - $330KUnited StatesPrincipal Engineer, LLM (Platform and Tooling)
Upstart
Remote$238K - $330KUnited States23d - 1d
AI Infrastructure Engineer
Together AI
$190K - $270KSan Francisco, CAAI Infrastructure Engineer
Together AI
$190K - $270KSan Francisco, CA1d