Senior Cloud Engineer
Job description
About the role
You will own the full lifecycle of the cloud infrastructure that directly powers Graphcore AI workloads and moves from requirements to production. This role focuses on deploying and operating private cloud services that validate advanced AI compute for both internal and external demands. You will work closely with the Cloud Platform Team to translate infrastructure requirements and user needs into reliable, deployed services. Decisions will be shaped by technical judgement and pace, especially when interacting with pre-release hardware and software. Clear communication is as valued as technical skill, and you will spend time each week on planning, code review, and debugging. Strong engineers are able to explain technical decisions in plain words and separate themselves through ownership and impact.
Key facts
What you'll do
Deploy and operate private cloud services to run Graphcore AI workloads and validate advanced AI compute capabilities.
Translate infrastructure requirements and product and user needs into reliable, deployed services that meet operational standards.
Apply judgement and pace to shape outcomes while working with pre-release hardware and software in a fast-moving environment.
Hold a Bachelor's degree or equivalent practical experience in a relevant subject as proof of foundational knowledge.
Bring solid software engineering or IT experience with a proven track record of delivery as an individual contributor.
Manage on-premises or private-cloud environments in production scenarios to ensure stability and performance.
Use strong Linux scripting and system administration skills across bash, Python, awk, sed, Ubuntu, RHEL and variants to automate and maintain systems.
Leverage Git, CI pipelines, and Infrastructure-as-Code tools such as Terraform/OpenTofu, Ansible or Packer to build and manage infrastructure.
Understand cloud services, virtualisation, networks, storage, monitoring and observability in practical, operational terms.
Collaborate across software, datacentre, and product teams to align infrastructure work with business and product goals.
Maintain and improve infrastructure reliability, security, and scalability to support demanding AI workloads.
Contribute to a culture of ownership, where engineers are responsible for end-to-end outcomes and continuous improvement.
Requirements
You hold a Bachelor's degree or equivalent practical experience in a relevant subject.
You bring solid software engineering or IT experience with proven delivery as an individual contributor.
You have managed on-premises or private-cloud environments in production scenarios.
You perform Linux scripting and system administration across bash, Python, awk, sed, Ubuntu, RHEL and variants strongly.
You use Git, CI pipelines and Infrastructure-as-Code tools such as Terraform/OpenTofu, Ansible or Packer.
You understand cloud services, virtualisation, networks, storage, monitoring and observability in practice.
You must have the legal right to work in the UK, as the role does not include visa sponsorship or support.
You are comfortable making decisions and moving quickly when working with pre-release hardware and software.
You communicate technical ideas clearly, both in writing and in conversation, to align with cross-functional partners.
Nice to have
Not applicable, as the source does not list any preferred additional qualifications or skills.
Practical notes
The role requires UK work eligibility, and there is no visa sponsorship or support provided by the company.
Typical interview steps include a recruiter screen, one or two technical rounds, possible coding problems, discussion of past projects, and system design questions.
Some loops may include a take-home task, and final rounds focus on team fit with opportunities for you to ask questions.
Interviewers value how you break down unfamiliar problems, so practicing aloud and reviewing past projects is beneficial.
Career growth and company context
Engineering careers at Graphcore usually progress from individual contributor to senior, staff, and principal levels, with some engineers moving into management roles that lead teams of five to twenty people. Others stay on the technical track, and growth is based on demonstrated impact rather than tenure alone. A typical engineering ladder has clear levels with defined expectations for scope, quality, and mentorship.
Graphcore has built a new type of processor for machine intelligence to accelerate machine learning and AI applications in a world of intelligent machines. As a Senior Cloud Engineer, you will play a key role in ensuring that the infrastructure supporting these processors is robust, scalable, and aligned with the needs of internal and external customers. You will work in an environment that values clarity, delivery, and collaboration, and you will be expected to take ownership of complex problems from initial requirements through to production operations. The role is well suited for engineers who enjoy working with emerging hardware, solving difficult infrastructure challenges, and contributing to a high-impact technical platform.