← All jobs

Infrastructure and MLOps Engineer

UNCLAIMED

Graphcore · Bristol · On-site

Full-time
0
transparency

This listing was indexed from Graphcore career page. Salary, benefits, progression, and satisfaction data are not verified. The trust score reflects what's missing.

About Graphcore

At Graphcore, we’re building the future of AI compute.

We’re a team of semiconductor, software and AI experts, with deep experience in creating the complete AI compute stack - from silicon and software to infrastructure at datacenter scale.

As part of the SoftBank Group, backed by significant long-term investment, we are delivering key technology into the fast-growing SoftBank AI ecosystem.To meet the vast and exciting AI opportunity, Graphcore is expanding its teams around the world.We are bringing together the brightest minds to solve the toughest problems, in a place where everyone has the opportunity to make an impact on the company, our products and the future of artificial intelligence.

Job Summary

Join our dynamic Software Infrastructure team and take a pivotal role in scaling and managing our infrastructure. You will develop essential tools and services that empower our broader software team. Your contributions will enhance the build, test, deployment, and productisation processes of our Machine Learning Software components. Work with our High-Performance Computing (HPC) AI platforms and gain invaluable experience in distributed systems

The Team

The Software Infrastructure team provides critical platforms and services for software development teams across the business. Our responsibilities include managing the CI platform and services, build engineering, component integration, and packaging and release systems. We operate in squads, fostering a culture of service ownership and empowerment for our engineers. We focus on long-term engineering solutions and strive to eliminate toil wherever possible.

Responsibilities and Duties

    Develop, own, and maintain tools and services to support AI research and engineering teams
    Deploy and maintain services with Kubernetes and Docker
    Manage our Cloud Infrastructure using tools such as Terraform

Candidate Profile

Essential:

    Knowledge of Python
    Familiarity with cloud services (e.g. AWS)
    Experience managing or developing in Linux environments
    Understanding of CI/CD principles
    Experience using Kubernetes (k8s)
    Experience of one of the following:

    maintaining machine learning applications.

    deploying ML orchestration tools (e.g. NV Ray, KFP, SkyPilot).

    managing ML accelerator hardware (e.g. DCGM).

Des

PythonGoJavaAwsDockerKubernetesTerraformCI/CD

What's missing

Verified salary range — not disclosed
Employee benefits — not disclosed
Career progression — not disclosed
Team satisfaction — not disclosed
Skill challenge — not disclosed

Company data (auto-enriched)

Employees600
Founded2016
FundingSeries E · £500M
Glassdoor3.8/5
Tech stack
PythonPyTorchC++CUDAPopART
0
Transparency
Salary0/30
Benefits0/20
Progression0/20
Satisfaction0/15
Challenge0/15
Apply on Graphcore site →
Redirects to employer's site · not tracked
Want to apply through ShowJob? Ask Graphcore to claim this listing for tracked applications and skill challenges.