Back to search
EngineeringNot provided by source

Senior Software Engineer - AI Compute Libraries & Performance

Graphcore · Not provided by source

Apply
Last seen by MeritLog September 12, 2026Source: GreenhouseSource version: greenhouse-job-board-v1

MeritLog read this listing from Graphcore's Greenhouse job board and last checked it on September 12, 2026.

Source: the employer's Greenhouse job board. Open the original listing for current details.

Job details

Work model
Not provided by source
Salary
Not listed by source
Location
Bristol, UK
Occupation
Software Developers(O*NET 15-1252.00)

Hiring context

How this role compares at Graphcore

Graphcore has 181 live roles in MeritLog’s catalog across 5 job families, and 142 of them are in engineering. 0 of those listings publish a pay range, a disclosure rate of 0%.

Counted across the job boards MeritLog tracks, at the time this page was served. Pay comparisons use only listings that publish a complete range in the same currency and period.

What the role asks for

What you'd do

  • Design and implement kernels for linear algebra and tensor ops (GEMM, batched GEMM, convolutions, reductions, elementwise and fused operations) in C++
  • Own performance and correctness - add microbenchmarks, regression tests, numerics validation
  • Profile and optimise across for next generation of AI hardware - threading, cache locality, memory layout, and kernel launch efficiency.
  • Debug issues, resolve bugs and generally improve the quality and functionality of the product
  • Actively engage in and support Agile ways of working within the team
  • Mentor colleagues within the team, sharing knowledge and providing guidance where appropriate
  • Excellent programming and scripting skills using C++ and Python
  • Understanding of processor architectures and profiling on Linux
  • Possess excellent written and oral communication skills, good work ethics, high sense of team-work
  • Love to produce quality work and be a team player

What they're asking for

  • Strong command of algorithmic performance - vectorisation, memory hierarchy, threading, lock-free patternsSkillPreferred
  • Hands-on with at least one BLAS/DNN stack and able to read/extend kernelsSkillPreferred
  • Comfort with CPU micro-optimisations and numerical stability/trade-offs across FP32/FP16/BF16/FP8SkillPreferred
  • Experience integrating native code into PyTorch or similar (custom ops, extensions, dispatch keys)SkillPreferred
  • ABI/API stability and packaging for Linux system (manylinux, wheels)SkillPreferred

Parsed by MeritLog from the employer’s own posting. The full description follows below.

Job description

About the job Build the compute kernels that make next-generation AI hardware perform at its limit. As a Senior Software Engineer, you will create high-performance AI compute libraries for Graphcore’s next-generation hardware. Your work will sit close to the hardware, where every design choice matters. You will own kernels for linear algebra and tensor operations, including GEMM, convolutions, reductions and fused operations. You will improve performance, correctness and numerical reliability across critical AI workloads. This role is for engineers who enjoy hard performance problems and care deeply about quality. You will help shape software that enables customers to get more from AI hardware. The team and culture You will join the ML Kernels and Runtime team, an expanding group focused on high-performance compute libraries. The team works close to the hardware and close to the product. Work happens through clear ownership, practical decision-making and fast technical feedback. Engineers profile, test, debug and improve the product with accountability for real outcomes. You will be expected to speak up, share knowledge and help others raise the bar. Mentoring, technical judgement and all-round leadership matter here Responsibilities and Duties • Design and implement kernels for linear algebra and tensor ops (GEMM, batched GEMM, convolutions, reductions, elementwise and fused operations) in C++ • Own performance and correctness - add microbenchmarks, regression tests, numerics validation • Profile and optimise across for next generation of AI hardware - threading, cache locality, memory layout, and kernel launch efficiency. • Debug issues, resolve bugs and generally improve the quality and functionality of the product • Actively engage in and support Agile ways of working within the team • Mentor colleagues within the team, sharing knowledge and providing guidance where appropriate Candidate Profile Essential • Excellent programming and scripting skills using C++ and Python • Understanding of processor architectures and profiling on Linux • Possess excellent written and oral communication skills, good work ethics, high sense of team-work • Love to produce quality work and be a team player Desirable • Strong command of algorithmic performance - vectorisation, memory hierarchy, threading, lock-free patterns • Hands-on with at least one BLAS/DNN stack and able to read/extend kernels • Comfort with CPU micro-optimisations and numerical stability/trade-offs across FP32/FP16/BF16/FP8 • Experience integrating native code into PyTorch or similar (custom ops, extensions, dispatch keys) • ABI/API stability and packaging for Linux system (manylinux, wheels) Benefits · Unlimited annual leave · Up to 5% matched pension · Phantom equity – share in Graphcore’s success · True flexibility in how and where you work · Office spaces designed for collaboration · Free food and an on-site barista · Health cash plan · Income protection · Life assurance · Along with other benefits you can choose from (private medical insurance, dental plan etc) We welcome people from all backgrounds and experiences and are committed to building an inclusive environment where everyone can do their best work. We’re an equal opportunity employer and recognise that everyone brings different strengths and perspectives. If you need any adjustments during the interview process, just let us know - we’re happy to support you. Join the team at Graphcore Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore brings together deep expertise to solve complex problems and deliver meaningful progress in AI compute. If you want to shape the compute libraries behind next-generation AI hardware, we’d love to hear from you. Apply now to help build what comes next.

Keep exploring

More Engineering roles

Search all jobs

Privacy choices

Analytics and advertising stay off unless you allow them. Private data stays out.

Read the privacy notice