<p><strong>.
Own the reliability of AI compute hardware where lab discovery becomes data centre reality.
Graphcore is building the hardware and systems infrastructure needed for the next generation of AI breakthroughs
This role helps prove those platforms work when complexity, scale and pressure increase.
</p> <p>You will lead advanced operational, diagnostic and engineering support across lab and data centre environments.
Your work will improve hardware bring-up, validation and troubleshooting for AI compute platforms.
</p> <p>You will diagnose failures across server blades, racks, power systems, thermal behaviour, network configuration and BIOS/BMC issues.
You will turn hard problems into clear root cause analysis, corrective action and better operating practice.
</p> <p>This is hands-on engineering with visible impact.
You will support new platforms from early bring-up through deployment, helping Graphcore move faster with confidence.
</p> <p>聽</p> <p><strong>The team and culture</strong></p> <p>You will work with Systems Engineering, Hardware Engineering, server engineering, firmware, platform architects and data centre operations.
Problems are solved close to the hardware, with clear evidence and direct ownership.
</p> <p>The team moves quickly during bring-up and validation cycles.
Engineers are expected to act fast, take responsibility and speak up when something needs attention.
</p> <p>Decisions are shaped by data, system behaviour and practical engineering judgement.
You will guide others, improve documentation and help raise the standard of hardware diagnostics.
</p> <p>What we're looking for</p> <p>路 Strong knowledge of server hardware architectures and board-level debugging<br>路 Experience isolating failures using system logs, telemetry, power data and thermal metrics<br>路 Hands-on experience with HPC systems, AI compute platforms or rack-scale infrastructure<br>路 Ability to lead structured root cause analysis and pr.