Explore internships, research positions, jobs, scholarships and fellowships from top organizations across India and worldwide.

Graphcore
About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary We are seeking an experienced Staff UEFI Engineer to design, develop, and deploy UEFI-based firmware for Graphcore’s hyperscale AI server platforms. This role focuses on developing system firmware responsible for platform initialization, hardware configuration, and reliability features across large-scale data center deployments. The Team Graphcore is a globally recognised leader in Artificial Intelligence computing systems. The company designs advanced semiconductors and data centre hardware that provide the specialised processing power needed to drive AI innovation, while delivering the efficiency required to support its broader adoption. The Firmware Engineering team develops the foundational firmware responsible for platform initialization, hardware configuration, and system reliability across Graphcore’s AI compute infrastructure. The team collaborates with silicon engineering, hardware design teams, BMC firmware teams, validation engineers, and platform architects to enable reliable server platform bring-up and operation. Responsibilities and Duties Design, develop, and deploy UEFI-based firmware for hyperscale server platforms. Collaborate with ODM partners throughout the design lifecycle from concept to mass production. Integrate UEFI firmware into CI/CD pipelines enabling automated builds, regression testing, and static analysis. Develop firmware functionality supporting platform initialization including CPU, memory, PCIe, and system interconnects. Implement system firmware security features including root of trust, secure boot chains, and signed firmware updates. Develop platform firmware features supporting server reliability, availability, and serviceability (RAS). Build firmware interfaces supporting telemetry, firmware updates, and system management capabilities. Collaborate with hardware, BMC, security, and validation teams to ensure full platform integration. Debug and perform root cause analysis for firmware and hardware issues across lab and production environments. Candidate Profile Essential Bachelor’s or Master’s degree in Electrical Engineering, Computer Engineering, Computer Science, or related discipline. 8+ years of experience developing UEFI or BIOS firmware for server platforms. Expertise with UEFI, TianoCore, and firmware architecture design. Strong experience developing firmware solutions for hyperscale or cloud data center environments. Strong programming skills in C/C++. Deep understanding of DDR memory training, cache coherency protocols, and PCIe subsystems. Strong knowledge of server platform architecture including power delivery, thermal management, sensors, and FRUs. Experience implementing CI/CD pipelines for firmware development. Experience debugging system firmware using logic analyzers, JTAG, GDB, and similar tools. Desirable Experience developing UEFI firmware for ARM-based server platforms. Experience with EDK II codebase development and upstream contribution. Experience working with ODM/JDM partners on server platform development programs. Experience delivering firmware for hyperscale cloud deployments and production environments. USA Benefits In addition to a competitive salary, Graphcore offers flexible working and a comprehensive benefits package designed to support your health, wellbeing and financial future. Our benefits include medical, dental and vision coverage, Flexible Spending Accounts (FSAs), Health Savings Accounts (HSAs), disability and life insurance, a 401(k) retirement plan, commuter benefits, wellness services and an Employee Assistance Programme (EAP). We welcome people of different backgrounds and experiences; we're committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments.
Graphcore
About Graphcore At Graphcore, we’re building the future of AI compute. We’re a team of semiconductor, software and AI experts, with deep experience in creating the complete AI compute stack - from silicon and software to infrastructure at datacenter scale. As part of the SoftBank Group, backed by significant long-term investment, we are delivering key technology into the fast-growing SoftBank AI ecosystem. To meet the vast and exciting AI opportunity, Graphcore is expanding its teams around the world. We are bringing together the brightest minds to solve the toughest problems, in a place where everyone has the opportunity to make an impact on the company, our products and the future of artificial intelligence Job Summary Applicants for this role should have strong experience designing, developing, and maintaining high-quality software systems. The role focuses on testing and validating a complex machine learning software stack, with particular emphasis on software architecture, automation, and engineering best practices. The ideal candidate is an experienced software engineer who values code quality, testability, and long-term maintainability, and enjoys building systems that other engineers rely on. This person will be comfortable working across large codebases, contributing to CI/CD infrastructure, and shaping technical direction through thoughtful design and mentoring in a technically demanding environment spanning ML frameworks, infrastructure, and AI accelerator hardware The Team The ML QA team is composed of highly skilled software engineers with a strong focus on automation, software quality, and data-driven validation. The team works closely with industry-standard machine learning frameworks and models, contributing to upstream open-source projects and collaborating across the wider software organization. Operating in a fast-paced environment, the team plays a critical role in ensuring reliability, performance, and maintainability across the ML software stack, helping to deliver robust and high-quality products to customers. Responsibilities and Duties Design, implement, and maintain robust test infrastructure and automation for a complex ML software stack. Architect and evolve test frameworks and tooling with a focus on scalability, maintainability, and developer experience. Build and maintain CI/CD pipelines targeting simulators, emulators (e.g. QEMU), and physical hardware. Create representative ML workloads and gain insights from their execution. (Numerical accuracy, performance analysis and benchmarking). Work closely with all Software development teams, supporting a culture of quality, security and maintainability. Review code and designs, setting a high bar for software engineering best practices. Mentor and support junior engineers, helping raise the overall technical capability of the team. Evaluate existing test strategies and infrastructure, identifying gaps and driving improvements aligned with team and organizational goals. Candidate Profile Essential: Experience in production-quality software engineering roles. Strong software design and architecture skills, with experience working on large or complex systems. Strong proficiency in Python, including experience building and maintaining production codebases. Solid experience with CI/CD systems and automated testing (preferably GitHub-based workflows). Experience working in Linux environments. Familiarity with C or C++, with the ability to read, debug, and reason about low-level code when needed. Proven ability to mentor junior engineers and influence engineering practices within a team. Strong problem-solving skills and a proactive, self-directed approach to work. Bachelor/Master's/PhD or equivalent experience in Computer Science, Maths, Machine Learning, Data Science, or related field. Desirable Exposure to machine learning frameworks such as PyTorch, JAX, Triton, TensorFlow Experience with distributed workload management systems such as Kubernetes, VLLM, Keras or MLOps pipelines Experience working with hardware simulators or emulators (e.g. QEMU). Experience developing for or working with FPGA-based systems. Experience with people management or mentoring Benefits In addition to a competitive salary, Graphcore offers flexible working, a generous annual leave policy, private medical insurance and health cash plan, a dental plan, pension (matched up to 5%), life assurance and income protection. We have a generous parental leave policy and an employee assistance programme (which includes health, mental wellbeing, and bereavement support). We offer a range of healthy food and snacks at our central Bristol office and have our own barista bar! We welcome people of different backgrounds and experiences; we’re committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments. Applicants for this position must hold the right to work in the UK. Unfortunately at this time, we are unable to provide visa sponsorship or support for visa applications
Graphcore
About Us Graphcore is a globally recognised leader in Artificial Intelligence computing systems. The company designs advanced semiconductors and data centre hardware that provide the specialised processing power needed to drive AI innovation, while delivering the efficiency required to support its broader adoption. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Job Summary Working within the logical design team, the silicon logical design engineer is responsible for a wide range of logical design tasks. The team is responsible for delivering the Microarchitecture and RTL design to implement the chip architecture specification for Graphcore Silicon, working closely with other engineers within the Silicon team. The successful candidate will be responsible for helping the team deliver high quality micro-architecture and RTL for Graphcore chips, working within the logical design team and with the broader Silicon team to ensure we meet the company objectives for Silicon delivery. Responsibilities and Duties Integrate IP and subsystems into top-level SoC designs Develop and maintain build and configuration environments Perform synthesis, linting, CDC/RDC, and timing checks at the SoC level Support verification and physical design teams through clean interface hand-offs Debug and resolve integration-related issues across multiple hierarchies Contribute to the continuous improvement of integration flows and automation Producing high quality microarchitecture and other documentation Ensure good communication between sites to maintain consistent working practises Candidate Profile Essential skills: Logical design experience in relevant industry Experience range 8-12 years in Semiconductor Industry/Product development exposure. Be highly motivated, a self-starter, and a team player Ability to work across teams and debugging issues seen to find root causes Degree in Computer Science, Engineering or related subject Ability to script in Python and/or TCL for automation and to solve design issues SystemVerilog Knowledge of digital design flows Desirable skills: Processor design Application specific blocks High-speed serial interfaces Complex third-party IP integration Arithmetic pipeline design Floating-point design Design for test Synthesis Timing Analysis Project Planning Power Integrity Silicon transistor Benefits: In addition to a competitive salary, Graphcore offers a competitive benefits package. We welcome people of different backgrounds and experiences; we’re committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments.
Graphcore
About Graphcore At Graphcore, we’re building the future of AI compute. We’re a team of semiconductor, software and AI experts, with deep experience in creating the complete AI compute stack - from silicon and software to infrastructure at datacenter scale. As part of the SoftBank Group, backed by significant long-term investment, we are delivering key technology into the fast-growing SoftBank AI ecosystem. To meet the vast and exciting AI opportunity, Graphcore is expanding its teams around the world. We are bringing together the brightest minds to solve the toughest problems, in a place where everyone has the opportunity to make an impact on the company, our products and the future of artificial intelligence. Job Summary Working within the Logical Design team, the Staff/Principal Silicon Logical Design Engineer is responsible for a wide range of logical design tasks which require working closely with other engineers within the Silicon department. This person is responsible for helping the team deliver high quality micro-architecture specifications and RTL for Graphcore chips, and assisting the rest of the Silicon teams to ensure we can meet the company objectives for silicon delivery. The Team The Logical Design team sits within the Silicon team. We are responsible for delivering the micro-architecture and RTL design to implement the chip architecture specification for Graphcore Silicon. Responsibilities and Duties Being part of the Logical Design team, producing high quality micro-architecture specification and RTL for Graphcore chips Ensuring good communication between different teams and across multiple sites Contributing to shared design infrastructure and flows Using EDA tools and the Graphcore design flow to design innovative technologies Contributing to block-level and chip-level checks and auditing (Linting, Synthesis, Timing Closure, CDC, RDC, Coverage) Working closely with Physical Design, Verification and DFT teams Essential skills: Degree in Computer Science, Engineering or a related subject Digital design experience in a relevant industry Be highly motivated, a self-starter, and a team player Ability to work across teams and finding solutions to problems Ability to solve design issues through programming, e.g. Python, Tcl Strong competency in SystemVerilog or VHDL Knowledge of digital design flows Capable of managing time and priorities Desirable: Processor design Application specific blocks High-speed serial interfaces Complex third-party IP integration Experience leading teams Arithmetic pipeline design Floating-point design Design for test Synthesis Timing Analysis Logical Equivalence Project Planning Power Integrity Silicon transistor knowledge Benefits In addition to a competitive salary, Graphcore offers flexible working, a generous annual leave policy, private medical insurance and health cash plan, a dental plan, pension (matched up to 5%), life assurance and income protection. We have a generous parental leave policy and an employee assistance programme (which includes health, mental wellbeing, and bereavement support). We offer a range of healthy food and snacks at our central Bristol office and have our own barista bar! We welcome people of different backgrounds and experiences; we’re committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments.
Graphcore
Manufacturing Test Engineer – Server Hardware Graphcore is a globally recognised leader in Artificial Intelligence computing systems. The company designs advanced semiconductors and data centre hardware that provide the specialised processing power needed to drive AI innovation, while delivering the efficiency required to support its broader adoption. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. We are opening a new AI Engineering Campus in Austin, which will play a central role in Graphcore's work building the future of AI computing. Role Overview We are seeking an experienced Manufacturing Test Engineer to support high-volume server manufacturing from board-level test through system-level production test. This role will work closely with an ODM manufacturing partner to define, implement, validate, and optimize the manufacturing test strategy for L6 board-level products, including ICT, MDA, and Board Functional Test, as well as support L10 system-level manufacturing test. The ideal candidate has strong experience in server hardware manufacturing, Linux-based test environments, diagnostic test coverage, fixture requirements, yield improvement, and root cause corrective action processes. This role requires both technical depth and hands-on manufacturing execution experience, with the ability to drive best practices across test development, factory readiness, quality planning, and ongoing production support. Key Responsibilities Manufacturing Test Strategy and Planning Work with ODM partners to define and execute the manufacturing test strategy for L6 board-level production. Develop and review test plans covering: In-Circuit Test, or ICT Manufacturing Defect Analyzer, or MDA Board Functional Test Diagnostic coverage requirements Manufacturing line test flow Failure detection and containment strategy Ensure test plans provide appropriate coverage for board-level defects, assembly issues, component-level failures, and functional performance requirements. Partner with hardware engineering, design validation, diagnostics, operations, quality, and ODM teams to align manufacturing test coverage with product risk areas. L6 Board-Level Test Development and Deployment Define requirements for board-level test stations, fixtures, test software, diagnostic content, and production test infrastructure. Support the development, validation, and release of board functional tests into the ODM manufacturing environment. Review ICT and MDA coverage reports and drive improvements to ensure adequate manufacturing defect detection. Define pass/fail criteria, test limits, data collection requirements, retest rules, and failure handling processes. Support bring-up, debug, and qualification of manufacturing test processes during NPI and production ramp. L10 System-Level Manufacturing Test Support development and deployment of L10 system-level manufacturing test processes. Port board functional tests into the L10 manufacturing environment where appropriate. Ensure L10 test coverage validates system-level integration, board functionality, firmware readiness, thermal behavior, power behavior, I/O functionality, and platform health. Work with ODM and internal engineering teams to ensure test execution is scalable, repeatable, and suitable for high-volume server production. Manufacturing Line and Fixture Requirements Specify manufacturing line requirements for test station configuration, test sequencing, data capture, networking, tooling, and operator workflow. Define requirements for test fixtures, cabling, adapters, load boards, debug interfaces, power delivery, signal access, and fixture maintenance. Ensure fixtures and test stations meet manufacturing requirements for reliability, repeatability, safety, ease of use, throughput, and serviceability. Drive fixture validation, correlation, preventive maintenance planning, and readiness for production ramp. Quality Planning, Yield, and Continuous Improvement Create and maintain an overall manufacturing quality plan focused on yield, defect containment, test coverage, and production readiness. Monitor manufacturing test yield, first-pass yield, failure pareto trends, retest rates, false failures, and escape risks. Lead technical investigations into manufacturing test failures, quality excursions, and yield loss. Drive structured root cause analysis and corrective action with ODM partners and internal stakeholders. Define and track corrective actions, containment plans, and long-term process improvements. Establish best practices for test development, test deployment, fixture readiness, failure analysis, data review, and manufacturing quality control. Cross-Functional Collaboration Serve as the primary technical interface between internal teams and ODM manufacturing test teams. Collaborate with hardware engineering, diagnostics, firmware, software, quality, supply chain, and operations teams. Support NPI builds, pilot builds, production ramp, and sustaining manufacturing activities. Communicate test readiness, risks, yield issues, corrective actions, and manufacturing quality status to program stakeholders. Travel to ODM manufacturing sites as needed to support build readiness, test deployment, debug, and ramp activities. DIFFERENTIATORS 8+ years of experience in manufacturing test engineering, hardware test engineering, or production test development. Experience supporting high-volume server manufacturing or similar complex compute, networking, storage, or data center hardware products. Strong understanding of board-level manufacturing test processes, including: ICT MDA Board Functional Test Diagnostic test execution Manufacturing defect detection Experience working directly with ODM, CM, or JDM manufacturing partners. Hands-on experience with Linux-based test environments, including test execution, scripting, log collection, and failure triage. Familiarity with server hardware architecture, including CPUs, memory, storage, networking, BMCs, firmware, power subsystems, and high-speed interfaces. Experience defining test fixture requirements and supporting fixture bring-up, validation, and production readiness. Strong understanding of manufacturing quality metrics, including first-pass yield, retest rate, defect paretos, failure analysis, and corrective action. Demonstrated ability to drive root cause analysis and corrective action across engineering and manufacturing teams. Ability to review test logs, identify failure signatures, isolate issues, and determine whether failures are related to hardware, firmware, software, test process, or fixture design. Strong written and verbal communication skills with the ability to clearly communicate technical issues, risks, and action plans. Preferred Qualifications Experience with L6 board-level and L10 system-level manufacturing processes. Experience porting board-level functional tests into system-level manufacturing environments. Familiarity with server diagnostics, BMC interfaces, BIOS/UEFI, firmware update flows, hardware health checks, and system stress testing. Experience with Python, Bash, or other scripting languages used in manufacturing test automation. Knowledge of manufacturing data systems, test result databases, yield dashboards, and factory analytics. Experience with high-volume NPI, EVT/DVT/PVT, pilot builds, and production ramp. Familiarity with test coverage analysis, DFT/DFM principles, and manufacturing escape prevention. Experience working with global manufacturing teams and offshore ODM sites. Key Success Measures Complete and production-ready L6 test plan covering ICT, MDA, and Board Functional Test. Successful deployment of board functional test into L6 and L10 manufacturing environments. Clearly defined manufacturing line, station, and fixture requirements. Strong diagnostic coverage aligned to product risk and manufacturing defect modes. Stable test processes with low false-failure rates and scalable execution time. Improved first-pass yield and reduced manufacturing defect escapes. Timely root cause identification and corrective action closure for yield and quality issues. Adoption of manufacturing test and quality best practices across ODM production lines. We welcome people of different backgrounds and experiences and are committed to building an inclusive work environment that makes Graphcore a great home for everyone. We are an equal opportunity employer and want to build a work environment where everyone is happy, productive and respectful so they can do their best work. If you have a disability or additional need that requires accommodation, just let us know. USA Benefits In addition to a competitive salary, Graphcore offers flexible working and a comprehensive benefits package designed to support your health, wellbeing and financial future. Our benefits include medical, dental and vision coverage, Flexible Spending Accounts (FSAs), Health Savings Accounts (HSAs), disability and life insurance, a 401(k) retirement plan, commuter benefits, wellness services and an Employee Assistance Programme (EAP). We welcome people of different backgrounds and experiences; we're committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments.
Graphcore
Staff Embedded SW/FW Engineer (Bringup) Graphcore is a globally recognised leader in Artificial Intelligence computing systems. The company designs advanced semiconductors and data centre hardware that provide the specialised processing power needed to drive AI innovation, while delivering the efficiency required to support its broader adoption. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. We are opening a new AI Engineering Campus in Bengaluru which will play a central role in Graphcore's work building the future of AI computing. Job Summary We have an exciting opportunity to be part of a collaborative, cross-functional development team developing C code used to validate cutting-edge, high-performance AI chips and platforms. You will play a critical role in supporting new product introductions and post-silicon validation. Working within the Post-Silicon Bringup team, you will be involved with bringing first silicon to life, developing code primarily in C to configure and exercise systems and sub-systems on new silicon devices, and working closely with many other teams to help it become a fully characterised and working product, reporting project status/progress to program management on a regular basis. You will have the opportunity to, and be responsible for, leading, mentoring, and providing technical guidance to other engineering team members. In this role, you can leverage your experience and industry knowledge to architect and drive implementation of continuous improvements to test infrastructure and processes. The Team The Post-Silicon Bringup team sits within the Architecture and Validation team, we are responsible for bringup and validation of new silicon when it returns from manufacture, enabling and supporting the production SW and FW teams to bring up their software and supporting the Silicon Characterisation team. Responsibilities and Duties Plan, design, develop and debug silicon bringup sequences and test in bare metal C/C++ on FPGA/Emulator prior to first silicon Deploy configuration sequences and validation tests on first silicon and debug them Develop automated test framework and regression test suites in Python to optimize validation efficiency Collaborate closely with engineers from many other disciplines on a variety of topics Work with Validation and Production Test engineering peers to implement best practices and continuous improvements to test methodologies Analyse test results, identify and debug failures/defects Contribute to shared test and validation infrastructure Provide feedback to architects Candidate Profile Essential: Strong experience in Bare metal / embedded C/C++ Good knowledge of digital ASICs. Be highly motivated, a self starter, and a team player Ability to work across teams and programming languages to find root causes of deep and complex issues Experience of the post-silicon validation process applied in digital ASIC environments Python, Linux Excellent communication skills and the ability to collaborate with others to solve problems. Excellent problem-solving, analytical & diagnostic skills Desirable: Driver level experience with one or more of the following is highly desirable: PCIe Ethernet Memory technologies (LPDDR, DDR, HBM, …) Other peripherals such as I2C, I3C, SPI, … Good knowledge of mixed-signal building blocks such as PLLs, high speed PHYs and IC control/communication protocols is highly desirable. Experience of Arm CPUs, System IP and debug tools. Experience of AMBA protocols. Understanding of ML applications and their workloads. Experience in Characterization, Failure Analysis, Test Development, Statistical analysis, and Customer Support Benefits: In addition to a competitive salary, Graphcore offers a competitive benefits package. We welcome people of different backgrounds and experiences; we’re committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments.
Graphcore
About Graphcore At Graphcore, we’re building the future of AI compute.We’re a team of semiconductor, software and AI experts, with deep experience in creating the complete AI compute stack - from silicon and software to infrastructure at datacenter scale. As part of the SoftBank Group, backed by significant long-term investment, we are delivering key technology into the fast-growing SoftBank AI ecosystem.To meet the vast and exciting AI opportunity, Graphcore is expanding its teams around the world.We are bringing together the brightest minds to solve the toughest problems, in a place where everyone has the opportunity to make an impact on the company, our products and the future of artificial intelligence. Job Summary We are looking for an experiences Staff Engineer to join our Cloud Platform Team and help develop and deploy clouds and services. Working closely with our colleagues in Software Platform, Datacentre Operations and Product Development teams, you will deploy services on our fleet of cutting-edge AI systems. As part of our Software Platform organisation, you will be involved in the cloud integration, validation, performance benchmarking, optimisation, and development of our high-performance AI solutions. These include in-house AI systems alongside off-the-shelf high-performance servers, switches and storage solutions. This is a hand-on technical role requiring a solid background in the use of cloud infrastructure, deployment using Infrastructure-as-Code, observability, high-performance networking and storage systems. You may have been working in an IT organisation, a datacentre, a cloud provider or as a developer of orchestration or cloud services. The Software Platform team at Graphcore We build Graphcore products into large-scale AI solutions for our customers and the Cloud Platform Team is responsible for providing such systems to both internal users via private clouds and customers via our own public clouds. Often the internal systems will be using and developing pre-release hardware and software, so it’s vital you are comfortable with unproven components. Responsibilities and Duties Operate and extend existing OpenStack-based cloud services and contribute to the deployment and development of new ones. Develop and operate end-user services on our clouds and support internal users in their use. You will turn end-user and product requirements into deployed services. Help to build automation to collect and analyse metrics and other observability data from the cloud services to support clear identification and reporting of any issues. Work with users to provide information of any product-related issues to Engineering and QA departments. Work with our Datacentre Operations Engineers to maintain and operate the fleet of AI systems at peak performance in our private clouds. Configure and test new Graphcore AI hardware and systems using Continuous Deployment and Infrastructure-as-code in internal and external datacentres. Drive corrective actions for systems that are not operating correctly, working with DC operations and Graphcore Engineering as required. Work with external vendors of off-the-shelf switches, servers and storage solutions to specify, benchmark and integrate 3rd party products into our Cloud Reference Design. Skills and Experience [ALL REQUIRED] Bachelor's degree or equivalent practical experience in a relevant subject. Solid infrastructure or IT experience with a proven track record of delivering technical output as an individual contributor. Experience managing or operating on-premises or private-cloud environments. Experience specifying, scoping, estimating and detailing work plans in an AGILE and SCRUM framework, including priorities, risks, issues, impacts and constraints. Strong proven Linux scripting ability (bash and python required). Strong proven Linux system administration (Ubuntu, RHEL and variants). Experience with a version control system (preferably Git) and using it to manage system configuration or automation. Experience with Continuous Integration or testing pipelines using GitLab, GitHub or similar. Hands-on experience deploying services into public or private clouds using Infrastructure-as-Code (IAC). A solid understanding of the technologies underpinning cloud services (APIs, virtualisation of CPUs, IO, systems), virtual networks, block storage, resource management and monitoring. Experience with IAC automation tools (e.g. Terraform/OpenTofu, Ansible, Packer). Experience with container deployment and management tools (e.g. docker, podman, apptainer). Experience with solutions for monitoring and observability. e.g. Grafana, Prometheus, OpenSearch/ElasticSearch, Loki, Mimir, OpenTelemetry, Fluentd ,Kafka Good communication and presentation skills, and experience dealing with end-users of IT or cloud services. An ability to work independently on critical infrastructure without oversight, and with a focus on end-user availability. Desirable but not required: Experience with OpenStack deployments or the technologies they rely on (e.g. Ceph, Open vSwitch, KVM, QEMU ). Experience with High Performance Computing (HPC) environments using SLURM or similar batch workload solutions. Strong skillset and experience in end-to-end deployment automation and CI of containerised services. Complete automation of pipelines for build, test, deploy, manage, alert, destroy, rebuild. Experience with managing production Kubernetes clusters and workloads. Experience with workload queue management systems (SLURM, LSF, Kueue). Experience with managed switch configuration (e.g. EOS, SONiC, DNOS). Programming experience with Python3 utilising classes and inheritance. Programming experience with Go. Benefits In addition to a competitive salary, Graphcore offers flexible working, a generous annual leave policy, private medical insurance and health cash plan, a dental plan, pension (matched up to 5%), life assurance and income protection. We have a generous parental leave policy and an employee assistance programme (which includes health, mental wellbeing, and bereavement support). We offer a range of healthy food and snacks at our central Bristol office and have our own barista bar! We welcome people of different backgrounds and experiences; we’re committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments. Sponsorship Applicants for this position must hold the right to work in the UK. Unfortunately at this time, we are unable to provide visa sponsorship or support for visa applications.
Graphcore
About Graphcore At Graphcore, we’re building the future of AI compute. We’re a team of semiconductor, software and AI experts, with deep experience in creating the complete AI compute stack - from silicon and software to infrastructure at datacenter scale. As part of the SoftBank Group, backed by significant long-term investment, we are delivering key technology into the fast-growing SoftBank AI ecosystem. To meet the vast and exciting AI opportunity, Graphcore is expanding its teams around the world. We are bringing together the brightest minds to solve the toughest problems, in a place where everyone has the opportunity to make an impact on the company, our products and the future of artificial intelligence. Job Summary We are looking for high-quality silicon physical design engineers to complement our existing exceptional team. We have a range of roles available with focus on those with extensive ranges of skills and experience although exceptional candidates with less experience will be considered. We want people who work collaboratively and proactively within a team focusing on collectively achieving our goals and creating the right engineering solutions. Good communication is essential, as is the ability to adapt and learn – we value the right characteristics more than specific experience. For the successful candidate we offer an open, honest and collaborative environment working on leading-edge designs at the most advanced nodes. Our engineers are not siloed, and they are trusted and encouraged to take ownership of their designs and problem solutions. You will become part of a team that looks for improvements to everything we do: our designs, our flows, our methodologies, our infrastructure. The Team The physical design team sits within the wider silicon design team which includes RTL, verification and DFT and with whom we collaborate extensively. Our work additionally involves strong links with architecture, packaging and product engineering. We are responsible for working with those teams to create high-quality RTL and then to build the final chip layout (e.g. GDSII) ensuring a signoff-quality design is delivered to the Foundry (e.g. TSMC). Responsibilities and Duties Applicants will be expected to contribute technically to the development of Graphcore's next generation of AI superchips, focusing on achieving robust, high-performance and power-efficient designs in leading-edge process technologies while meeting ambitious development schedules. Contributions are expected to span multiple areas and involve: using state-of-the-art EDA tools and in-house Graphcore flows to deliver final designs that are of sign-off quality enhancing existing flows and in-house tools/APIs to support new features and/or methodologies analysis and/or debugging of complex engineering problems leading to workable solutions developing an understanding of emerging technical issues and applying that knowledge to optimise in-house flows and methodologies Candidates will be expected to work closely both with other teams within Graphcore and with 3rd party support engineers/contractors, ensuring good communication between all parties, and to contribute meaningfully to the overall efficiency and success of the Physical Team. Candidate Profile Essential Skills and Experience: A Degree in Electronic/Electrical Engineering, Computer Science or related subject Be highly motivated, a self-starter, and a team player Enjoy taking responsibility and improving skills and knowledge Excellent problem-solving skills for debugging issues seen and finding root causes Ability to program/script (required to solve design issues, typically in Tcl and Python) Experience in 7nm or smaller technologies A good breadth of experience with physical design flows including: Floorplanning/Budgeting, Synthesis, Place and Route, Clock CTS, Timing Analysis, Logical Equivalence, Physical Verification (DRC/LVS/ERC) Desirable Skills and Experience: In depth expertise in one or more aspects of physical design flows Structuring of builds to maximise PPA Silicon transistor knowledge including std cell libraries and/or memories 2D and 3D design (CoWoS, UCIe etc.) High speed, high power and/or full reticle chip design Ethernet, PCIe, LPDDR, HBM interfaces Power integrity and optimization 2nm or 3nm technologies Chip finishing (pad rings, chip level LVS/DRC/ERC) Design for test Team management Project Planning Benefits In addition to a competitive salary, Graphcore offers flexible working, a generous annual leave policy, private medical insurance and health cash plan, a dental plan, pension (matched up to 5%), life assurance and income protection. We have a generous parental leave policy and an employee assistance programme (which includes health, mental wellbeing, and bereavement support). We offer a range of healthy food and snacks at our central Bristol office and have our own barista bar! We welcome people of different backgrounds and experiences; we’re committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments
Graphcore
About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary We are seeking an experienced Senior UEFI Engineer to design, develop, and deploy UEFI-based firmware for Graphcore’s hyperscale AI server platforms. This role focuses on developing system firmware responsible for platform initialization, hardware configuration, and reliability features across large-scale data center deployments. The Team Graphcore is a globally recognised leader in Artificial Intelligence computing systems. The company designs advanced semiconductors and data centre hardware that provide the specialised processing power needed to drive AI innovation, while delivering the efficiency required to support its broader adoption. The Firmware Engineering team develops the foundational firmware responsible for platform initialization, hardware configuration, and system reliability across Graphcore’s AI compute infrastructure. The team collaborates with silicon engineering, hardware design teams, BMC firmware teams, validation engineers, and platform architects to enable reliable server platform bring-up and operation. Responsibilities and Duties Design, develop, and deploy UEFI-based firmware for hyperscale server platforms. Collaborate with ODM partners throughout the design lifecycle from concept to mass production. Integrate UEFI firmware into CI/CD pipelines enabling automated builds, regression testing, and static analysis. Develop firmware functionality supporting platform initialization including CPU, memory, PCIe, and system interconnects. Implement system firmware security features including root of trust, secure boot chains, and signed firmware updates. Develop platform firmware features supporting server reliability, availability, and serviceability (RAS). Build firmware interfaces supporting telemetry, firmware updates, and system management capabilities. Collaborate with hardware, BMC, security, and validation teams to ensure full platform integration. Debug and perform root cause analysis for firmware and hardware issues across lab and production environments. Candidate Profile Essential Bachelor’s or Master’s degree in Electrical Engineering, Computer Engineering, Computer Science, or related discipline. 6+ years of experience developing UEFI or BIOS firmware for server platforms. Expertise with UEFI, TianoCore, and firmware architecture design. Strong experience developing firmware solutions for hyperscale or cloud data center environments. Strong programming skills in C/C++. Deep understanding of DDR memory training, cache coherency protocols, and PCIe subsystems. Strong knowledge of server platform architecture including power delivery, thermal management, sensors, and FRUs. Experience implementing CI/CD pipelines for firmware development. Experience debugging system firmware using logic analyzers, JTAG, GDB, and similar tools. Desirable Experience developing UEFI firmware for ARM-based server platforms. Experience with EDK II codebase development and upstream contribution. Experience working with ODM/JDM partners on server platform development programs. Experience delivering firmware for hyperscale cloud deployments and production environments. USA Benefits In addition to a competitive salary, Graphcore offers flexible working and a comprehensive benefits package designed to support your health, wellbeing and financial future. Our benefits include medical, dental and vision coverage, Flexible Spending Accounts (FSAs), Health Savings Accounts (HSAs), disability and life insurance, a 401(k) retirement plan, commuter benefits, wellness services and an Employee Assistance Programme (EAP). We welcome people of different backgrounds and experiences; we're committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments.
Graphcore
About Graphcore At Graphcore, we’re building the future of AI compute.We’re a team of semiconductor, software and AI experts, with deep experience in creating the complete AI compute stack - from silicon and software to infrastructure at datacenter scale.As part of the SoftBank Group, backed by significant long-term investment, we are delivering key technology into the fast-growing SoftBank AI ecosystem.To meet the vast and exciting AI opportunity, Graphcore is expanding its teams around the world.We are bringing together the brightest minds to solve the toughest problems, in a place where everyone has the opportunity to make an impact on the company, our products and the future of artificial intelligence. Job Summary We turn large-scale system measurements into decisions. Our team runs workloads across clusters of machines and collects detailed performance data. The challenge isn’t running the measurements—it’s deciding what they mean, and whether a system is good enough to enter production. You will work with results from real systems and help answer questions like: Is this system behaving as expected? Is performance stable enough to trust? Does this meet the criteria to enter production? This work includes systems engineering and analysis. It involves understanding variability, repeatability, and the differences between signal and noise. You won’t be confined to a single role. You may shape measurements, influence how they are run, or improve how results are interpreted. You are free to specialise, but the team is responsible for leaving no gaps. This is not a dashboarding or reporting role in the traditional sense. The goal is to produce outputs that support real engineering decisions. We’re looking for engineers who: Think carefully about uncertainty and evidence Prefer clarity over presentation Are comfortable challenging conclusions when data is weak Selection criteria: Our engineers typically bring significant practical experience and sound engineering judgement. Depth in one area is valued, but the ability to work across boundaries is equally important. Essential Strong software engineering experience, typically gained across multiple projects or systems over several years Experience working in Linux-based environments, ideally with distributed or high-performance systems Proficiency in Python Experience with automation and CI/CD systems (e.g. GitLab CI, Jenkins, GitHub Actions) Ability to design, implement, and run experiments or tests that produce meaningful results Ability to interpret results and communicate findings clearly, with an emphasis on accuracy and usefulness to decision-making Comfortable working in areas where requirements are not fully defined and judgement is required Desirable Experience working with large-scale or distributed systems (e.g. clusters, cloud platforms, HPC environments) Experience with performance, reliability, or systems-level testing/measurement Familiarity with pytest or similar frameworks for structured test/measurement execution Experience analysing system behaviour under load (compute, network, or ML workloads) Experience working with containerisation, orchestration, or provisioning systems (e.g. Docker, Kubernetes, OpenStack) Proficiency in other applications programming languages (e.g. C++)7. Exposure to data analysis, statistics, or interpreting variability in results Benefits In addition to a competitive salary, Graphcore offers flexible working, a generous annual leave policy, private medical insurance and health cash plan, a dental plan, pension (matched up to 5%), life assurance and income protection. We have a generous parental leave policy and an employee assistance programme (which includes health, mental wellbeing, and bereavement support). We offer a range of healthy food and snacks at our central Bristol office and have our own barista bar! We welcome people of different backgrounds and experiences; we’re committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments. Sponsorship Applicants for this position must hold the right to work in the UK. Unfortunately at this time, we are unable to provide visa sponsorship or support for visa applications.
Graphcore
About Graphcore At Graphcore, we’re building the future of AI compute.We’re a team of semiconductor, software and AI experts, with deep experience in creating the complete AI compute stack - from silicon and software to infrastructure at datacenter scale.As part of the SoftBank Group, backed by significant long-term investment, we are delivering key technology into the fast-growing SoftBank AI ecosystem.To meet the vast and exciting AI opportunity, Graphcore is expanding its teams around the world.We are bringing together the brightest minds to solve the toughest problems, in a place where everyone has the opportunity to make an impact on the company, our products and the future of artificial intelligence. Job Summary We decide whether entire racks of machines are good enough to enter production. Our team measures and evaluates large-scale Linux systems—from a single rack to data-centre scale. We don’t just run benchmarks; we determine whether a system behaves correctly, and whether it is reliable enough to be trusted. The work spans designing workloads, building execution systems, and interpreting results. A measurement is only useful if it leads to a decision. You won’t be confined to a narrow role. Some engineers here focus on infrastructure, others on workloads or analysis—but everyone contributes to understanding system behaviour at scale. You are free to specialise, but as a team we are responsible for leaving no gaps. You might find yourself: Expanding measurement coverage from small clusters to full racks Designing workloads that expose system behaviour Building systems to run experiments across large clusters Interpreting results and defining what “good enough” looks like This is not a pure infrastructure or data role. The work combines systems engineering, measurement, and judgement. We’re looking for engineers who are comfortable working where the right answer isn’t obvious, and where careful measurement matters more than output volume. Selection criteria: Our engineers typically bring significant practical experience and sound engineering judgement. Depth in one area is valued, but the ability to work across boundaries is equally important. Essential Strong software engineering experience, typically gained across multiple projects or systems over several years Experience working in Linux-based environments, ideally with distributed or high-performance systems Proficiency in Python Experience with automation and CI/CD systems (e.g. GitLab CI, Jenkins, GitHub Actions) Ability to design, implement, and run experiments or tests that produce meaningful results Ability to interpret results and communicate findings clearly, with an emphasis on accuracy and usefulness to decision-making Comfortable working in areas where requirements are not fully defined and judgement is required Desirable Experience working with large-scale or distributed systems (e.g. clusters, cloud platforms, HPC environments) Experience with performance, reliability, or systems-level testing/measurement Familiarity with pytest or similar frameworks for structured test/measurement execution Experience analysing system behaviour under load (compute, network, or ML workloads) Experience working with containerisation, orchestration, or provisioning systems (e.g. Docker, Kubernetes, OpenStack) Proficiency in other applications programming languages (e.g. C++) Exposure to data analysis, statistics, or interpreting variability in results Benefits In addition to a competitive salary, Graphcore offers flexible working, a generous annual leave policy, private medical insurance and health cash plan, a dental plan, pension (matched up to 5%), life assurance and income protection. We have a generous parental leave policy and an employee assistance programme (which includes health, mental wellbeing, and bereavement support). We offer a range of healthy food and snacks at our central Bristol office and have our own barista bar! We welcome people of different backgrounds and experiences; we’re committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments. Sponsorship Applicants for this position must hold the right to work in the UK. Unfortunately at this time, we are unable to provide visa sponsorship or support for visa applications.
Graphcore
About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary We are looking for an experienced System Level Test Engineer to join our Product Test and Diagnosis Department (PTD). In this role, you will contribute to the development and deployment of System Level Test (SLT) solutions for next-generation AI processors. Working closely with hardware, software, validation, and manufacturing teams, you will develop test content, automation, diagnostics, and characterization capabilities that support silicon bring-up, yield learning, and manufacturing deployment. The ideal candidate will have strong technical foundations in semiconductor test and validation, excellent debug skills, and a passion for improving product quality and manufacturability. The Team The Product Test and Diagnostics team’s role is to detect and manage hardware defects that arise from the manufacture and use of our products. This covers chips, boards and finished systems and takes place both in the manufacturing sites and in the field. Responsibilities and Duties Develop and maintain SLT test content, automation, diagnostics, and characterization workloads for silicon bring-up and manufacturing readiness activities. Execute characterization, voltage/frequency margining, reliability screening, and performance correlation activities while analyzing results and identifying improvement opportunities. Support development and deployment of SLT hardware including test boards, sockets, instrumentation, and thermal solutions. Integrate validation and bring-up workloads into the SLT environment and support development of manufacturing screening solutions. Debug failures, perform root-cause analysis, and support correlation activities across Simulation, ATE, Bench, SLT, and manufacturing environments. Develop software tools, scripts, and automation flows for test execution, telemetry collection, data analysis, and reporting. Contribute to improvements in test coverage, test-time optimization, DPPM reduction, and manufacturing readiness. Work closely with Design, DFT, Silicon Validation, Product Engineering, Manufacturing, and OSAT teams to support product development and deployment activities. Support transfer of SLT solutions into production environments and assist with manufacturing ramp activities. Mentor junior engineers and contribute to continuous improvement of SLT methodologies and best practices. Minimum Qualifications: Experience in semiconductor test, validation, characterization, post-silicon bring-up, or manufacturing environments. Understanding of semiconductor fundamentals, digital systems, power delivery, signal integrity, and board-level debugging. Familiarity with DFT concepts including scan, JTAG, BIST, MBIST, boundary scan, and manufacturing test methodologies. Experience with Python, Tcl, Shell, or similar scripting languages for automation and data analysis. Strong debug, problem-solving, root-cause analysis, and data interpretation skills. Experience working with cross-functional teams including Design, DFT, Validation, Product Engineering, or Manufacturing. Bachelor's degree in Electrical Engineering, Computer Engineering, or related field. Preferred Qulaifications Familiarity with advanced packaging, chiplets, or high-performance computing products. Experience with manufacturing test and production environments, along with OSAT engagement. USA Benefits In addition to a competitive salary, Graphcore offers flexible working and a comprehensive benefits package designed to support your health, wellbeing and financial future. Our benefits include medical, dental and vision coverage, Flexible Spending Accounts (FSAs), Health Savings Accounts (HSAs), disability and life insurance, a 401(k) retirement plan, commuter benefits, wellness services and an Employee Assistance Programme (EAP). We welcome people of different backgrounds and experiences; we're committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments.
Graphcore
About the job Own the reliability strategy behind advanced AI silicon built for datacenter scale. Graphcore is expanding the hardware platforms that will power the next generation of AI compute. As Senior Semiconductor Reliability Engineer, you will lead reliability strategy across advanced silicon nodes and packaging technologies. You will shape reliability from technology selection and design enablement through bring-up, qualification and manufacturing. Your work will help complex, high-performance devices move from ambitious design into robust production. You will assess risk, guide sign-off, lead qualification and drive failure analysis across silicon, interconnect and package technologies. This is a rare role at the centre of device physics, design, packaging and manufacturing. The team and culture You will work across silicon, packaging, manufacturing and quality, connecting technical detail with product-level decisions. Work moves through direct ownership, clear evidence and fast escalation when risks need attention. Decisions are grounded in data, modelling, qualification results and open technical debate. You will be expected to speak up, challenge assumptions and turn complex findings into clear next steps. This is a role for someone who thinks beyond individual tests. You will help build stronger reliability methods and mentor others as Graphcore scales advanced AI hardware platforms. What we’re looking for · Bachelor’s or Master’s degree in Electrical Engineering, Materials Science, Physics, Mechanical Engineering, or a related field · Strong semiconductor reliability experience across silicon, BEOL, interconnect or advanced package technologies · Deep understanding of mechanisms such as EM, TDDB, BTI and HCI · Hands-on experience with HTOL, thermal cycling, thermal shock, reflow testing and JEDEC reliability standards · Ability to use X-ray and CSAM analysis to diagnose package defects and support root-cause analysis · Strong technical judgement, communication skills and ownership across foundries, OSATs, labs and internal teams While we have outlined a set of requirements, we value transferable skills and diverse experiences. If you meet most of our essential criteria, we encourage you to submit an application and showcase how your background makes you a strong candidate. Benefits · Flexible working: Balance your work and personal life with greater flexibility · Generous leave: Take time to rest, recharge and enjoy life outside of work · Retirement planning support: Up to 5% matched pension · Phantom equity: Share in Graphcore’s success · Workplace experience: Enjoy thoughtfully designed office spaces for collaboration, with free food and an on-site barista to support your day · Peace of mind protection: Income protection and life assurance to provide financial security for you and your loved ones · Flexible benefits: Tailor your benefits package with a choice of additional options, including private medical insurance and dental cover · Optional benefits: Dental cover, health cash plan, private medical insurance, cycle to work scheme, give as you earn We welcome people from all backgrounds and experiences and are committed to building an inclusive environment where everyone can do their best work. We’re an equal opportunity employer and recognise that everyone brings different strengths and perspectives. If you need any adjustments during the interview process, just let us know - we’re happy to support you. Sponsorship We offer sponsorship for this role. Join the Team at Graphcore Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore brings together deep expertise to solve complex problems and deliver meaningful progress in AI compute. If you want to shape the reliability foundations of advanced AI silicon at Graphcore, we’d love to hear from you. Apply now to be part of the journey.
Graphcore
Salary Range: PLN 260,400 - 352,200 + Benefits + Equity Subject to alignment to the responsibilities and duties of the role. Location: Gdańsk - Hybrid Working Policy - 2-3 Days per week in office About Graphcore At Graphcore, we’re building the future of AI compute.We’re a team of semiconductor, software and AI experts, with deep experience in creating the complete AI compute stack - from silicon and software to infrastructure at datacenter scale.As part of the SoftBank Group, backed by significant long-term investment, we are delivering key technology into the fast-growing SoftBank AI ecosystem.To meet the vast and exciting AI opportunity, Graphcore is expanding its teams around the world.We are bringing together the brightest minds to solve the toughest problems, in a place where everyone has the opportunity to make an impact on the company, our products and the future of artificial intelligence. Job Summary As a Senior QA Engineer within the Management & Observability team, you will be responsible for validating Graphcore's end-to-end telemetry and observability platform. Working closely with Telemetry and Observability engineers, you will design, develop and automate comprehensive test strategies covering telemetry generation, data collection, processing, storage, visualization and alerting. Your work will ensure that Graphcore's observability solutions are reliable, scalable and production-ready for both internal engineering teams and customers. You will contribute throughout the software development lifecycle by defining quality standards, building automated test frameworks and validating distributed systems operating at scale. Responsibilities and Duties Define and implement the end-to-end quality strategy for Graphcore's telemetry and observability platform. Design, develop and maintain automated functional, integration, system and regression tests covering the complete telemetry lifecycle - from telemetry generation to dashboards, APIs and alerting. Develop automated validation frameworks for telemetry pipelines, data quality, metrics, logs, traces and time-series data. Design realistic test environments capable of validating large-scale distributed deployments and production-like workloads. Work closely with software engineers throughout design and implementation to ensure testability, reliability and quality are built into every component. Validate performance, scalability, resilience and fault recovery of telemetry and observability solutions under realistic operating conditions. Integrate automated testing into CI/CD pipelines and continuously improve test coverage, execution time and release quality. Investigate defects through root-cause analysis, working with engineering teams to resolve complex system-level issues. Develop quality metrics, test reports and release readiness criteria to support engineering and product decisions. Contribute to continuous improvement of testing methodologies, automation frameworks and engineering best practices. Skills and Experience BSc or MSc degree in Computer Science, Computer Engineering or equivalent practical experience. 5–8 years of experience in Software QA, Test Automation or Software Engineering. Experience designing automated test frameworks for distributed systems. Experience testing cloud-native or infrastructure software running on Linux. Experience with Python programming. Experience building automated integration and system tests. Familiarity with CI/CD platforms and automated testing pipelines. Experience with Kubernetes, Docker and containerized environments. Understanding of distributed systems, networking and API testing. Experience validating REST and gRPC APIs. Strong debugging and root-cause analysis skills. Excellent written and verbal communication skills. Desirable: Experience testing observability platforms based on Prometheus, Grafana, OpenTelemetry, ClickHouse, Kafka or Elastic Stack. Experience validating telemetry pipelines and large-scale time-series data. Experience with performance, scalability and resiliency testing. Familiarity with Infrastructure as Code technologies such as Terraform or Ansible. Experience testing AI infrastructure, HPC platforms or cloud infrastructure. Experience with one additional programming language such as Go or C++. Knowledge of modern observability practices including metrics, logs and distributed tracing. Benefits In addition to a competitive salary, annual leave policy, medical and dental health plans, a gym card and employee pension (matched up to 4%). We review our benefits on a yearly basis to ensure we offer a valuable and rewarding benefits programme to our employees. We welcome people of different backgrounds and experiences; we’re committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments. Sponsorship Applicants for this position must hold the right to work in the Poland. Unfortunately at this time, we are unable to provide visa sponsorship or support for visa applications.
Graphcore
Silicon Physical Design Engineer Multiple roles across different levels Graphcore is a globally recognised leader in Artificial Intelligence computing systems. The company designs advanced semiconductors and data centre hardware that provide the specialised processing power needed to drive AI innovation, while delivering the efficiency required to support its broader adoption. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. We are opening a new AI Engineering Campus in Bengaluru which will play a central role in Graphcore's work building the future of AI computing. The physical design team sits within the wider silicon design team which includes RTL, verification and DFT. Our work also involves strong links with architecture, packaging and product engineering. We are responsible for working with those teams to create high-quality RTL and building the final chip layout (e.g. GDSII) ensuring a signoff-quality design is delivered to the Foundry (e.g. TSMC). We are looking to hire high-quality silicon physical design engineers to join our team. The successful candidate will support the team with achieving our goals and creating the right engineering solutions. We are a collaborative team and good communication is essential, as is the ability to adapt and learn. For the successful candidate we offer an open, honest and collaborative environment working on leading-edge designs at the most advanced nodes. Our engineers are not siloed, and they are trusted and encouraged to take ownership of their designs and problem solutions. You will be part of a team that looks for improvements to everything we do: our designs, our flows, our methodologies, our infrastructure. Responsibilities and Duties Applicants will be expected to contribute technically to the development of Graphcore's next generation of AI superchips, focusing on achieving robust, high-performance and power-efficient designs in leading-edge process technologies while meeting ambitious development schedules Contributions are expected to span multiple areas and involve: using state-of-the-art EDA tools and in-house Graphcore flows to deliver final designs that are of sign-off quality enhancing existing flows and in-house tools/APIs to support new features and/or methodologies analysis and/or debugging of complex engineering problems leading to workable solutions developing an understanding of emerging technical issues and applying that knowledge to optimise in-house flows and methodologies Candidates will be expected to work closely both with other teams within Graphcore and with 3rd party support engineers/contractors, ensuring good communication between all parties, and to contribute meaningfully to the overall efficiency and success of the Physical Team Essential skills: A Degree in Electronic/Electrical Engineering, Computer Science or related subject with 3+ years of Industry experience Be highly motivated, a self-starter, and a team player Excellent problem-solving skills for debugging issues seen and finding root causes Ability to program/script (required to solve design issues, typically in Tcl and Python) Experience in 7nm or smaller technologies A good breadth of experience with physical design flows including: Floorplanning/Budgeting Synthesis Place and Route Clock CTS Logical Equivalence Timing Analysis Physical Verification (DRC/LVS/ERC) Desirable skills: In depth expertise in one or more aspects of physical design flows Structuring of builds to maximise PPA Silicon transistor knowledge including std cell libraries and/or memories 2D and 3D design (CoWoS, UCIe etc.) High speed, high power and/or full reticle chip design Ethernet, PCIe, LPDDR, HBM interfaces Power integrity and optimization 2nm or 3nm technologies Chip finishing (pad rings, chip level LVS/DRC/ERC) Design for test Team management Project Planning In addition to a competitive salary, Graphcore offers a competitive benefits package. We welcome people of different backgrounds and experiences; we’re committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments.
Graphcore
Senior: PLN 260,400 - 352,200 Staff: PLN 350,700 - 474,400 Subject to alignment to the responsibilities and duties of the role - we currently have multiple positions available at both Senior and Staff level About Graphcore At Graphcore, we’re building the future of AI compute. We’re a team of semiconductor, software and AI experts, with deep experience in creating the complete AI compute stack - from silicon and software to infrastructure at datacenter scale. As part of the SoftBank Group, backed by significant long-term investment, we are delivering key technology into the fast-growing SoftBank AI ecosystem. To meet the vast and exciting AI opportunity, Graphcore is expanding its teams around the world. We are bringing together the brightest minds to solve the toughest problems, in a place where everyone has the opportunity to make an impact on the company, our products and the future of artificial intelligence Job Summary As a Senior Machine Learning Engineer in the Applied AI team at Graphcore, you will contribute to advancing AI technology by developing and optimising AI models tailored to our specialised hardware. You will work on large scale systems where performance is critical to the success of our projects. Working closely with the Software development and Research teams, you will play a critical role in identifying opportunities to innovate and differentiate Graphcore’s technology. We seek engineers with strong technical skills and an understanding of AI model implementation at scale, eager to make a tangible impact in this rapidly evolving field. The Team The Applied AI team’s role is to be proxies for our customers, we need to understand the latest AI models, applications, and software to ensure that Graphcore’s technology works seamlessly with the AI ecosystem and at scale. We build reference applications, contribute to key software libraries e.g. optimising kernels for efficiency on our hardware, and collaborate with the Research team to develop and publish novel ideas in domains such as efficient compute, model scaling and distributed training and inference of AI models for multiple modalities and applications. If you're excited about advancing the next generation of AI models on cutting-edge hardware, we’d love to hear from you! Responsibilities and Duties Implement latest machine learning models and optimise them for performance and accuracy, scaling to 1000s of accelerators. Test and evaluate new internal software releases, provide feedback to software engineering teams, make necessary code fixes, and conduct code reviews. Benchmark models and key ML techniques to identify performance bottlenecks and improve model efficiency. Design and conduct experiments on novel AI methods, implement them and evaluate results. Collaborate with Research, Software, and Product teams to define, build, and test Graphcore’s next generation of AI hardware. Engage with AI community and keep in touch with the latest developments in AI. Candidate Profile Essential: Bachelor/Master's/PhD or equivalent experience in Machine Learning, Computer Science, Maths, Data Science, or related field. Proficiency in deep learning frameworks like PyTorch/JAX. Strong Python or C++ software development skills Expertise in deep learning from model training to optimisation and evaluation. Capable of designing, executing and reporting from ML experiments. Developed deep understanding of performance bottlenecks and how to overcome them. Ability to move quickly in a dynamic environment Enjoy cross-functional work collaborating with other teams. Strong communicator - able to explain complex technical concepts to different audiences. Desirable: Experience in one or more of: MLOps for Kubernetes-based clusters Building production systems with large language models Efficient computing based on low-precision arithmetic. Experience writing C++/Triton/CUDA kernels for performance optimisation of ML models. Experience in distributed training or inference of ML models across 64+ accelerators. Familiarity with HPC systems and networking including Infiniband, NVLink, RoCE technologies. Have contributed to open-source projects or published research papers in relevant fields. Knowledge of cloud computing platforms. Keen to present, publish and deliver talks in the AI community. Benefits In addition to a competitive salary, Graphcore offers annual leave policy, medical and dental health plans, a gym card, and employee pension (matched up to 4%). We review our benefits on a yearly basis to ensure we offer a valuable and rewarding benefits programme to our employees. We welcome people of different backgrounds and experiences; we’re committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments.
Graphcore
About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary We are seeking a Senior Firmware Validation Engineer to support validation and quality assurance for the rack-level firmware stack across Graphcore’s ARM-based server platforms. This role focuses on validating firmware components including SoC firmware (EDK II/UEFI), OpenBMC firmware, rack management services, and platform-level infrastructure used in hyperscale AI server deployments. The Team Graphcore is a globally recognised leader in Artificial Intelligence computing systems. The company designs advanced semiconductors and data centre hardware that provide the specialised processing power needed to drive AI innovation, while delivering the efficiency required to support its broader adoption. The Platform Validation team ensures the reliability and quality of Graphcore’s firmware and system software stack across server nodes and rack-scale AI infrastructure. The team collaborates closely with firmware engineering, silicon teams, hardware engineering, and ODM partners to validate complex platform management stacks and ensure production readiness. Responsibilities and Duties Define and execute validation strategy for rack-level firmware stacks across ARM server platforms. Develop validation plans and automated test frameworks for platform bring-up and firmware lifecycle management. Integrate automated test cases for rack-level firmware components into CI/CD pipelines. Validate firmware update frameworks including signed updates, redundancy mechanisms, and rollback protection. Drive validation of platform security features including Root of Trust, secure boot, and TPM integration. Participate in system-level debugging and root cause analysis across firmware, hardware, and platform integration. Develop automation frameworks and regression testing pipelines supporting firmware validation. Collaborate with silicon vendors, ODM partners, and platform engineering teams during bring-up and manufacturing ramp. Candidate Profile Essential Bachelor’s or Master’s degree in Electrical Engineering, Computer Engineering, Computer Science, or equivalent experience. 6+ years of experience in firmware or platform validation for server or data center systems. Experience validating ARM server firmware stacks including UEFI/EDK II and OpenBMC platforms. Strong understanding of server architecture including power delivery, thermals, networking, and rack infrastructure. Experience validating firmware security features including Root of Trust and secure boot. Strong familiarity with firmware lifecycle management and firmware update frameworks. Experience building automation frameworks and CI/CD pipelines for firmware validation. Desirable Experience validating rack-scale firmware platforms in hyperscale or AI cloud environments. Hands-on experience with EDK II/UEFI validation and OpenBMC system testing. Experience validating firmware for liquid-cooled or high-density server platforms. Experience building hardware-in-the-loop (HIL) or rack-level automated validation environments. Experience validating high-speed interconnects such as PCIe in large-scale deployments. Familiarity with hardware debug tools including JTAG, GDB, and logic analyzers. Experience validating platform management protocols such as Redfish, PLDM, MCTP, and IPMI. USA Benefits In addition to a competitive salary, Graphcore offers flexible working and a comprehensive benefits package designed to support your health, wellbeing and financial future. Our benefits include medical, dental and vision coverage, Flexible Spending Accounts (FSAs), Health Savings Accounts (HSAs), disability and life insurance, a 401(k) retirement plan, commuter benefits, wellness services and an Employee Assistance Programme (EAP). We welcome people of different backgrounds and experiences; we're committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments.
Graphcore
About Graphcore At Graphcore, we’re building the future of AI compute.We’re a team of semiconductor, software and AI experts, with deep experience in creating the complete AI compute stack - from silicon and software to infrastructure at datacenter scale.As part of the SoftBank Group, backed by significant long-term investment, we are delivering key technology into the fast-growing SoftBank AI ecosystem.To meet the vast and exciting AI opportunity, Graphcore is expanding its teams around the world.We are bringing together the brightest minds to solve the toughest problems, in a place where everyone has the opportunity to make an impact on the company, our products and the future of artificial intelligence. Job Summary We are looking for an experienced Silicon Test Engineer to join our Product Test and Diagnosis Department (PTD). This is a pivotal role and will involve building a team of engineers to develop System Level Test (SLT) capability within the company. Working closely with a cross-functional team you will implement SLT tests for a family of next generation AI Processors. The ideal candidate should have a focus on quality and demonstrate a good understanding of the importance of production test on the success of a product. They will have a proven Functional Test or ATE Test Engineering background, and will have a pragmatic, hands-on and flexible approach to a fast-changing environment. The Team The Product Test and Diagnostics team’s role is to detect and manage hardware defects that arise from the manufacture and use of our products. This covers chips, boards and finished systems and takes place both in the manufacturing sites and in the field. Responsibilities and Duties Managing a team of engineers to implement SLT test solutions for a multi-die AI Processor product Working with the Silicon, Test and DFT teams to define manufacturing test solutions Working with key industry leaders to develop an SLT system capable of meeting the products needs Understanding of top, and board level DFT strategies within a 3D, multi-die product Understanding of implementation details for an Incoming Quality Inspection (IQC) stage, and the inter-department complexities Defining test fixtures for SLT Test Solutions: Structural and functional test equipment, interface PCBs, sockets, cooling systems, etc Writing Production, Engineering and Characterisation Test Programs on Advanced SLT platforms Working with other departments to balance test coverage and fault detection across all test stages in the flow; ATE, board, blade, and in-field test Correlation to bench measurements Transferring test solutions to the offshore vendors, as part of an NPI release Work collaboratively with other engineering teams to ensure successful silicon, and board bring-up and help identify any silicon failures to enhance yield learning and improvement Candidate Profile Essential: A pragmatic approach, the ability to improvise, pitch in, and get the job done Experience in production test of large digital devices Very good knowledge of digital integrated circuits and advanced test techniques Very good understanding of DFT components such as (i)JTAG (IEEE 1149.x), BIST, MBIST, etc Understanding of functional test of complex systems Experience in transferring to volume manufacture to OSATs and Contract Manufacturers Ability to work within high pressure cross-department and cross-company teams Ability to mentor junior members of the team Strong scripting and debugging skills using programming languages like Python, TCL, Perl, Shell etc Degree level qualifications in electronics or a related field Good communication skills Desirable Proven competence on SLT testers Experience of thermal control of high-power devices in a production environment. Benefits In addition to a competitive salary, Graphcore offers flexible working and a comprehensive benefits package designed to support your health, wellbeing and financial future. Our benefits include medical, dental and vision coverage, Flexible Spending Accounts (FSAs), Health Savings Accounts (HSAs), disability and life insurance, a 401(k) retirement plan, commuter benefits, wellness services and an Employee Assistance Programme (EAP). We welcome people of different backgrounds and experiences; we're committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments.
Graphcore
About us Graphcore is a globally recognised leader in Artificial Intelligence computing systems. The company designs advanced semiconductors and data centre hardware that provide the specialised processing power needed to drive AI innovation, while delivering the efficiency required to support its broader adoption. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Job Summary Working within the Silicon verification team, the silicon verification engineer is responsible for a wide range of tasks within the silicon verification team. This person is responsible for verification activities within Graphcore, helping the silicon team meet the company objectives for quality silicon delivery. The Team The verification team sits within the Silicon design team. We are responsible for ensuring that the RTL created by the logical design team and used by the physical design team matches the architecture specification for Graphcore silicon. Responsibilities and Duties Verification activities within the verification team Ensuring good communication between sites Verification planning, specification and closure of functional coverage Providing feedback to architects Test generation and failure diagnosis/triage Contributing to shared verification infrastructure Candidate Profile Essential: verification experience in relevant industry Proven leadership and planning skills Be highly motivated, a self starter, and a team player Ability to work across teams and programming languages to find root causes of deep and complex issues Ability to research along with the knowledge to solve complex problems Presents technical and functional knowledge to design experiments/ projects that contribute to overall direction of team Exercises independent judgment in developing methods, techniques and evaluation criteria for obtaining results Highly influential on colleagues with ability to explain difficult concepts Experience of the verification process applied in CPU and/or ASIC environments System Verilog, Python, C++, Linux 12 – 14 years-experience in engineering background Desirable UVM SVA Assembly languages LLVM, GCC DVCS e.g. Git SGE or other DRMS XML and XPath/XSLT Web programming – HTML/DOM, Javascript, SQL
Graphcore
About the job Turn complex manufacturing data into higher yield, stronger quality and better products at scale. As Principal Product Engineer, you will help Graphcore manufacture advanced products with optimised yield, quality and cost. You will guide product performance through launch, production ramp and sustained high volume manufacturing. Your work will influence chip manufacturing, test, PCBA and system level production. You will help teams find the signals that improve reliability, reduce cost and protect quality. You will partner with OSATs, test engineering, manufacturing, quality, data engineering and analytics teams. Together, you will release new products and solve difficult yield, test and manufacturing challenges. This role suits someone who enjoys deep technical ownership and practical impact. You will turn manufacturing data into action across silicon, boards, blades and rack scale systems. The team and culture You will join the Product Test and Diagnosis team within Manufacturing Operations. The team spans Bristol, Cambridge, India, the US and Taiwan, working close to products and production partners. Work happens through direct ownership, technical depth and clear decisions. You will define best practice, challenge assumptions and move quickly when yield or quality needs attention. The team connects silicon, board level assemblies, server blades and rack scale systems. Decisions are grounded in data, engineering judgement and accountability for manufacturing outcomes. What we're looking for · Strong semiconductor product engineering, test engineering or manufacturing engineering capability in high volume production · Deep understanding of semiconductor, PCBA and system level manufacturing and testing · Strong statistical analysis skills, including SPC, DOE, regression, hypothesis testing and six sigma methods · Proven ability to analyse and visualise large manufacturing datasets using Python, SQL, JMP or similar tools · Experience driving yield, quality or cost improvements across complex manufacturing processes · Ability to work across test, manufacturing, quality, data and external partner teams with clear technical ownership While we have outlined a set of requirements, we value transferable skills and diverse experiences. If you meet most of our essential criteria, we encourage you to submit an application and showcase how your background makes you a strong candidate. Benefits · Flexible working: Balance your work and personal life with greater flexibility · Generous leave: Take time to rest, recharge and enjoy life outside of work · Retirement planning support: Up to 5% matched pension · Phantom equity: Share in Graphcore’s success · Workplace experience: Enjoy thoughtfully designed office spaces for collaboration, with free food and an on-site barista to support your day · Peace of mind protection: Income protection and life assurance to provide financial security for you and your loved ones · Flexible benefits: Tailor your benefits package with a choice of additional options, including private medical insurance and dental cover · Optional benefits: Dental cover, health cash plan, private medical insurance, cycle to work scheme, give as you earn We welcome people from all backgrounds and experiences and are committed to building an inclusive environment where everyone can do their best work. We’re an equal opportunity employer and recognise that everyone brings different strengths and perspectives. If you need any adjustments during the interview process, just let us know - we’re happy to support you. Sponsorship Applicants must have the legal right to work in the UK. Unfortunately, we are unable to provide visa sponsorship or support visa applications for this role. Join the Team at Graphcore Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore brings together deep expertise to solve complex problems and deliver meaningful progress in AI compute. If you want to turn manufacturing insight into better products at scale, we’d love to hear from you. Apply now to help shape product engineering at Graphcore.
Graphcore
Principal Post Silicon Validation Engineer (Bringup) Salary $241,100 - $326,100 + Phantom Equity + Benefits Graphcore is a globally recognised leader in Artificial Intelligence computing systems. The company designs advanced semiconductors and data centre hardware that provide the specialised processing power needed to drive AI innovation, while delivering the efficiency required to support its broader adoption. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. We are opening a new AI Engineering Campus in Austin which will play a central role in Graphcore's work building the future of AI computing. We are looking to hire Post-Silicon Validation Engineers to join our collaborative, cross-functional development team validating cutting edge, high performance AI chips and platforms. You will play a critical role in supporting new product introductions and post-silicon validation. Working within the Post-Silicon Validation team, you will be involved with bringing first silicon to life, functionally validating it and working closely with many other teams to help it become a fully characterised and working product, reporting project status/progress to program management on a regular basis. You will have the opportunity to, and be responsible for, leading, mentoring, and providing technical guidance to other engineering team members. In this role, you can leverage your experience and industry knowledge to architect and drive implementation of continuous improvements to test infrastructure and processes. The Post-Silicon Validation team sits within the Architecture and Validation team, we are responsible for validation of new silicon when it returns from manufacture, enabling and supporting the production SW and FW teams to bring up their software and also supporting the Silicon Characterisation team. Responsibilities and Duties Plan, design, develop and debug silicon validation tests in bare metal C/C++ on FPGA/Emulator prior to first silicon Deploy silicon validation tests on first silicon and debugging them Develop automated test framework and regression test suites in Python to optimize validation efficiency Collaborate closely with engineers from many other disciplines on a variety of topics Work with Validation and Production Test engineering peers to implement best practices and continuous improvements to test methodologies Analyse test results, identify and debug failures/defects Contribute to shared test and validation infrastructure Provide feedback to architects Essential skills: Strong experience in Bare metal / embedded C/C++ Good knowledge of digital ASICs Be highly motivated, a self starter, and a team player Ability to work across teams and programming languages to find root causes of deep and complex issues Experience of the post-silicon validation process applied in digital ASIC environments Python, Linux Excellent communication skills and the ability to collaborate with others to solve problems Excellent problem-solving, analytical & diagnostic skills Desirable skills: Driver level experience with one or more of the following is highly desirable: PCIe Ethernet Memory technologies (LPDDR, DDR, HBM, …) Other peripherals such as I2C, I3C, SPI, … Good knowledge of mixed-signal building blocks such as PLLs, high speed PHYs and IC control/communication protocols is highly desirable Experience of Arm CPUs, System IP and debug tools Experience of AMBA protocols Understanding of ML applications and their workloads Experience in Characterization, Failure Analysis, Test Development, Statistical analysis, and Customer Support USA Benefits In addition to a competitive salary, Graphcore offers flexible working and a comprehensive benefits package designed to support your health, wellbeing and financial future. Our benefits include medical, dental and vision coverage, Flexible Spending Accounts (FSAs), Health Savings Accounts (HSAs), disability and life insurance, a 401(k) retirement plan, commuter benefits, wellness services and an Employee Assistance Programme (EAP). We welcome people of different backgrounds and experiences; we're committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments.
Graphcore
Graphcore is a globally recognized leader in Artificial Intelligence computing systems. The company designs advanced semiconductors and data centre hardware that provide the specialized processing power needed to drive AI innovation, while delivering the efficiency required to support its broader adoption. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. We are opening a new AI Engineering Campus in Austin, which will play a central role in Graphcore's work building the future of AI computing. RESPONSIBILITIES Responsible for the maintaining all the network fabrics and network topologies operating at peak performance on Graphcore Data centers housing AI server platforms. Collaborate closely with our Network architecture team through all phases of the design and development lifecycle — from concept to deployment — ensuring timely, high-quality, and stable high-speed network fabrics. Primary support of deployed internal fabrics in our AI Datacenters. Direct escalation for Hardware, Link and, Performance issues on network fabrics. Escalation support for customer network fabrics (aligned to reference arch.) Triage day-to-day network topologies (scale-up, scale-out, and front-end networks) issues of Fabrics deployed in Graphcore Datacenters. REQUIREMENTS Bachelor’s, master’s degree or equivalent experience in Network Engineering/Security, Information Technology, Computer Science, Computer engineering, or a related field. 9+ years of hands-on experience in network engineering and 3+ years supporting networking on AI or Hyperscale Datacenters. Knowledge of Remote Direct Memory Access (RDMA) and its implementations, specifically ROCEv2 or InfiniBand Knowledge of Lossless network – Experience with congestion control and lossless Ethernet technologies including PDF, ECN and DCQCN. Versed in High-Bandwidth network infrastructure including experience with operating 400G and 800G network fabrics and tackling challenges at the physical layer for networks operating at these speeds. This includes familiarity with specific optical transceivers. Experience sustaining non-blocking, multi-tier CLOS networks (i.e., Spine / Leaf / Super-Spine) that are optimized for high-density GPU clusters. Extensive troubleshooting experience for data center routing protocols such as BGP and OSPF as well as troubleshooting of VXLAN implementations. Cross-Functional collaboration- The candidate must excel at working closely with network architects, systems hardware engineers, AI software engineers, and Datacenter infrastructure teams to optimize the full stack from the application layer down to the Network infrastructure that includes controllers (NICs, SmartNICs, DPUs), network switches, and cabling and optical transceiver operation in our Datacenter. DIFFERENTIATORS Deep hands-on experience with NetDevOps to automate tasks using tools such as Terraform or Ansible to manage network configurations and state. Possess strong scripting skills such as Python, Go, or JSON to build custom automation, interact with APIs, and develop internal network tools. Deep hands-on experience with modern, high-performance hardware from network switch and NIC providers such as Broadcom, Arista, Cisco, or Pensando. Knowledgeable in Advanced Network Telemetry, beyond traditional SNMP. Knowledgeable in telemetry protocols such as gNMI, Sflow, Netflow, Redfish, and Prometheus to characterize and detect transient congestion events, micro-bursts in the network fabric or defective network hardware. We welcome people of different backgrounds and experiences and are committed to building an inclusive work environment that makes Graphcore a great home for everyone. We are an equal opportunity employer and want to build a work environment where everyone is happy, productive and respectful so they can do their best work. If you have a disability or additional need that requires accommodation, just let us know. pproach to interview and encourage you to chat to us if you require any reasonable adjustments. In addition to a competitive salary, Graphcore offers flexible working and a comprehensive benefits package designed to support your health, wellbeing and financial future. Our benefits include medical, dental and vision coverage, Flexible Spending Accounts (FSAs), Health Savings Accounts (HSAs), disability and life insurance, a 401(k) retirement plan, commuter benefits, wellness services and an Employee Assistance Programme (EAP). We welcome people of different backgrounds and experiences; we're committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments.
Graphcore
About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software, and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers, and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary Reporting to the Validation leadership team, the Principal Execution and Quality Validation Engineer will be responsible for driving validation execution, automation, product quality, and release readiness across Graphcore silicon and platform technologies. The role combines deep expertise in hardware validation, test automation, and quality engineering with a strong focus on execution excellence. Working closely with architecture, design, verification, firmware, software, systems engineering, and validation teams, the successful candidate will develop scalable validation methodologies, improve test coverage and execution efficiency, and ensure products meet the highest standards of functionality, reliability, and performance before customer deployment. As a recognized technical leader within the validation organisation, this role will influence validation strategies, automation roadmaps, and quality practices across multiple projects and engineering disciplines. The Team The Validation Execution and Quality team sits within the Validation organisation and is responsible for improving validation effectiveness, test execution efficiency, product quality, and release readiness across Graphcore silicon and platform products. The team develops validation methodologies, automation frameworks, test execution infrastructure, and quality processes that enable engineering teams to execute validation activities at scale. Working closely with validation, architecture, firmware, software, systems, and design teams, the group helps ensure products are thoroughly validated before release. Responsibilities and Duties Define and drive validation strategies for complex silicon, subsystem, and platform technologies, ensuring comprehensive coverage across functional, performance, reliability, and stress testing Leverage advanced hardware knowledge to plan, optimize, verify, and test critical electronic systems, including silicon, FPGA-based platforms, test systems, and associated validation infrastructure Develop and execute validation plans that ensure products meet quality, reliability, and performance objectives before release Create detailed test plans from high-level validation requirements and test cases, ensuring timely execution and high-fidelity results Design, develop, and maintain automation frameworks that enable rapid execution of test plans, regression suites, and validation activities Lead silicon and platform validation activities from bring-up through product readiness Conduct complex analysis of validation results and system behaviour to identify issues and opportunities for improvement Collaborate with architecture, design, verification, firmware, software, systems engineering, and validation teams to implement new validation requirements and improve test coverage Integrate new validation capabilities, methodologies, and test solutions into existing validation infrastructure to improve productivity, quality, and execution efficiency Develop automated data collection, analysis, and reporting solutions to improve engineering visibility and decision-making Evaluate complex design features to identify functional, compatibility, integration, and reliability risks Investigate and drive resolution of complex validation failures through structured debug and root-cause analysis methodologies Assess validation techniques, workflows, and engineering processes, introducing automation and process improvements to maximize productivity Support power, performance, and system-level validation activities through development of scalable validation infrastructure and methodologies Influence validation roadmaps, infrastructure development, and engineering best practices across the validation organisation Act as the final validation authority for assigned areas, ensuring issues are identified and resolved before products reach customers Drive continuous improvement initiatives that improve validation quality, execution efficiency, and engineering productivity Mentor engineers on validation methodologies, automation frameworks, test execution strategies, and quality practices Candidate Profile Essential: Strong experience in silicon validation, hardware validation, system validation, or related engineering disciplines Deep understanding of validation methodologies, test planning, test execution, and quality processes Strong experience developing automation frameworks for validation, regression execution, and test orchestration Expert-level Python programming skills for automation, data analysis, and infrastructure development Strong understanding of hardware systems, silicon bring-up, and validation workflows Experience translating high-level requirements into structured validation plans, test cases, and measurable quality outcomes Experience analysing validation results and driving root-cause investigations for complex technical issues Strong understanding of validation coverage, quality metrics, release readiness criteria, and risk assessment methodologies Experience collaborating across architecture, design, verification, firmware, software, systems, and validation teams Excellent analytical, diagnostic, and problem-solving skills Exceptional communication skills and the ability to influence technical direction across multidisciplinary engineering teams Ability to independently lead complex technical initiatives and drive adoption of new validation methodologies and tools Desirable Experience with post-silicon bring-up and system integration activities Experience developing large-scale regression systems and automated test execution infrastructure Experience with FPGA-based validation environments Experience automating laboratory equipment and validation environments Familiarity with CI/CD systems and continuous testing methodologies Experience developing dashboards, reporting systems, or engineering analytics tools Experience with C, C++, or embedded software development Experience working on AI accelerator, high-performance compute, or semiconductor products Experience defining organisational validation standards, quality metrics, and engineering best practices
Graphcore
About Us Graphcore is a globally recognised leader in Artificial Intelligence computing systems. The company designs advanced semiconductors and data centre hardware that provide the specialised processing power needed to drive AI innovation, while delivering the efficiency required to support its broader adoption. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. We are opening a new AI Engineering Campus in Austin, which will play a central role in Graphcore's work building the future of AI computing. Job Overview: Graphcore is seeking a Principal Engineer, Power Engineering to aid in defining the architecture of the power delivery components from grid to chip. You will work closely with multiple cross-functional teams, including hardware engineers, firmware developers, and data center operations, to ensure the infrastructure we build meets customer requirements as well as attaining standards of performance, reliability, and scalability. This role will be based in Austin or San Jose and could potentially be performed onsite or remotely within the US. Responsibilities: Provide thought leadership, planning, and design for power delivery from data center power drop through delivery to systems & chips Establish power roadmap. Work with vendors to ensure their planned future technologies are pulled into our system engineering roadmap AND ensure that our (system engineering) required technological innovation is being driven into their (power vendor) roadmap. Work to improve reliability and efficiency of power engineering component over time. Rack-level power components. Competency with specification of hardware and firmware for rack-level power shelves. Competency with design for resiliency. Set power policies in the data center for: power buffer/oversubscription, scheduling of jobs based on power, and monitoring (power quality faults) / accommodation faults. Resolve telemetry and data needed to support power policy algorithms. Ensure robustness of IT power in the presence of grid disturbances and make sure that IT power loads are compatible with grid power requirements (“good grid power citizen”) Board and system-level power regulation components. Understand and able to apply current VRM, TLDR, and POL technologies. Help to develop new technologies to increase efficiency. Understand manufacturing capabilities and limitations of new technology. Drive Provide guidance to the power equipment manufacturers, ODM and engineering teams on the required testing and validation procedures for power hardware and firmware components. Develop and test hardware prototypes as needed to support the exploration and validation of new concepts in power delivery. Tackle and resolve sophisticated power hardware and firmware issues as needed. Ensure compliance with industry standards and regulations. Required Skills and Experience : Bachelor’s or Master’s or PhD degree in Electrical Engineering with a specialized knowledge of Power Engineering. demonstrated ability in power system and delivery architecture, design, and development. Experience with rack-level hardware design, including servers, storage, networking, and power distribution. Proficiency in power hardware and firmware development and debugging. Excellent problem-solving and analytical skills. Strong communication and teamwork skills. Ability to work in a fast-paced, multifaceted environment. "Nice to Have" Skills and Experience: Familiarity with new technologies in AI and data center infrastructure. Comfortable meeting with and engaging directly with customers (internal and external) during requirements gathering and solution development Benefits In addition to a competitive salary, Graphcore offers a competitive benefits package. We welcome people of different backgrounds and experiences; we’re committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments.
Graphcore
About us Graphcore is one of the world’s leading innovators in artificial intelligence compute. We are developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and support the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of a family of companies responsible for some of the world’s most transformative technologies. Together, we share a bold vision to enable advanced artificial intelligence and ensure its benefits are accessible to everyone. Graphcore brings together AI researchers, silicon designers, software engineers and systems architects to solve complex technical challenges and deliver innovative computing solutions. Job Summary The Principal Electrical Engineer will be a technical authority within Data Center Engineering, leading the architecture and delivery of safe, resilient and scalable electrical infrastructure for high-density AI computing environments. Working with internal teams, data center developers, utilities, consultants and equipment partners, this role will guide projects from early technical studies through design, construction, commissioning, operation and lifecycle improvement. The successful candidate must reside in, or be willing to relocate to, Austin, Texas. Approximately 10% travel may be required. The Team The Data Center Engineering team is responsible for defining and enabling the infrastructure needed to deploy and operate Graphcore’s computing systems at scale. The team works across electrical, mechanical, thermal, controls, systems and operational disciplines, collaborating with external engineering and construction partners to deliver reliable, efficient and maintainable data center environments. Responsibilities and Duties Act as the technical authority for electrical engineering across data center infrastructure projects, from the utility or on-site power source through to the IT rack. Lead electrical architecture development through conceptual studies, detailed design, construction, commissioning, operation and lifecycle improvement. Develop electrical standards, specifications, reference designs, qualification plans and acceptance criteria for power distribution and conversion systems. Define electrical protection and safety approaches for AC and DC systems, including fault detection, interruption, isolation, grounding, stored energy, insulation monitoring, lockout and safe maintenance. Develop resilient power solutions for highly dynamic AI and high-performance computing loads, including ride-through, short-duration energy buffering, backup generation, energy storage and utility interconnection. Evaluate established and emerging power conversion, protection, control and energy-storage technologies, making clear recommendations based on safety, reliability, efficiency and maintainability. Lead technical due diligence, design reviews, failure-mode assessments, factory and site acceptance testing, site inspections and commissioning activities. Direct engineering consultants and technology partners, ensuring that designs, calculations, specifications and construction documentation meet project requirements. Develop and manage owner’s project requirements and review consultant deliverables, submittals, requests for information and proposed technical changes. Collaborate with construction partners during the delivery of laboratory and data hall spaces, supporting site reviews, technical issue resolution and project risk management. Partner with teams across IT, silicon, mechanical, thermal, controls, operations and sustainability to resolve cross-disciplinary design challenges. Support operational readiness, incident investigation and root-cause analysis, identifying practical improvements to reliability, safety and system performance. Engage effectively with authorities, utilities and relevant industry bodies to support compliant project delivery. Maintain current knowledge of developments in AI and high-performance computing infrastructure, electrical equipment, engineering standards, codes and regional utility requirements. Mentor engineers and provide technical guidance across projects without direct people-management responsibility. Candidate Profile Essential Bachelor’s or master’s degree in Electrical Engineering, a closely related discipline, or equivalent relevant experience. Substantial experience delivering mission-critical electrical infrastructure in hyperscale data centers, colocation facilities, semiconductor facilities, utilities or comparable high-availability environments. Demonstrated ability to lead a significant technical area, make independent engineering decisions and influence outcomes across multiple disciplines and external partners. Deep knowledge of medium- and low-voltage AC systems and high-power DC distribution systems. Strong understanding of power electronics and conversion technologies, including rectifiers, inverters, isolated and non-isolated DC/DC conversion, controls and galvanic isolation. Strong knowledge of electrical fault behaviour and protection, including grounding, insulation monitoring, selective coordination, interruption, arc hazards, stored energy and safe maintenance. Broad technical knowledge of utility and medium-voltage systems, transformers, switchgear, uninterruptible power supplies, generators, battery energy storage, power distribution units, busway, monitoring and controls. Experience designing electrical interfaces for high-density computing or other highly dynamic mission-critical loads. Proficiency with ETAP or SKM PowerTools, together with experience using an electromagnetic-transient or power-electronics simulation environment such as PSCAD, EMTP-RV, PLECS, PSIM or MATLAB/Simulink. Strong familiarity with applicable North American and international electrical requirements, including relevant NFPA, IEEE, IEC, UL, CSA and local standards. Experience directing consultants, reviewing technical deliverables and supporting construction, testing and commissioning activities. Excellent communication and influencing skills, with the ability to explain complex engineering decisions clearly to technical and non-technical stakeholders. Sound engineering judgement and the ability to lead cross-functional work in a fast-moving environment with competing priorities. Willingness to travel approximately 10% as required. Desirable Professional Engineer licence or equivalent international professional certification. Experience with behind-the-meter generation, microgrids or other on-site energy systems. Experience developing, testing or deploying high-power DC data center architectures. Hands-on experience with emerging transformer, converter or electrical-protection technologies. Experience integrating battery energy storage, supercapacitors, fuel cells, renewable generation or demand-response capabilities. Experience designing Tier III or Tier IV data centers, or infrastructure with comparable availability requirements. Participation in relevant industry standards groups or technical forums. Experience with reliability engineering, FMEA or FMECA, functional safety, controls, telemetry or operational-technology cybersecurity. Familiarity with engineering and design tools such as Visio, Bluebeam, AutoCAD or Revit. In addition to a competitive salary, Graphcore offers flexible working and a comprehensive benefits package designed to support your health, wellbeing and financial future. Our benefits include medical, dental and vision coverage, Flexible Spending Accounts (FSAs), Health Savings Accounts (HSAs), disability and life insurance, a 401(k) retirement plan, commuter benefits, wellness services and an Employee Assistance Programme (EAP). We welcome people of different backgrounds and experiences; we're committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments.
Graphcore
About Graphcore At Graphcore, we’re building the future of AI compute.We’re a team of semiconductor, software and AI experts, with deep experience in creating the complete AI compute stack - from silicon and software to infrastructure at datacenter scale. As part of the SoftBank Group, backed by significant long-term investment, we are delivering key technology into the fast-growing SoftBank AI ecosystem.To meet the vast and exciting AI opportunity, Graphcore is expanding its teams around the world.We are bringing together the brightest minds to solve the toughest problems, in a place where everyone has the opportunity to make an impact on the company, our products and the future of artificial intelligence. Job Summary We are looking for an experienced Principal Engineer to join our Cloud Platform Team and help develop and deploy clouds and services. Working closely with our colleagues in Software Platform, Datacentre Operations and Product Development teams, you will deploy services on our fleet of cutting-edge AI systems. As part of our Software Platform organisation, you will be involved in the cloud integration, validation, performance benchmarking, optimisation, and development of our high-performance AI solutions. These include in-house AI systems alongside off-the-shelf high-performance servers, switches and storage solutions. This is a hand-on technical role requiring a solid background in the use of cloud infrastructure, deployment using Infrastructure-as-Code, observability, high-performance networking and storage systems. You may have been working in an IT organisation, a datacentre, a cloud provider or as a developer of orchestration or cloud services. The Software Platform team at Graphcore We build Graphcore products into large-scale AI solutions for our customers and the Cloud Platform Team is responsible for providing such systems to both internal users via private clouds and customers via our own public clouds. Often the internal systems will be using and developing pre-release hardware and software, so it’s vital you are comfortable with unproven components. Responsibilities and Duties Operate and extend existing OpenStack-based cloud services and contribute to the deployment and development of new ones. You will be responsible for significant technical initiatives and projects, mentoring a small team of more junior engineers in best engineering practices. Develop and operate end-user services on our clouds and support internal users in their use. You will turn end-user and product requirements into deployed services. Help to build automation to collect and analyse metrics and other observability data from the cloud services to support clear identification and reporting of any issues. Work with users to provide information of any product-related issues to Engineering and QA departments. Work with our Datacentre Operations Engineers to maintain and operate the fleet of AI systems at peak performance in our private clouds. Configure and test new Graphcore AI hardware and systems using Continuous Deployment and Infrastructure-as-code in internal and external datacentres. Drive corrective actions for systems that are not operating correctly, working with DC operations and Graphcore Engineering as required. Work with external vendors of off-the-shelf switches, servers and storage solutions to specify, benchmark and integrate 3rd party products into our Cloud Reference Design. Skills and Experience [ALL REQUIRED] Bachelor's degree or equivalent practical experience in a relevant subject. Solid infrastructure or IT experience with a proven track record of delivering technical output as an individual contributor. Experience managing or operating on-premises or private-cloud environments. Experience specifying, scoping, estimating and detailing work plans in an AGILE and SCRUM framework, including priorities, risks, issues, impacts and constraints. Expert-level, proven Linux scripting ability (bash and python required). Expert-level, proven Linux system administration (Ubuntu, RHEL and variants). Experience with a version control system (preferably Git) and using it to manage system configuration or automation. Experience with Continuous Integration or testing pipelines using GitLab, GitHub or similar. Hands-on experience deploying services into public or private clouds using Infrastructure-as-Code (IAC). A solid understanding of the technologies underpinning cloud services (APIs, virtualisation of CPUs, IO, systems), virtual networks, block storage, resource management and monitoring. Expert with IAC automation tools (e.g. Terraform/OpenTofu, Ansible, Packer). Experience with container deployment and management tools (e.g. docker, podman, apptainer). Experience with solutions for monitoring and observability. e.g. Grafana, Prometheus, OpenSearch/ElasticSearch, Loki, Mimir, OpenTelemetry, Fluentd ,Kafka Excellent communication and presentation skills, and experience dealing with end-users of IT or cloud services. An ability to work independently and lead others on critical infrastructure without oversight, and with a focus on end-user availability. Desirable but not required: Experience with OpenStack deployments or the technologies they rely on (e.g. Ceph, Open vSwitch, KVM, QEMU ). Experience with High Performance Computing (HPC) environments using SLURM or similar batch workload solutions. Strong skillset and experience in end-to-end deployment automation and CI of containerised services. Complete automation of pipelines for build, test, deploy, manage, alert, destroy, rebuild. Experience with managing production Kubernetes clusters and workloads. Experience with workload queue management systems (SLURM, LSF, Kueue). Experience with managed switch configuration (e.g. EOS, SONiC, DNOS). Programming experience with Python3 utilising classes and inheritance. Programming experience with Go. Benefits In addition to a competitive salary, Graphcore offers flexible working, a generous annual leave policy, private medical insurance and health cash plan, a dental plan, pension (matched up to 5%), life assurance and income protection. We have a generous parental leave policy and an employee assistance programme (which includes health, mental wellbeing, and bereavement support). We offer a range of healthy food and snacks at our central Bristol office and have our own barista bar! We welcome people of different backgrounds and experiences; we’re committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments. Sponsorship Applicants for this position must hold the right to work in the UK. Unfortunately at this time, we are unable to provide visa sponsorship or support for visa applications.
Graphcore
At Graphcore, we’re building the future of AI compute. We’re a team of semiconductor, software and AI experts, with deep experience in creating the complete AI compute stack - from silicon and software to infrastructure at datacenter scale. As part of the SoftBank Group, backed by significant long-term investment, we are delivering key technology into the fast-growing SoftBank AI ecosystem.To meet the vast and exciting AI opportunity, Graphcore is expanding its teams around the world.We are bringing together the brightest minds to solve the toughest problems, in a place where everyone has the opportunity to make an impact on the company, our products and the future of artificial intelligence. Job Summary We're looking for a Linux Engineering Lead to help shape and operate the Linux platforms that underpin our engineering environments. This is a hands-on technical leadership role. You'll lead a small team of engineers while driving automation, reliability and operational excellence across the Linux infrastructure used by Graphcore's engineering organisation. You will be responsible for leading incident response, driving operational improvements, and setting standards for how Linux systems are managed and supported across the organization. While the role includes leadership responsibilities, it will initially require a hands-on approach, including direct involvement in troubleshooting, system support, and automation efforts, while building team capability and scaling processes. Collaborating intimately with engineering groups, platform engineers, and infrastructure experts, you will guarantee systems stay stable, efficient, and consistent with changing business and product delivery requirements. The Team You’ll be joining a multi-disciplinary team with strong technical skills and a very supportive culture. We work closely together, regularly share knowledge, and your skills will make a direct impact on our business. It’s an exciting and pivotal moment for us right now, with plenty of new projects ahead. If you're looking to solve interesting problems and see your work deliver real-world results, this is the team for you! Responsibilities and Duties Lead and mentor a team of Linux engineers Own the reliability, performance and scalability of engineering Linux environments Drive adoption of Infrastructure-as-Code and GitOps practices Build automation that reduces operational overhead and improves consistency Lead major incident response and root cause analysis activities Partner with software, platform and infrastructure teams to support evolving engineering requirements Establish standards, tooling and processes that enable systems to scale efficiently Improve observability, monitoring and operational visibility across the estate Help define the future direction of Linux platform engineering at Graphcore Candidate Profile Essential Significant experience administering Linux environments at scale Strong troubleshooting skills across systems, networking, storage and applications Experience leading engineers, projects or operational initiatives Strong automation and scripting skills (Python, Bash or similar) Experience with Infrastructure-as-Code and configuration management tools such as Ansible, Terraform or Puppet Experience working with Git-based workflows and CI/CD pipelines Experience managing production incidents and driving operational improvements Excellent communication and stakeholder management skills Desirable Experience supporting AI, HPC or large-scale engineering environments Experience with observability platforms and monitoring systems Experience working alongside platform engineering, SRE or DevOps teams Knowledge of identity and access management Experience building or scaling operational processes We welcome people of different backgrounds and experiences; we’re committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments.
Graphcore
About Graphcore At Graphcore, we’re building the future of AI compute. We’re a team of semiconductor, software and AI experts, with deep experience in creating the complete AI compute stack - from silicon and software to infrastructure at datacenter scale. As part of the SoftBank Group, backed by significant long-term investment, we are delivering key technology into the fast-growing SoftBank AI ecosystem.To meet the vast and exciting AI opportunity, Graphcore is expanding its teams around the world.We are bringing together the brightest minds to solve the toughest problems, in a place where everyone has the opportunity to make an impact on the company, our products and the future of artificial intelligence. Job Summary Join our dynamic Software Infrastructure team and take a pivotal role in scaling and managing our infrastructure. You will develop essential tools and services that empower our broader software team. Your contributions will enhance the build, test, deployment, and productisation processes of our Machine Learning Software components. Work with our High-Performance Computing (HPC) AI platforms and gain invaluable experience in distributed systems The Team The Software Infrastructure team provides critical platforms and services for software development teams across the business. Our responsibilities include managing the CI platform and services, build engineering, component integration, and packaging and release systems. We operate in squads, fostering a culture of service ownership and empowerment for our engineers. We focus on long-term engineering solutions and strive to eliminate toil wherever possible. Responsibilities and Duties Develop, own, and maintain tools and services to support AI research and engineering teams Deploy and maintain services with Kubernetes and Docker Manage our Cloud Infrastructure using tools such as Terraform Candidate Profile Essential: Knowledge of Python Familiarity with cloud services (e.g. AWS) Experience managing or developing in Linux environments Understanding of CI/CD principles Experience using Kubernetes (k8s) Experience of one of the following: maintaining machine learning applications. deploying ML orchestration tools (e.g. NV Ray, KFP, SkyPilot). managing ML accelerator hardware (e.g. DCGM). Desirable Experience with Infrastructure as Code (IaC) tools (e.g. Terraform/OpenTofu) Experience with GitHub Actions Experience with modern observability tooling (e.g. Prometheus) Experience with Grafana Knowledge of Go/Java/C++ (or similar language) Benefits In addition to a competitive salary, Graphcore offers flexible working, a generous annual leave policy, private medical insurance and health cash plan, a dental plan, pension (matched up to 5%), life assurance and income protection. We have a generous parental leave policy and an employee assistance programme (which includes health, mental wellbeing, and bereavement support). We offer a range of healthy food and snacks at our central Bristol office and have our own barista bar! We welcome people of different backgrounds and experiences; we’re committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments.
Graphcore
Graphcore is a globally recognised leader in Artificial Intelligence computing systems. The company designs advanced semiconductors and data centre hardware that provide the specialised processing power needed to drive AI innovation, while delivering the efficiency required to support its broader adoption. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. We are expending our labs in Bristol which will play a central role in Graphcore's work building the future of AI computing. We are looking for a Validation and Debug engineer to work with Architecture, Silicon Engineering, Hardware Engineering and the Firmware team to help create and execute the bring-up and debug of our cutting-edge system platforms. As such you will need to develop a detailed understanding of our silicon and platform products. Your role will be to create and execute Validation and Debug plans of our silicon devices and platforms in systems to show that they will operate correctly over all conditions in the final product. This Validation will also involve debugging of hardware and firmware. You will concentrate on mainly manual testing of systems in the lab with the aid of some automation. Responsibilities and Duties Perform Validation and Debug for new hardware platforms. Debug PCB, power, signal integrity, and interface issues. Validate and debug high-speed memory interfaces and slow speed interfaces LPDDR5-9600 memory interfaces SPI/I2C/I3C/UART Support LPDDR5-9600 initialization, training, timing optimization, margin analysis, and stability validation. Analyse DDR timing margins, signal integrity, eye diagrams, and training behaviour. Identify root causes and drive corrective actions across hardware and firmware domains. Create bring-up documentation, test procedures, and debug reports. Working closely with: PCB design engineers Power engineers Signal integrity engineers Embedded software teams Essential skills Hardware Debug LPDDR validation and tuning Signal integrity analysis DDR training and margin analysis Protocol-level debugging Root-cause analysis Low-level system integration An ability to work independently without daily oversight. Independently performs board bring-up and subsystem debug Drives issue resolution across teams. Use advanced lab instrumentation including: High-bandwidth oscilloscopes BERT systems VNA/TDR equipment Hands-on experience with: LPDDR5 memory subsystems LPDDR5 compliance testing Desirable skills: Knowledge of PCBA and system level technologies Ethernet validation PCIe Background knowledge of ATE systems and capabilities. You may not be directly involved with such testing. The ability to code or script automation and data analysis using appropriate coding languages such as Python, LabView and BASH. Benefits: In addition to a competitive salary, Graphcore offers a competitive benefits package. We welcome people of different backgrounds and experiences; we’re committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments.
Graphcore
At Graphcore, we’re building the future of AI compute. We’re a team of semiconductor, software and AI experts, with deep experience in creating the complete AI compute stack - from silicon and software to infrastructure at datacenter scale. As part of the SoftBank Group, backed by significant long-term investment, we are delivering key technology into the fast-growing SoftBank AI ecosystem.To meet the vast and exciting AI opportunity, Graphcore is expanding its teams around the world.We are bringing together the brightest minds to solve the toughest problems, in a place where everyone has the opportunity to make an impact on the company, our products and the future of artificial intelligence. Job Summary Our Engineering Labs are where new silicon, systems, and platforms are brought to life, tested, and scaled. We’re looking for an experienced Engineering Lab Infrastructure Lead to help us build and operate the technical foundations that support this work. You will lead the Engineering Lab Support function, ensuring our labs, infrastructure, and services remain reliable, scalable, and effective for the engineers developing Graphcore’s next generation technologies. Initially, this is a highly hands-on role. You’ll work directly with engineering teams, supporting lab environments, managing infrastructure, automating workflows, and solving complex technical problems. As the function grows, you’ll help build and lead a small team while defining the processes, standards, and service model that will support the organisation long term. You’ll work closely with silicon, hardware, systems, and software engineering teams, making a direct impact on the speed and effectiveness of product development. The Team You’ll be joining a multidisciplinary team with strong technical skills and a very encouraging culture. We work closely together and regularly share knowledge, and your skills will make a direct impact on our business. It’s an exciting and pivotal moment for us right now, with plenty of new projects ahead. If you're looking to solve interesting problems and see your work deliver real-world results, this is the team for you. Responsibilities and Duties Leading the development of Engineering Lab infrastructure and support services Acting as the technical escalation point for complex infrastructure and lab issues Managing Linux-based servers and engineering environments Supporting hardware bring-up, validation, and testing activities Designing and improving operational processes, tooling, and automation Maintaining and improving network, storage, and compute infrastructure within engineering labs Managing infrastructure through configuration management and Infrastructure-as-Code practices Building strong relationships with engineering teams and understanding their evolving requirements Developing knowledge bases, documentation, and operational standards Recruiting, mentoring, and leading a small team of Lab Infrastructure Engineers as the function grows Essential Strong Linux systems administration experience across Debian and/or Red Hat environments Experience supporting engineering, research, laboratory, HPC, or data-centre environments Solid networking knowledge including routing, VLANs, VPNs, and troubleshooting complex connectivity issues Experience managing physical infrastructure including servers, rack-mounted equipment, BMCs, firmware, and out-of-band management Experience with configuration management and automation tools such as Ansible, Puppet, or similar Familiarity with authentication and access-management systems such as LDAP, RADIUS, or Active Directory integrations Strong troubleshooting skills with a structured and methodical approach to problem solving Excellent communication skills and a customer-focused mindset Desirable Container technologies such as Docker, containerd, or Kubernetes Monitoring and observability platforms such as Prometheus, Grafana, Zabbix, OpenTelemetry, or similar Python scripting and automation CI/CD tooling including GitLab or GitHub Actions Experience supporting hardware development, silicon validation, embedded systems, or electronics laboratories Performance analysis and troubleshooting across compute, storage, and network infrastructure Web infrastructure technologies including NGINX, HAProxy, or load balancing platforms We welcome people of different backgrounds and experiences; we’re committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments.
Showing 1–8 of 30 opportunities
Get the latest internships, jobs, scholarships and research opportunities delivered to your inbox.
Get early access to verified notifications & deadlines.
Only relevant and verified career opportunities.
We value your inbox. Zero marketing clutter.
10,000+ students & scholars actively advancing.