跳到主要内容
OOfficialJobs
菜单
官方来源官方来源职位

首席工程师

机器翻译
查看雇主原标题Principal Engineer

Graphcore · Austin, Texas, United States; Milpitas, California, 美国 · salary, Graphcore offers flexible working and a comprehensive benefits package designed to support your health, wellbeing and financial future. Our benefits inc

职位信息来自雇主公开的招聘页面。申请前请务必在雇主官网核实详情。

为什么值得关注?

发现指数 46/100,仅依据与该职位一起存储的证据计算。

46/100 发现指数
  • 新的雇主官方职位

分数构成

  • 时效性 (随职位发布时间变化)+18
  • 雇主官方来源+15
  • 稀有职位+5
  • 公司来源健康度+8

该职位未包含:已披露薪资、远程职位、提及签证担保、提及搬迁、未出现在监控的职位板上。

这些理由来自雇主自己的职位描述与我们核实过的来源检查结果。除了已存储的信号之外,我们不做任何推测。

职位描述

英文原文

该职位由雇主以英文发布,暂无中文版本,下面完整显示英文原文。 查看官方职位页面.

职位描述

About us

Graphcore is one of the world’s leading innovators in Artificial Intelligence compute.

It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry.

As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone.

Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives, spanning AI research specialists, silicon designers, software engineers and systems architects.

Job Summary

We are looking for an experienced Principal Engineer to join our System Management team and help lead the development of critical interfaces used by internal and external customers to manage system state. You will provide technical leadership within assigned areas of System Management, guide architecture and implementation choices, mentor engineers and translate broader technical direction into effective execution. This is a hands-on engineering role for someone who can lead complex technical work, improve reliability and operational readiness, and collaborate effectively across multiple engineering disciplines.

The Team

The System Management team sits within the Software Platform group and helps build Graphcore products into large-scale AI solutions for our customers.

The team is responsible for developing the interfaces between hardware, AI software and frameworks, as well as providing interfaces for public and private cloud environments. This includes system management capabilities that abstract complex hardware administration and enable reliable deployment and operation at scale.

As one of the first teams to work with new hardware and software, we regularly solve complex system-level problems in environments where components and interfaces are still evolving. The role requires strong technical judgement, adaptability and an ability to work effectively across engineering teams.

Responsibilities and Duties

• Convert agreed System Management direction into technical plans, engineering priorities and deliverable work for assigned areas.

• Provide technical leadership for architecture and design decisions, building alignment across collaborating teams and documenting important technical trade-offs.

• Act as a technical authority for assigned areas of System Management, leading the delivery of large and complex engineering initiatives and coordinating technical plans, dependencies,

岗位职责

• Convert agreed System Management direction into technical plans, engineering priorities and deliverable work for assigned areas.

• Provide technical leadership for architecture and design decisions, building alignment across collaborating teams and documenting important technical trade-offs.

• Act as a technical authority for assigned areas of System Management, leading the delivery of large and complex engineering initiatives and coordinating technical plans, dependencies, risks and decisions.

• Provide technical direction and mentoring to engineers working across system management, hardware lifecycle management, deployment automation and production operations.

• Take responsibility for key technical outcomes across the full software lifecycle, including design, implementation, automated testing, integration, deployment, observability and production readiness.

• Identify systemic reliability, scalability and operability issues and lead practical improvements across the platform.

• Collaborate with Hardware, Firmware, Platform Software and Datacenter Operations teams to diagnose system-level issues and improve end-to-end product behaviour.

• Improve engineering standards and working practices, including CI/CD, Infrastructure-as-Code, automated testing, release safety and learning from operational incidents.

• Act as a senior technical escalation point for complex issues while creating reusable knowledge, tooling and automation that reduce future operational effort.

Candidate Profile

任职要求

• Bachelor’s degree or equivalent practical experience in a relevant subject.

• Substantial experience designing, building and operating Linux-based infrastructure or distributed systems.

• Demonstrated experience providing technical leadership for complex engineering initiatives involving multiple teams or stakeholder groups.

• Experience influencing architecture and technical decisions across team boundaries without relying on formal authority.

• Experience translating broad technical goals into scoped plans, milestones, technical decisions, risks and delivery priorities.

• Strong experience developing RESTful APIs and programming in Go, with Bash and Python used for systems automation.

• Deep practical experience with Kubernetes, container runtimes and operating production workloads.

• Hands-on experience with Infrastructure-as-Code, source control and CI/CD technologies such as Terraform/OpenTofu, Ansible, GitLab, GitHub Actions and Git.

• Experience with hardware-management interfaces such as Redfish, IPMI or equivalent management systems.

• Strong Linux systems engineering, troubleshooting and operational debugging capability.

• Demonstrated ability to develop other engineers through technical mentoring, design reviews and coaching.

• Clear communication skills, with the ability to persuade, build alignment and bring stakeholders together around practical technical outcomes.

Desirable

• Experience using AI coding assistants effectively within professional engineering workflows.

• Experience developing Kubernetes operators and custom resources.

• Experience with High Performance Computing environments using SLURM, LSF or similar workload-management systems.

• Experience with virtualisation technologies such as Open vSwitch, KVM and QEMU.

• Experience with distributed object, block and file storage technologies such as Ceph.

• Experience with monitoring and observability platforms such as Grafana, Prometheus, OpenSearch/Elasticsearch, Loki, Mimir or OpenTelemetry.

• Experience configuring managed network switches using technologies such as EOS, SONiC or DNOS.

• Experience supporting AI infrastructure or PyTorch workloads.

In addition to a competitive salary, Graphcore offers flexible working and a comprehensive benefits package designed to support your health, wellbeing and financial future. Our benefits include medical, dental and vision coverage, Flexible Spending Accounts (FSAs), Health Savings Accounts (HSAs), disability and life insurance, a 401(k) retirement plan, commuter benefits, wellness services and an Employee Assistance Programme (EAP). We welcome people of different backgrounds and experiences; we're committed to building an inclusive work environment that makes Graphcore a great home for everyone. We offer an equal opportunity process and understand that there are visible and invisible differences in all of us. We can provide a flexible approach to interview and encourage you to chat to us if you require any reasonable adjustments.

Graphcore 的更多职位

公司主页
官方来源最新
Austin, Texas, 美国全职salary, Graphcore offers flexible working and a comprehensive benefits package designed to support your health, wellbeing and financial future. Our benefits inc
英文原文

About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation…

未出现在监控的职位板上
首次发现于2小时前
已核实2小时前
官方来源最新
Austin, Texas, 美国全职salary, Graphcore offers flexible working and a comprehensive benefits package designed to support your health, wellbeing and financial future. Our benefits inc
英文原文

We are looking for a disciplined and dynamic, Lead System Engineer – compute blade and rack Validation to join our growing compute rack validation team. As a diligent leader in Systems Engineering, y…

未出现在监控的职位板上
首次发现于2小时前
已核实2小时前
官方来源最新
Austin, Texas, 美国全职salary, Graphcore offers flexible working and a comprehensive benefits package designed to support your health, wellbeing and financial future. Our benefits inc
英文原文

About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation…

未出现在监控的职位板上
首次发现于2小时前
已核实2小时前

Senior Systems Engineer原文

Graphcore · DC Engineering*

官方来源最新
Austin, Texas, 美国全职salary, Graphcore offers flexible working and a comprehensive benefits package designed to support your health, wellbeing and financial future. Our benefits inc
英文原文

About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation o…

未出现在监控的职位板上
首次发现于2小时前
已核实2小时前

其他公司的相似职位

搜索这类职位
Mexico - Mexico City全职未披露薪资
英文原文

To get the best candidate experience, please consider applying for a maximum of 3 roles within 12 months to ensure you are not duplicating efforts. Job Category Sales Job Details About Sal…

官方来源职位
首次发现于8小时前
已核实8小时前

Principal Customer Engineer, Majors原文

Cloudflare · Solution Engineering

官方来源最新
Distributed全职€148k – €204k
英文原文

About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of websites and other Internet propertie…

官方来源职位
首次发现于8小时前
已核实8小时前
官方来源最新
Distributed全职€148k – €204k
英文原文

About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of websites and other Internet propertie…

官方来源职位
首次发现于8小时前
已核实8小时前
Long Beach, CA全职Base salary is just one component of our total rewards package at Rocket Lab. Employees may also receive company equity and access to a robust benefits package
英文原文

ABOUT ROCKET LAB Rocket Lab is the end-to-end space company building rockets, spacecraft, and critical subsystems that keep the world connected, protected, expand humanity’s reach to the Moon, Mar…

未出现在监控的职位板上
首次发现于3天前
已核实8小时前