跳到主要内容
OOfficialJobs
菜单
官方来源官方来源职位

培训:机器学习框架工程师

机器翻译
查看雇主原标题Training: ML Framework Engineer

OpenAI · San Francisco · $295k – $500k

职位信息来自雇主公开的招聘页面。申请前请务必在雇主官网核实详情。

为什么值得关注?

发现指数 73/100,仅依据与该职位一起存储的证据计算。

73/100 发现指数
  • 新的雇主官方职位
  • 已披露薪资
  • 检测到搬迁支持关键词
  • 稀有职位匹配

分数构成

  • 时效性 (随职位发布时间变化)+18
  • 雇主官方来源+15
  • 已披露薪资+15
  • 提及搬迁+6
  • 稀有职位+11
  • 公司来源健康度+8

该职位未包含:远程职位、提及签证担保、未出现在监控的职位板上。

这些理由来自雇主自己的职位描述与我们核实过的来源检查结果。除了已存储的信号之外,我们不做任何推测。

职位描述

机器翻译

关于团队 Training Runtime 设计核心的分布式机器学习训练运行时,为从早期研究实验到前沿规模模型训练的一切提供支持。我们肩负双重使命:加速研究人员的工作并实现前沿规模,我们正在构建一个统一、模块化的运行时,既能满足研究人员当前的需求,又能伴随他们沿着扩展曲线不断前进。 我们的工作聚焦于三大支柱:高性能、异步、零拷贝的张量及优化器状态感知的数据移动;高性能、高可用、容错的训练框架(训练循环、状态管理、弹性检查点、确定性编排和可观测性);以及针对长期运行、特定作业和用户提供进程的分布式进程管理。 我们将经过验证的大规模能力整合到一个可组合、面向开发者的运行时中,使团队能够快速迭代并在任何规模下可靠运行,同时与模型栈、研究和平台团队紧密合作。对我们而言,成功的衡量标准是同时提升训练吞吐量(模型训练的速度)和研究人员吞吐量(想法转化为实验和产品的速度)。 关于该职位 作为一名 Training: ML Framework Engineer,你将致力于提升我们内部训练框架的训练吞吐量,同时让研究人员能够试验新想法。这需要良好的工程能力(例如设计、实现和优化最先进的 AI 模型)、编写无 bug 的机器学习代码(这出乎意料地困难!),以及深入了解超级计算机的性能。在该职位所追求的所有项目中,最终目标都是推动该领域向前发展。 我们正在寻找热爱优化性能、理解分布式系统,并且无法容忍代码中存在 bug 的人。由于我们的训练框架用于使用大量 GPU 的大规模训练,这里的性能改进将产生巨大影响。 该职位位于加利福尼亚州旧金山。我们采用每周 3 天到办公室的混合工作模式,并为新员工提供搬迁援助。

岗位职责

作为一名 Training: ML Framework Engineer,你将致力于提升我们内部训练框架的训练吞吐量,同时让研究人员能够试验新想法。这需要良好的工程能力(例如设计、实现和优化最先进的 AI 模型)、编写无 bug 的机器学习代码(这出乎意料地困难!),以及深入了解超级计算机的性能。在该职位所追求的所有项目中,最终目标都是推动该领域向前发展。 我们正在寻找热爱优化性能、理解分布式系统,并且无法容忍代码中存在 bug 的人。由于我们的训练框架用于使用大量 GPU 的大规模训练,这里的性能改进将产生巨大影响。 该职位位于加利福尼亚州旧金山。我们采用每周 3 天到办公室的混合工作模式,并为新员工提供搬迁援助。 在该职位中,你将: • 在我们内部训练框架中应用最新技术,为我们的训练实现出色的硬件效率

• 分析和优化我们的训练框架

• 与研究人员合作,使他们能够开发下一代模型

如果你符合以下条件,你可能会在该职位中如鱼得水: • 运行过小规模 ML 实验

• 热爱探究系统的工作原理,并不断提出如何让系统更快、同时最大限度降低复杂性和维护负担的想法

• 拥有扎实的软件工程技能,并精通 Python

关于 OpenAI OpenAI 是一家 AI 研究和部署公司,致力于确保通用人工智能造福全人类。我们不断拓展 AI 系统能力的边界,并寻求通过我们的产品将其安全地部署到世界各地。AI 是一种极其强大的工具,其创建必须以安全和人类需求为核心;为实现我们的使命,我们必须包容并重视构成人类完整光谱的众多不同视角、声音和经历。 我们是提供平等机会的雇主,不会基于种族、宗教、肤色、国籍、性别、性取向、年龄、退伍军人身份、残疾、遗传信息或其他适用的受法律保护特征进行歧视。 如需更多信息,请参阅 OpenAI 的《平权行动与平等就业机会政策声明》。 申请人的背景调查将依照适用法律进行,对于有逮捕或定罪记录的合格申请人,将根据相关法律考虑其就业,包括针对美国候选人的《旧金山公平机会条例》、《洛杉矶县雇主公平机会条例》和《加州公平机会法》。对于未建制洛杉矶县的员工:我们合理认为,犯罪史可能与以下工作职责存在直接、不利和负面的关系,可能导致有条件录用通知被撤回:保护委托给你的计算机硬件免遭盗窃、丢失或损坏;在雇佣终止或任务结束时归还你持有的所有计算机硬件(包括其中包含的数据);以及对专有、机密和非公开信息保密。此外,工作职责要求访问安全且受保护的信息技术系统,并承担相关的数据安全义务。 如需通知 OpenAI 你认为该职位发布不合规,请通过此表单提交报告。与职位发布合规无关的询问将不予回复。 我们致力于为残障申请人提供合理便利,可通过此链接提出请求。 OpenAI 全球申请人隐私政策 在 OpenAI,我们相信人工智能有潜力帮助人们解决巨大的全球性挑战,我们希望 AI 带来的益处能够被广泛共享。加入我们,共同塑造技术的未来。

以上内容由机器翻译自动生成,可能存在错误;投递前请以雇主原文为准。

查看雇主原文

职位描述

About the Team Training Runtime designs the core distributed machine-learning training runtime that powers everything from early research experiments to frontier-scale model runs. With a dual mandate to accelerate researchers and enable frontier scale, we’re building a unified, modular runtime that meets researchers where they are and moves with them up the scaling curve. Our work focuses on three pillars: high-performance, asynchronous, zero-copy tensor and optimizer-state-aware data movement; performant, high-uptime, fault-tolerant training frameworks (training loop, state management, resilient checkpointing, deterministic orchestration, and observability); and distributed process management for long-lived, job-specific and user-provided processes. We integrate proven large-scale capabilities into a composable, developer-facing runtime so teams can iterate quickly and run reliably at any scale, partnering closely with model-stack, research, and platform teams. Success for us is measured by raising both training throughput (how fast models train) and researcher throughput (how fast ideas become experiments and products). About the Role As a Training: ML Framework Engineer, you will work on improving the training throughput for our internal training framework, while enabling researchers to experiment with new ideas. This requires good engineering (for example designing, implementing, and optimizing state-of-the-art AI models), writing bug-free machine learning code (surprisingly difficult!), and acquiring deep knowledge of the performance of supercomputers. In all the projects this role pursues, the ultimate goal is to push the field forward. We’re looking for people who love optimizing performance, understanding distributed systems, and who cannot stand having bugs in their code. Since our training framework is used for large runs with massive numbers of GPUs, performance improvements here will have a large impact. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new

岗位职责

As a Training: ML Framework Engineer, you will work on improving the training throughput for our internal training framework, while enabling researchers to experiment with new ideas. This requires good engineering (for example designing, implementing, and optimizing state-of-the-art AI models), writing bug-free machine learning code (surprisingly difficult!), and acquiring deep knowledge of the performance of supercomputers. In all the projects this role pursues, the ultimate goal is to push the field forward. We’re looking for people who love optimizing performance, understanding distributed systems, and who cannot stand having bugs in their code. Since our training framework is used for large runs with massive numbers of GPUs, performance improvements here will have a large impact. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: • Apply the latest techniques in our internal training framework to achieve impressive hardware efficiency for our training runs

• Profile and optimize our training framework

• Work with researchers to enable them to develop the next generation of models

You might thrive in this role if you: • Have run small scale ML experiments

• Love figuring out how systems work and continuously come up with ideas for how to make them faster while minimizing complexity and maintenance burden

• Have strong software engineering skills and are proficient in Python

About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.

OpenAI 的更多职位

公司主页
官方来源最新
San Francisco远程全职$293k – $325k
英文原文

About the Role OpenAI’s Industrial Compute organization is responsible for ensuring our compute infrastructure scales efficiently to support millions of users and increasingly sophisticated AI mode…

未出现在监控的职位板上
首次发现于17小时前
已核实10小时前

模型策略经理

OpenAI · Safety Systems, Model Policy

官方来源最新
San Francisco远程全职$266k – $335k
英文原文

About the Team Our Safety Systems team is at the forefront of OpenAI's mission to build and deploy safe AGI, driving our commitment to AI safety and fostering a culture of trust and transparency.…

未出现在监控的职位板上提及搬迁
首次发现于17小时前
已核实10小时前

应用人工智能工程师

OpenAI · Go To Market, Technical Success

官方来源最新
新加坡远程全职未披露薪资
英文原文

About the Team OpenAI’s Applied AI Engineering team helps organizations turn frontier AI capabilities into safe, reliable, and high-impact production systems. We work with customer executives, prod…

未出现在监控的职位板上提及搬迁
首次发现于17小时前
已核实10小时前

战略财务,算力

OpenAI · Strategic Finance, Strategic Finance

官方来源最新
San Francisco全职$234k – $260k
英文原文

About the Team The Compute & Infrastructure Strategy team handles strategy and execution of OpenAI’s compute roadmap. This team’s key responsibilities span financial analysis & reporting, capacity…

未出现在监控的职位板上
首次发现于17小时前
已核实10小时前