跳到主要内容
OOfficialJobs
菜单
官方来源官方来源职位

软件工程师,模型推理

机器翻译
查看雇主原标题Software Engineer, Model Inference

OpenAI · San Francisco · $266k – $500k

职位信息来自雇主公开的招聘页面。申请前请务必在雇主官网核实详情。

为什么值得关注?

发现指数 57/100,仅依据与该职位一起存储的证据计算。

57/100 发现指数
  • 新的雇主官方职位
  • 已披露薪资

分数构成

  • 时效性 (随职位发布时间变化)+18
  • 雇主官方来源+15
  • 已披露薪资+15
  • 稀有职位+1
  • 公司来源健康度+8

该职位未包含:远程职位、提及签证担保、提及搬迁、未出现在监控的职位板上。

这些理由来自雇主自己的职位描述与我们核实过的来源检查结果。除了已存储的信号之外,我们不做任何推测。

职位描述

机器翻译

关于团队 我们的推理团队通过我们的产品将 OpenAI 最强大的研究和技术带给全世界。我们赋能消费者、企业和开发者,让他们能够使用和访问我们最先进的 AI 模型,使他们能够做到以前从未做到的事情。我们专注于高性能、高效的模型推理,并通过模型推理加速研究进展。 关于职位 我们正在寻找一位工程师,希望将世界上最大、最强大的 AI 模型进行优化,使其适用于高吞吐量、低延迟和高可用性的生产与研究环境。 在这个职位中,你将: • 与机器学习研究人员、工程师和产品经理并肩合作,将我们最新的技术投入生产。

• 与研究人员并肩合作,通过出色的工程实现先进研究。

• 引入新技术、工具和架构,以提升我们模型推理栈的性能、延迟、吞吐量和效率。

• 构建工具,让我们能够洞察瓶颈和不稳定的来源,然后设计并实施解决方案,以解决最高优先级的问题。

• 优化我们的代码和 Azure VM 集群,以充分利用我们硬件的每一个 FLOP 和每一 GB 的 GPU RAM。

如果你具备以下条件,你可能会在这个职位中如鱼得水: • 了解现代 ML 架构,并对如何优化其性能(尤其是推理性能)有直觉。

• 能够端到端地负责问题,并愿意补齐完成工作所需的任何知识。

• 至少拥有 5 年专业软件工程经验。

• 已经熟悉或能够快速熟悉 PyTorch、NVidia GPU 以及优化它们的软件栈(例如 NCCL、CUDA),以及 InfiniBand、MPI、NVLink 等 HPC 技术。

• 拥有架构、构建、观测和调试生产分布式系统的经验。加分

岗位职责

我们正在寻找一位工程师,希望将世界上最大、最强大的 AI 模型进行优化,使其适用于高吞吐量、低延迟和高可用性的生产与研究环境。 在这个职位中,你将: • 与机器学习研究人员、工程师和产品经理并肩合作,将我们最新的技术投入生产。

• 与研究人员并肩合作,通过出色的工程实现先进研究。

• 引入新技术、工具和架构,以提升我们模型推理栈的性能、延迟、吞吐量和效率。

• 构建工具,让我们能够洞察瓶颈和不稳定的来源,然后设计并实施解决方案,以解决最高优先级的问题。

• 优化我们的代码和 Azure VM 集群,以充分利用我们硬件的每一个 FLOP 和每一 GB 的 GPU RAM。

如果你具备以下条件,你可能会在这个职位中如鱼得水: • 了解现代 ML 架构,并对如何优化其性能(尤其是推理性能)有直觉。

• 能够端到端地负责问题,并愿意补齐完成工作所需的任何知识。

• 至少拥有 5 年专业软件工程经验。

• 已经熟悉或能够快速熟悉 PyTorch、NVidia GPU 以及优化它们的软件栈(例如 NCCL、CUDA),以及 InfiniBand、MPI、NVLink 等 HPC 技术。

• 拥有架构、构建、观测和调试生产分布式系统的经验。如果有性能关键型分布式系统的工作经验,则加分。

• 曾因规模快速增长而多次需要重建或大幅重构生产系统。

• 自我驱动,乐于找出最重要的问题并着手解决。

• 态度谦逊,乐于帮助同事,并愿意为团队成功付出一切努力。

关于 OpenAI OpenAI 是一家 AI 研究与部署公司,致力于确保通用人工智能造福全人类。我们不断突破 AI 系统能力的边界,并寻求通过我们的产品将其安全地部署到全世界。AI 是一种极其强大的工具,其创建必须以安全和人类需求为核心;为了实现我们的使命,我们必须包容并重视构成人类全貌的众多不同视角、声音和经历。 我们是一家机会平等的雇主,不会基于种族、宗教、肤色、国籍、性别、性取向、年龄、退伍军人身份、残疾、遗传信息或其他适用的受法律保护特征进行歧视。 如需更多信息,请参阅 OpenAI 的平权行动与平等就业机会政策声明。 对申请人的背景调查将依据适用法律进行,对于美国候选人,有逮捕或定罪记录的合格申请人将依据这些法律获得就业考虑,包括《旧金山公平机会条例》、《洛杉矶县雇主公平机会条例》和《加州公平机会法》。对于未建制洛杉矶县的员工:我们合理认为,犯罪历史可能与以下工作职责存在直接、不利和负面的关系,并可能导致有条件录用通知被撤回:保护委托给你的计算机硬件免遭盗窃、丢失或损坏;在雇佣终止或任务结束时归还你持有的所有计算机硬件(包括其中包含的数据);以及维护专有、机密和非公开信息的保密性。此外,工作职责要求访问安全且受保护的信息技术系统以及相关的数据安全义务。 如需通知 OpenAI 你认为该职位发布不合规,请通过此表单提交报告。与职位发布合规无关的询问将不会得到回复。 我们致力于为残障申请人提供合理便利,可通过此链接提出请求。 OpenAI 全球申请人隐私政策 在 OpenAI,我们相信人工智能有潜力帮助人们解决巨大的全球挑战,我们希望 AI 带来的益处能够被广泛共享。加入我们,共同塑造技术的未来。

以上内容由机器翻译自动生成,可能存在错误;投递前请以雇主原文为准。

查看雇主原文

职位描述

About the Team Our Inference team brings OpenAI’s most capable research and technology to the world through our products. We empower consumers, enterprise and developers alike to use and access our start-of-the-art AI models, allowing them to do things that they’ve never been able to before. We focus on performant and efficient model inference, as well as accelerating research progression via model inference. About the Role We are looking for an engineer who wants to take the world's largest and most capable AI models and optimize them for use in a high-volume, low-latency, and high-availability production and research environment. In this role, you will: • Work alongside machine learning researchers, engineers, and product managers to bring our latest technologies into production.

• Work alongside researchers to enable advanced research through awesome engineering.

• Introduce new techniques, tools, and architecture that improve the performance, latency, throughput, and efficiency of our model inference stack.

• Build tools to give us visibility into our bottlenecks and sources of instability and then design and implement solutions to address the highest priority issues.

• Optimize our code and fleet of Azure VMs to utilize every FLOP and every GB of GPU RAM of our hardware.

You might thrive in this role if you: • Have an understanding of modern ML architectures and an intuition for how to optimize their performance, particularly for inference.

• Own problems end-to-end, and are willing to pick up whatever knowledge you're missing to get the job done.

• Have at least 5 years of professional software engineering experience.

• Have or can quickly gain familiarity with PyTorch, NVidia GPUs and the software stacks that optimize them (e.g. NCCL, CUDA), as well as HPC technologies such as InfiniBand, MPI, NVLink, etc.

• Have experience architecting, building, observing, and debugging production distributed systems. Bon

岗位职责

We are looking for an engineer who wants to take the world's largest and most capable AI models and optimize them for use in a high-volume, low-latency, and high-availability production and research environment. In this role, you will: • Work alongside machine learning researchers, engineers, and product managers to bring our latest technologies into production.

• Work alongside researchers to enable advanced research through awesome engineering.

• Introduce new techniques, tools, and architecture that improve the performance, latency, throughput, and efficiency of our model inference stack.

• Build tools to give us visibility into our bottlenecks and sources of instability and then design and implement solutions to address the highest priority issues.

• Optimize our code and fleet of Azure VMs to utilize every FLOP and every GB of GPU RAM of our hardware.

You might thrive in this role if you: • Have an understanding of modern ML architectures and an intuition for how to optimize their performance, particularly for inference.

• Own problems end-to-end, and are willing to pick up whatever knowledge you're missing to get the job done.

• Have at least 5 years of professional software engineering experience.

• Have or can quickly gain familiarity with PyTorch, NVidia GPUs and the software stacks that optimize them (e.g. NCCL, CUDA), as well as HPC technologies such as InfiniBand, MPI, NVLink, etc.

• Have experience architecting, building, observing, and debugging production distributed systems. Bonus point if worked on performance-critical distributed systems.

• Have needed to rebuild or substantially refactor production systems several times over due to rapidly increasing scale.

• Are self-directed and enjoy figuring out the most important problem to work on.

• Have a humble attitude, an eagerness to help your colleagues, and a desire to do whatever it takes to make the team succeed.

About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.

OpenAI 的更多职位

公司主页
官方来源最新
San Francisco远程全职$293k – $325k
英文原文

About the Role OpenAI’s Industrial Compute organization is responsible for ensuring our compute infrastructure scales efficiently to support millions of users and increasingly sophisticated AI mode…

未出现在监控的职位板上
首次发现于14小时前
已核实6小时前

模型策略经理

OpenAI · Safety Systems, Model Policy

官方来源最新
San Francisco远程全职$266k – $335k
英文原文

About the Team Our Safety Systems team is at the forefront of OpenAI's mission to build and deploy safe AGI, driving our commitment to AI safety and fostering a culture of trust and transparency.…

未出现在监控的职位板上提及搬迁
首次发现于14小时前
已核实6小时前

应用人工智能工程师

OpenAI · Go To Market, Technical Success

官方来源最新
新加坡远程全职未披露薪资
英文原文

About the Team OpenAI’s Applied AI Engineering team helps organizations turn frontier AI capabilities into safe, reliable, and high-impact production systems. We work with customer executives, prod…

未出现在监控的职位板上提及搬迁
首次发现于14小时前
已核实6小时前

战略财务,算力

OpenAI · Strategic Finance, Strategic Finance

官方来源最新
San Francisco全职$234k – $260k
英文原文

About the Team The Compute & Infrastructure Strategy team handles strategy and execution of OpenAI’s compute roadmap. This team’s key responsibilities span financial analysis & reporting, capacity…

未出现在监控的职位板上
首次发现于14小时前
已核实6小时前

其他公司的相似职位

搜索这类职位
India - Bangalore全职未披露薪资
英文原文

To get the best candidate experience, please consider applying for a maximum of 3 roles within 12 months to ensure you are not duplicating efforts. Job Category Software Engineering Job Detail…

未出现在监控的职位板上
首次发现于6小时前
已核实6小时前

Systems Engineer原文

Cloudflare · Engineering

官方来源最新
混合办公混合办公全职未披露薪资
英文原文

About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of websites and other Internet propertie…

官方来源职位
首次发现于6小时前
已核实6小时前
官方来源最新
混合办公混合办公全职未披露薪资
英文原文

About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of websites and other Internet propertie…

官方来源职位
首次发现于6小时前
已核实6小时前

Software Engineer原文

Coinbase · Engineering - Frontend

官方来源最新
Remote - 加拿大远程全职未披露薪资
英文原文

Ready to do the most impactful work of your career? At Coinbase , we are uncompromising on our mission to increase economic freedom. The bar is high, the environment is intense, and we like it that w…

官方来源职位
首次发现于6小时前
已核实6小时前