跳到主要内容
OOfficialJobs
菜单
官方来源官方来源职位

软件工程师,推理 - 多模态

机器翻译
查看雇主原标题Software Engineer, Inference - Multi Modal

OpenAI · San Francisco · $295k – $555k

职位信息来自雇主公开的招聘页面。申请前请务必在雇主官网核实详情。

为什么值得关注?

发现指数 57/100,仅依据与该职位一起存储的证据计算。

57/100 发现指数
  • 新的雇主官方职位
  • 已披露薪资

分数构成

  • 时效性 (随职位发布时间变化)+18
  • 雇主官方来源+15
  • 已披露薪资+15
  • 稀有职位+1
  • 公司来源健康度+8

该职位未包含:远程职位、提及签证担保、提及搬迁、未出现在监控的职位板上。

这些理由来自雇主自己的职位描述与我们核实过的来源检查结果。除了已存储的信号之外,我们不做任何推测。

职位描述

机器翻译

关于团队 OpenAI 的推理团队负责在多种平台上部署我们最先进的模型——包括我们的 GPT 模型、4o 图像生成和 Whisper。我们的工作确保这些模型在生产环境中可用、高性能且可扩展,我们与研究团队紧密合作,将下一代模型推向世界。我们是一支规模小、行动迅速的工程师团队,专注于提供世界一流的开发者体验,同时推动 AI 能力的边界。 我们正在扩展到多模态推理,构建为处理图像、音频和其他非文本模态的模型提供服务所需的基础设施。这些工作负载本质上更加异构和实验性,涉及多样的模型规模和交互、更复杂的输入/输出格式,以及与产品和研究的更紧密协调。 关于该职位 我们正在寻找一名软件工程师,帮助我们大规模地为 OpenAI 的多模态模型提供服务。你将加入一个小团队,负责构建可靠、高性能的基础设施,在生产环境中为实时音频、图像和其他 MM 工作负载提供服务。 这项工作本质上是跨职能的:你将直接与训练这些模型的研究人员以及定义新交互模态的产品团队合作。你将构建并优化系统,让用户能够生成语音、理解图像,并以远超文本的方式与模型交互。 在这个职位中,你将: • 为大规模多模态模型设计和实现推理基础设施。

• 优化系统,以实现图像和音频输入输出的高吞吐、低延迟交付。

• 使实验性研究工作流能够过渡为可靠的生产服务。

• 与研究人员、基础设施团队和产品工程师紧密合作,部署最先进的能力。

• 为系统级改进做出贡献,包括 GPU 利用率、张量并行和硬件抽象层。

如果你具备以下条件,你可能会在这个职位中如鱼得水: • 拥有为 LLM 或多模态模型构建和扩展推理系统的经验。

• 曾处理过基于 GPU 的

岗位职责

我们正在寻找一名软件工程师,帮助我们大规模地为 OpenAI 的多模态模型提供服务。你将加入一个小团队,负责构建可靠、高性能的基础设施,在生产环境中为实时音频、图像和其他 MM 工作负载提供服务。 这项工作本质上是跨职能的:你将直接与训练这些模型的研究人员以及定义新交互模态的产品团队合作。你将构建并优化系统,让用户能够生成语音、理解图像,并以远超文本的方式与模型交互。 在这个职位中,你将: • 为大规模多模态模型设计和实现推理基础设施。

• 优化系统,以实现图像和音频输入输出的高吞吐、低延迟交付。

• 使实验性研究工作流能够过渡为可靠的生产服务。

• 与研究人员、基础设施团队和产品工程师紧密合作,部署最先进的能力。

• 为系统级改进做出贡献,包括 GPU 利用率、张量并行和硬件抽象层。

如果你具备以下条件,你可能会在这个职位中如鱼得水: • 拥有为 LLM 或多模态模型构建和扩展推理系统的经验。

• 曾处理过基于 GPU 的 ML 工作负载,并理解大型模型的性能动态,尤其是处理图像或音频等复杂数据时。

• 喜欢实验性、快速演进的工作,并与研究团队紧密合作。

• 能够自如应对涉及网络、分布式计算和高吞吐数据处理的系统。

• 熟悉 vLLM、TensorRT-LLM 或自定义模型并行系统等推理工具。

• 能够端到端地负责问题,并乐于在模糊、快速变化的环境中工作。

加分项: • 有在生产环境中处理图像生成或音频合成模型的经验。

• 接触过分布式 ML 训练或系统高效的模型设计。

关于 OpenAI OpenAI 是一家 AI 研究与部署公司,致力于确保通用人工智能造福全人类。我们推动 AI 系统能力的边界,并寻求通过我们的产品将其安全地部署到世界。AI 是一种极其强大的工具,其创建必须以安全和人类需求为核心,而为了实现我们的使命,我们必须包容并重视构成人类全貌的众多不同视角、声音和经历。 我们是一家提供平等机会的雇主,我们不会基于种族、宗教、肤色、国籍、性别、性取向、年龄、退伍军人身份、残疾、遗传信息或其他适用的受法律保护特征进行歧视。 如需更多信息,请参阅 OpenAI 的平权行动和平等就业机会政策声明。 申请人的背景调查将根据适用法律进行,对于美国候选人,有逮捕或定罪记录的合格申请人将根据这些法律获得就业考虑,包括《旧金山公平机会条例》、《洛杉矶县雇主公平机会条例》和《加州公平机会法》。对于未建制洛杉矶县的员工:我们合理认为犯罪历史可能与以下工作职责存在直接、不利和负面的关系,可能导致撤回有条件录用通知:保护委托给你的计算机硬件免遭盗窃、丢失或损坏;在雇佣终止或任务结束时归还你持有的所有计算机硬件(包括其中包含的数据);以及维护专有、机密和非公开信息的保密性。此外,工作职责要求访问安全和受保护的信息技术系统以及相关的数据安全义务。 如需通知 OpenAI 你认为该职位发布不合规,请通过此表单提交报告。与职位发布合规无关的询问将不予回复。 我们致力于为有残疾的申请人提供合理的便利,可通过此链接提出请求。 OpenAI 全球申请人隐私政策 在 OpenAI,我们相信人工智能有潜力帮助人们解决巨大的全球挑战,我们希望 AI 的好处能够被广泛分享。加入我们,共同塑造技术的未来。

以上内容由机器翻译自动生成,可能存在错误;投递前请以雇主原文为准。

查看雇主原文

职位描述

About the Team OpenAI’s Inference team powers the deployment of our most advanced models - including our GPT models, 4o Image Generation, and Whisper - across a variety of platforms. Our work ensures these models are available, performant, and scalable in production, and we partner closely with Research to bring the next generation of models into the world. We're a small, fast-moving team of engineers focused on delivering a world-class developer experience while pushing the boundaries of what AI can do. We’re expanding into multimodal inference, building the infrastructure needed to serve models that handle image, audio, and other non-text modalities. These workloads are inherently more heterogeneous and experimental, involving diverse model sizes and interactions, more complex input/output formats, and tighter coordination with product and research. About the Role We’re looking for a software engineer to help us serve OpenAI’s multimodal models at scale. You’ll be part of a small team responsible for building reliable, high-performance infrastructure for serving real-time audio, image, and other MM workloads in production. This work is inherently cross-functional: you’ll collaborate directly with researchers training these models and with product teams defining new modalities of interaction. You'll build and optimize the systems that let users generate speech, understand images, and interact with models in ways far beyond text. In this role, you will: • Design and implement inference infrastructure for large-scale multimodal models.

• Optimize systems for high-throughput, low-latency delivery of image and audio inputs and outputs.

• Enable experimental research workflows to transition into reliable production services.

• Collaborate closely with researchers, infra teams, and product engineers to deploy state-of-the-art capabilities.

• Contribute to system-level improvements including GPU utilization, tensor parallelism, and hardware abstraction layers.

You might thrive in this role if you: • Have experience building and scaling inference systems for LLMs or multimodal models.

• Have worked with GPU-base

岗位职责

We’re looking for a software engineer to help us serve OpenAI’s multimodal models at scale. You’ll be part of a small team responsible for building reliable, high-performance infrastructure for serving real-time audio, image, and other MM workloads in production. This work is inherently cross-functional: you’ll collaborate directly with researchers training these models and with product teams defining new modalities of interaction. You'll build and optimize the systems that let users generate speech, understand images, and interact with models in ways far beyond text. In this role, you will: • Design and implement inference infrastructure for large-scale multimodal models.

• Optimize systems for high-throughput, low-latency delivery of image and audio inputs and outputs.

• Enable experimental research workflows to transition into reliable production services.

• Collaborate closely with researchers, infra teams, and product engineers to deploy state-of-the-art capabilities.

• Contribute to system-level improvements including GPU utilization, tensor parallelism, and hardware abstraction layers.

You might thrive in this role if you: • Have experience building and scaling inference systems for LLMs or multimodal models.

• Have worked with GPU-based ML workloads and understand the performance dynamics of large models, especially with complex data like images or audio.

• Enjoy experimental, fast-evolving work and collaborating closely with research.

• Are comfortable dealing with systems that span networking, distributed compute, and high-throughput data handling.

• Have familiarity with inference tooling like vLLM, TensorRT-LLM, or custom model parallel systems.

• Own problems end-to-end and are excited to operate in ambiguous, fast-moving spaces.

Nice to Have: • Experience working with image generation or audio synthesis models in production.

• Exposure to distributed ML training or system-efficient model design.

About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.

OpenAI 的更多职位

公司主页
官方来源最新
San Francisco远程全职$293k – $325k
英文原文

About the Role OpenAI’s Industrial Compute organization is responsible for ensuring our compute infrastructure scales efficiently to support millions of users and increasingly sophisticated AI mode…

未出现在监控的职位板上
首次发现于14小时前
已核实7小时前

模型策略经理

OpenAI · Safety Systems, Model Policy

官方来源最新
San Francisco远程全职$266k – $335k
英文原文

About the Team Our Safety Systems team is at the forefront of OpenAI's mission to build and deploy safe AGI, driving our commitment to AI safety and fostering a culture of trust and transparency.…

未出现在监控的职位板上提及搬迁
首次发现于14小时前
已核实7小时前

应用人工智能工程师

OpenAI · Go To Market, Technical Success

官方来源最新
新加坡远程全职未披露薪资
英文原文

About the Team OpenAI’s Applied AI Engineering team helps organizations turn frontier AI capabilities into safe, reliable, and high-impact production systems. We work with customer executives, prod…

未出现在监控的职位板上提及搬迁
首次发现于14小时前
已核实7小时前

战略财务,算力

OpenAI · Strategic Finance, Strategic Finance

官方来源最新
San Francisco全职$234k – $260k
英文原文

About the Team The Compute & Infrastructure Strategy team handles strategy and execution of OpenAI’s compute roadmap. This team’s key responsibilities span financial analysis & reporting, capacity…

未出现在监控的职位板上
首次发现于14小时前
已核实7小时前

其他公司的相似职位

搜索这类职位
官方来源最新
Hybrid - San Francisco, New York City混合办公全职$208k – $312k
英文原文

About Vercel: Vercel is the agentic infrastructure company. We free people and agents to ship what’s next. For more than a decade, Vercel has shaped how the web is built. As the team behind Next…

未出现在监控的职位板上
首次发现于41分钟前
已核实41分钟前

Staff Software Engineer, Event logging原文

Airbnb · Software Engineering

官方来源最新
美国全职未披露薪资
英文原文

Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every c…

未出现在监控的职位板上
首次发现于41分钟前
已核实41分钟前

Senior Staff Software Engineer, Trust原文

Airbnb · Software Engineering

官方来源最新
Remote - US远程全职未披露薪资
英文原文

Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every c…

未出现在监控的职位板上
首次发现于41分钟前
已核实41分钟前
官方来源最新
美国全职未披露薪资
英文原文

Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every c…

未出现在监控的职位板上
首次发现于41分钟前
已核实41分钟前