软件工程师,数据基础设施
查看雇主原标题
Software Engineer, Data InfrastructureOpenAI · San Francisco; New York City; Seattle; Mountain View · $266k – $445k
职位信息来自雇主公开的招聘页面。申请前请务必在雇主官网核实详情。
为什么值得关注?
发现指数 63/100,仅依据与该职位一起存储的证据计算。
- 新的雇主官方职位
- 已披露薪资
- 检测到搬迁支持关键词
分数构成
- 时效性 (随职位发布时间变化)+18
- 雇主官方来源+15
- 已披露薪资+15
- 提及搬迁+6
- 稀有职位+1
- 公司来源健康度+8
该职位未包含:远程职位、提及签证担保、未出现在监控的职位板上。
这些理由来自雇主自己的职位描述与我们核实过的来源检查结果。除了已存储的信号之外,我们不做任何推测。
职位描述
机器翻译关于团队 OpenAI 的数据平台负责拥有支撑关键产品、研究和分析工作流的基础数据栈。我们运营着一些生产环境中规模最大的 Spark 计算集群;在 Iceberg 和 Delta 上设计和构建数据湖及元数据系统,并以艾字节级架构为愿景;在 Kafka 和 Flink 上运行高吞吐量流处理平台;通过 Airflow 提供编排;并支持诸如 Chronon 之类的 ML 特征工程工具。我们的使命是提供可靠、安全且高效的大规模数据访问,并加速智能的、AI 辅助的数据工作流。 加入我们,共同构建和运营这些支撑 OpenAI 产品、研究和分析的核心平台。 我们不仅仅是在扩展基础设施——我们正在重新定义人们与数据交互的方式。我们的愿景包括智能界面和 AI 辅助工作流,使处理数据变得更快、更可靠、更直观。 关于该职位 该职位专注于构建和运营支持大规模计算集群和存储系统的数据基础设施,专为高性能和可扩展性而设计。你将帮助设计、构建和运营 OpenAI 下一代数据基础设施。你将扩展并加固大数据计算和存储平台,构建并支持高吞吐量流处理系统,构建并运营低延迟数据摄取,为 ML 和分析实现安全且受治理的数据访问,并为极端规模下的可靠性和性能进行设计。 你将承担全生命周期责任:架构、实现、生产运营以及参与 on-call。 你曾作为平台支持过 Spark、Kafka、Flink、Airflow、Trino 或 Iceberg。你精通 Terraform 等基础设施工具,拥有调试大规模分布式系统的经验,并对在 AI 领域解决数据基础设施问题充满热情。 该职位位于加利福尼亚州旧金山。我们采用每周 3 天到办公室的混合工作模式,并为新员工提供搬迁协助。 在该职位中,你将: • 设计、构建和维护数据基础设施系统,例如分布式计算、数据编排、分布式存储、流处理基础设施、机器学习基础设施,同时确保可扩展性、可靠性和安全性
• 确保我们的数据平台能够扩展 o
岗位职责
该职位专注于构建和运营支持大规模计算集群和存储系统的数据基础设施,专为高性能和可扩展性而设计。你将帮助设计、构建和运营 OpenAI 下一代数据基础设施。你将扩展并加固大数据计算和存储平台,构建并支持高吞吐量流处理系统,构建并运营低延迟数据摄取,为 ML 和分析实现安全且受治理的数据访问,并为极端规模下的可靠性和性能进行设计。 你将承担全生命周期责任:架构、实现、生产运营以及参与 on-call。 你曾作为平台支持过 Spark、Kafka、Flink、Airflow、Trino 或 Iceberg。你精通 Terraform 等基础设施工具,拥有调试大规模分布式系统的经验,并对在 AI 领域解决数据基础设施问题充满热情。 该职位位于加利福尼亚州旧金山。我们采用每周 3 天到办公室的混合工作模式,并为新员工提供搬迁协助。 在该职位中,你将: • 设计、构建和维护数据基础设施系统,例如分布式计算、数据编排、分布式存储、流处理基础设施、机器学习基础设施,同时确保可扩展性、可靠性和安全性
• 确保我们的数据平台能够扩展数个数量级,同时保持可靠和高效
• 通过为你的工程师同事和团队成员提供出色的数据工具和系统,加速公司生产力
• 与产品、研究和分析团队合作,构建解锁新功能和体验的技术基础能力
• 对你所构建系统的可靠性负责,包括参与关键事件的 on-call 轮值
如果你符合以下条件,你可能会在该职位中如鱼得水: • 拥有 4 年以上数据基础设施工程经验,或
• 拥有 4 年以上基础设施工程经验,并对数据有强烈兴趣
• 以构建和运营可扩展、可靠、安全的系统为荣
• 能够适应模糊性和快速变化
• 拥有内在的学习渴望并填补技能缺口,同时具备同样出色的才能,能够清晰简洁地与他人分享所学
该职位仅位于我们的旧金山总部。我们为新员工提供搬迁协助。
关于 OpenAI OpenAI 是一家 AI 研究和部署公司,致力于确保通用人工智能造福全人类。我们推动 AI 系统能力边界,并寻求通过我们的产品将其安全地部署到世界。AI 是一种极其强大的工具,其创建必须以安全和人类需求为核心,而为了实现我们的使命,我们必须包容并重视构成人类全貌的众多不同视角、声音和经历。 我们是提供平等机会的雇主,我们不会基于种族、宗教、肤色、国籍、性别、性取向、年龄、退伍军人身份、残疾、遗传信息或其他适用的受法律保护特征进行歧视。 如需更多信息,请参阅 OpenAI 的平权行动和 equal employment opportunity 政策声明。 申请人的背景调查将根据适用法律进行,对于美国候选人,有逮捕或定罪记录的合格申请人将根据这些法律获得就业考虑,包括《旧金山公平机会条例》、《洛杉矶县雇主公平机会条例》和《加利福尼亚公平机会法》。对于未建制洛杉矶县的员工:我们合理认为犯罪历史可能与以下工作职责存在直接、不利和负面的关系,可能导致有条件录用通知被撤回:保护委托给你的计算机硬件免遭盗窃、丢失或损坏;在雇佣终止或任务结束时归还你持有的所有计算机硬件(包括其中包含的数据);以及维护专有、机密和非公开信息的机密性。此外,工作职责要求访问安全且受保护的信息技术系统以及相关的数据安全义务。 如需通知 OpenAI 你认为该职位发布不合规,请通过此表单提交报告。与职位发布合规无关的询问将不会得到回复。 我们致力于为有残疾的申请人提供合理便利,可通过此链接提出请求。 OpenAI 全球申请人隐私政策 在 OpenAI,我们相信人工智能有潜力帮助人们解决巨大的全球挑战,我们希望 AI 带来的益处能够被广泛共享。加入我们,共同塑造技术的未来。
以上内容由机器翻译自动生成,可能存在错误;投递前请以雇主原文为准。
查看雇主原文
职位描述
About the Team Data Platform at OpenAI owns the foundational data stack powering critical product, research, and analytics workflows. We operate some of the largest Spark compute fleets in production; design, and build data lakes and metadata systems on Iceberg and Delta with a vision toward exabyte-scale architecture; run high throughput streaming platforms on Kafka and Flink; provide orchestration with Airflow; and support ML feature engineering tooling such as Chronon. Our mission is to deliver reliable, secure, and efficient data access at scale and accelerate intelligent, AI assisted data workflows. Join us to build and operate these core platforms that underpin OpenAI products, research, and analytics. We’re not just scaling infrastructure – we’re redefining how people interact with data. Our vision includes intelligent interfaces and AI-assisted workflows that make working with data faster, more reliable, and more intuitive. About the Role This role focuses on building and operating data infrastructure that supports massive compute fleets and storage systems, designed for high performance and scalability. You’ll help design, build, and operate the next generation of data infrastructure at OpenAI. You will scale and harden big data compute and storage platforms, build and support high-throughput streaming systems, build and operate low latency data ingestions, enable secure and governed data access for ML and analytics, and design for reliability and performance at extreme scale. You will take full lifecycle ownership: architecture, implementation, production operations, and on-call participation. You’ve supported Spark, Kafka, Flink, Airflow, Trino, or Iceberg as platforms. You’re well-versed in infrastructure tooling like Terraform, experienced in debugging large-scale distributed systems, and excited about solving data infrastructure problems in the AI space. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: • Design, build, and maintain data infrastructure systems such as distributed compute, data orchestration, distributed storage, streaming infrastructure, machine learning infrastructure while ensuring scalability, reliability, and security
• Ensure our data platform can scale by o
岗位职责
This role focuses on building and operating data infrastructure that supports massive compute fleets and storage systems, designed for high performance and scalability. You’ll help design, build, and operate the next generation of data infrastructure at OpenAI. You will scale and harden big data compute and storage platforms, build and support high-throughput streaming systems, build and operate low latency data ingestions, enable secure and governed data access for ML and analytics, and design for reliability and performance at extreme scale. You will take full lifecycle ownership: architecture, implementation, production operations, and on-call participation. You’ve supported Spark, Kafka, Flink, Airflow, Trino, or Iceberg as platforms. You’re well-versed in infrastructure tooling like Terraform, experienced in debugging large-scale distributed systems, and excited about solving data infrastructure problems in the AI space. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: • Design, build, and maintain data infrastructure systems such as distributed compute, data orchestration, distributed storage, streaming infrastructure, machine learning infrastructure while ensuring scalability, reliability, and security
• Ensure our data platform can scale by orders of magnitude while remaining reliable and efficient
• Accelerate company productivity by empowering your fellow engineers & teammates with excellent data tooling and systems
• Collaborate with product, research and analytics teams to build the technical foundations capabilities that unlock new features and experiences
• Own the reliability of the systems you build, including participation in an on-call rotation for critical incidents
You might thrive in this role if you: • Have 4+ years in data infrastructure engineering OR
• Have 4+ years in infrastructure engineering with a strong interest in data
• Take pride in building and operating scalable, reliable, secure systems
• Are comfortable with ambiguity and rapid change
• Have an intrinsic desire to learn and fill in missing skills, and an equally strong talent for sharing learnings clearly and concisely with others
This role is exclusively based in our San Francisco HQ. We offer relocation assistance to new employees. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.