软件工程师,模型路由与推理
查看雇主原标题
Software Engineer, Model Routing & InferenceCursor (Anysphere) · New York; San Francisco
职位信息来自雇主公开的招聘页面。申请前请务必在雇主官网核实详情。
为什么值得关注?
发现指数 42/100,仅依据与该职位一起存储的证据计算。
- 新的雇主官方职位
分数构成
- 时效性 (随职位发布时间变化)+18
- 雇主官方来源+15
- 稀有职位+1
- 公司来源健康度+8
该职位未包含:已披露薪资、远程职位、提及签证担保、提及搬迁、未出现在监控的职位板上。
这些理由来自雇主自己的职位描述与我们核实过的来源检查结果。除了已存储的信号之外,我们不做任何推测。
职位描述
机器翻译我们的使命是实现编码自动化。我们旅程的第一步是为专业程序员打造最好的工具,结合富有创造力的研究、设计和工程。我们的组织非常扁平,团队规模小且人才密集。我们尤其喜欢求真、充满热情且富有创造力的人。我们享受激烈的辩论、疯狂的想法,以及交付代码。
岗位职责
作为 SpaceXAI 模型路由与推理团队的软件工程师,你将构建支撑产品中每一次 AI 交互的推理平台。 该团队负责完整的推理路径:让 Cursor 的 AI 在世界上少有团队能够达到的规模下更快、更可靠、更具成本效益。每一个 agent 会话、每一次 tab 补全,以及每一条聊天消息,都会流经你的技术栈。 示例项目包括…… • 构建并演进我们的推理网关,这是一个对所有提供商 API 语义的统一抽象,使模型接入变成一次配置变更。
• 设计智能的跨提供商故障转移,使任何单一提供商的中断都不会造成用户可见的服务降级。
• 设计路由背压和准入控制,使流量高峰不会级联到各提供商。
你可能适合,如果 • 你在构建高吞吐、低延迟的分布式系统方面有深厚经验,尤其是在推理服务、流量路由或实时数据管道方面。
• 你能够自如地在大规模场景下权衡成本/性能取舍(GPU 利用率、提供商经济性、容量规划)。
• 你有扎实的软件工程基础,并乐于交付能够处理数百万请求的生产系统。
• 你能在灰色地带做出良好判断:当不存在唯一“正确”答案时,权衡可靠性、成本、延迟和用户体验。
申请 如果看起来合适,我们会联系你安排 2-3 场简短的技术面试。之后,我们会安排一次在我们办公室的现场面试,届时你将参与一个小项目、讨论想法,并与团队见面。 #LI-DNI
以上内容由机器翻译自动生成,可能存在错误;投递前请以雇主原文为准。
查看雇主原文
职位描述
Our mission is to automate coding. The first step in our journey is to build the best tool for professional programmers, using a combination of inventive research, design, and engineering. Our organization is very flat, and our team is small and talent dense. We particularly like people who are truth-seeking, passionate, and creative. We enjoy spirited debate, crazy ideas, and shipping code.
岗位职责
As a Software Engineer on the Model Routing & Inference team at SpaceXAI, you'll build the inference platform that powers every AI interaction in the product. This team owns the full inference path: making Cursor's AI faster, more reliable, and more cost-effective at a scale few teams in the world get to operate at. Every agent session, every tab completion, and every chat message flows through your stack. Example projects include... • Building and evolving our inference gateway, a single abstraction over every provider's API semantics, so model onboarding becomes a config change.
• Designing intelligent cross-provider failover so no single provider outage causes user-visible degradation.
• Designing routing backpressure and admission control so traffic spikes don't cascade into providers.
You may be a fit if • You have deep experience building high-throughput, low-latency distributed systems, especially in inference serving, traffic routing, or real-time data pipelines.
• You're comfortable reasoning about cost/performance tradeoffs at scale (GPU utilization, provider economics, capacity planning).
• You have strong software engineering fundamentals and enjoy shipping production systems that handle millions of requests.
• You make good calls in the gray area: weighing reliability, cost, latency, and user experience when there isn't a single "right" answer.
Applying If there appears to be a fit, we'll reach to schedule 2-3 short technicals. After, we'll schedule an onsite in our office, where you'll work on a small project, discuss ideas, and meet the team. #LI-DNI