Inference Engineer, Robotics

Remote San Francisco FullTime

Posted 1y ago

Job Location

San Francisco

Tech Stack

Remote Work Policy

Fully remote

Employment Type

FullTime

Categories

Applied AI Engineer

About the job

We are seeking a GPU Inference Engineer to enhance model serving efficiency for our Robotics research. This high-impact role involves driving initiatives to optimize inference performance and scalability, as well as assisting researchers in developing inference-friendly models. This position is crucial for scaling the team's goals, enabling leadership to focus on higher-leverage initiatives by building a stronger technical foundation.

Responsibilities

  • Improve model serving, inference performance, and system efficiency through engineering efforts.
  • Optimize kernel and data movement for enhanced system throughput and reliability.
  • Collaborate with research and product teams to ensure effective model performance at scale.
  • Design, build, and enhance critical serving infrastructure to support Robotics growth and reliability.

Requirements

  • Deep expertise in model performance optimization, particularly at the inference layer.
  • Strong background in kernel-level systems, data movement, and low-level performance tuning.
  • Excitement for scaling high-performing AI systems serving real-world, multimodal workloads.
  • Ability to navigate ambiguity, set technical direction, and drive complex initiatives to completion.

Benefits

  • Relocation assistance to new employees.

About OpenAI

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.