Inference Engineer, Robotics
Remote • San Francisco • FullTime
Posted 1y ago
Remote Work Policy
Fully remote
Employment Type
FullTime
Categories
Applied AI Engineer
About the job
We are seeking a GPU Inference Engineer to enhance model serving efficiency for our Robotics research. This high-impact role involves driving initiatives to optimize inference performance and scalability, as well as assisting researchers in developing inference-friendly models. This position is crucial for scaling the team's goals, enabling leadership to focus on higher-leverage initiatives by building a stronger technical foundation.
Responsibilities
- Improve model serving, inference performance, and system efficiency through engineering efforts.
- Optimize kernel and data movement for enhanced system throughput and reliability.
- Collaborate with research and product teams to ensure effective model performance at scale.
- Design, build, and enhance critical serving infrastructure to support Robotics growth and reliability.
Requirements
- Deep expertise in model performance optimization, particularly at the inference layer.
- Strong background in kernel-level systems, data movement, and low-level performance tuning.
- Excitement for scaling high-performing AI systems serving real-world, multimodal workloads.
- Ability to navigate ambiguity, set technical direction, and drive complex initiatives to completion.
Benefits
- Relocation assistance to new employees.