Rust Jobs

179 open roles mentioning Rust

Software Engineer, GPT Infrastructure

4mo ago
OpenAI

OpenAI

We are seeking a software engineer to help build the platform that qualifies and optimizes inference workloads across heterogeneous compute environments. You will develop both OpenAI-hosted services and secure partner-side software for running long-lived optimization workflows. These workflows generate candidate kernels, runtime configurations, and serving-stack changes; compile and execute them on target hardware; verify their correctness; measure their performance; and use the results to guide further optimization. You will work across model architecture, distributed execution, compilers, runtimes, networking, and accelerator systems. A central part of the role is turning research prototypes and one-off hardware bring-up efforts into reliable, reusable infrastructure with clear contracts, reproducible results, strong observability, and well-defined security boundaries.

San Francisco hybrid FullTime
OpenAIPythonGo +2 more

Member of Technical Staff (Software Engineer, Computer)

4mo ago
P

Perplexity AI

Perplexity is seeking builders to join its highly leveraged engineering team to create innovative products that accelerate human productivity. As an engineer, you will build products and systems that define how AI empowers humans to think, decide, and act. You will ship fast and work across product, infrastructure, and frontier-level advancements in technology, focusing on craftsmanship, ownership, entrepreneurship, scholarship, and partnership. This role is for individuals who identify problems, design solutions, and ship them, acting with urgency and a founder's mindset to deliver for users.

San Francisco onsite FullTime
PythonTypeScriptJavaScript +4 more

Software Engineer, Applied AI

4mo ago
M

Mercor

As a Software Engineer on Applied AI, you will build, deploy, and operate systems that bridge frontier AI research and data delivery. This is a high-ownership, deeply technical role where you will tackle ill-defined problems, prototype rapidly with researchers and customers, and transition systems from early experiments to reliable, scalable production. You will manage projects end-to-end, including requirements gathering, creating data pipelines, and enhancing model-adjacent infrastructure, while collaborating with frontier AI labs and internal teams to deliver impactful applied AI solutions.

San Francisco onsite FullTime
PythonGoRust

Backend Engineer, Chanakya

5mo ago
s

sarvam

Sarvam is building India's full-stack sovereign AI platform, focusing on research, models, infrastructure, and applications to make AI work for India. As a Backend Engineer, you will design and implement the core system infrastructure powering Sarvam's 'atoms', including MCP servers, on-prem deployment tooling, document ingestion pipelines, agentic backends, and API layers. These systems operate in constrained environments, and the architecture decisions you make will have a lasting impact. You will ship real systems and be accountable for their performance in the field.

Bengaluru onsite FullTime
LangChainLlamaIndexWeaviate +5 more

Senior Full-Stack Engineer

5mo ago
N

Nabla

Nabla is seeking a Senior Full-Stack Engineer to join a cross-functional squad and contribute end-to-end to a specific part of the product. This role involves taking a leading role in building and scaling Nabla’s AI-powered healthcare platform, working across the stack from back-end services to web, desktop, and mobile applications. You will collaborate closely with Product, Design, and Machine Learning teams to deliver high-impact features that directly serve healthcare professionals, with the goal of restoring the human connection at the heart of healthcare by streamlining clinical documentation.

Paris office onsite FullTime
KubernetesPythonTypeScript +5 more

Member of Technical Staff (AI Inference Engineer)

5mo ago
P

Perplexity AI

We are seeking an engineer to join our team responsible for building and running the inference engine behind Perplexity's queries. This role involves deploying dozens of model architectures at scale, managing tight latency and cost budgets, and working with a stack including Rust, Python, CUDA, and CuTe DSL. You will contribute to supporting new models, migrating GPU kernels, developing a Rust-native serving runtime, optimizing performance, and enhancing reliability and observability.

San Francisco onsite FullTime
KubernetesPythonRust +4 more

Member of Technical Staff (AI Inference Engineer)

5mo ago
P

Perplexity AI

We are seeking an AI Inference Engineer to join our dynamic team. This role is central to Perplexity's operations, as you will build and manage the inference engine that powers every query. You will deploy a variety of model architectures at scale, focusing on meeting stringent latency and cost requirements. Our technology stack includes Rust, Python, CUDA, and the CuTe DSL.

London onsite FullTime
KubernetesPythonRust +4 more

Internship - Search Machine Learning Engineer

5mo ago
P

Perplexity AI

Perplexity is seeking a Search Machine Learning Engineer Intern to contribute to the development of next-generation search technologies, specifically focusing on retrieval and ranking. This internship offers a hands-on opportunity to collaborate with experienced engineers, enhance search quality, experiment with novel models, and implement features that directly influence user search and information discovery experiences. The program is designed for a 12-24 week full-time engagement, conducted in person at our London office.

$12k - $24k

London onsite FullTime
PythonRustRAG +4 more

Staff / Principal Machine Learning Engineer, Serving - Switzerland

5mo ago
i

inworld

Inworld is seeking a Staff/Principal Machine Learning Engineer to join their research lab focused on building top-ranked real-time voice models. These models power large consumer-facing AI applications across various sectors. The role involves optimizing real-time inference, developing best-in-class APIs and products, and contributing to the research and development of state-of-the-art models. The ideal candidate is a fast learner who thrives in ambiguity and can demonstrate a strong portfolio of built, broken, and understood systems. This position emphasizes impact, shipping stable code, and a deep understanding of the underlying logic behind engineering decisions.

Switzerland remote FullTime
KubernetesPythonGo +2 more

Staff / Principal Machine Learning Engineer, Serving - UK

5mo ago
i

inworld

Inworld is seeking a Staff/Principal Machine Learning Engineer to join their research lab focused on building top-ranked real-time voice models. These models power large consumer-facing AI applications across various sectors, reaching hundreds of millions of end-users. The role involves optimizing real-time inference, developing state-of-the-art models, and creating best-in-class APIs and products. The ideal candidate thrives in ambiguity, learns quickly, and can demonstrate a strong track record of building, breaking, and understanding complex systems. This position emphasizes impact, shipping stable and reliable systems, and a deep understanding of the underlying logic behind engineering decisions.

£140k - £200k

UK onsite FullTime
KubernetesPythonGo +2 more

Senior / Lead Machine Learning Engineer, Serving - Serbia

5mo ago
i

inworld

Inworld is a research lab building top-ranked real-time voice models used in large consumer-facing AI applications. We are seeking a Senior/Lead Machine Learning Engineer to optimize real-time inference and create best-in-class APIs and products. This role involves tackling complex, ambiguous problems and driving impact through shipped, stable systems. We value engineers who are comfortable with ambiguity, have a bias for action, and obsess over performance, latency, and reliability as core product features.

Serbia onsite FullTime
KubernetesPythonGo +2 more

Senior / Lead Machine Learning Engineer, Serving - Germany

5mo ago
i

inworld

Inworld is seeking a Senior / Lead Machine Learning Engineer to join our team in Germany. We are a research lab building top-ranked real-time voice models used in large-scale AI applications across various sectors. Our work involves research, development, optimizing real-time inference, and creating best-in-class APIs and products. We are looking for fast learners with a strong ability to build, break, and understand systems, who thrive in ambiguity and can demonstrate their past work. This role emphasizes full-cycle ownership, taking models from research to production-ready serving.

Germany onsite FullTime
KubernetesPythonGo +2 more

Staff / Principal Machine Learning Engineer, Serving - USA

5mo ago
i

inworld

Inworld is seeking a Staff/Principal Machine Learning Engineer to join their team in Mountain View, USA. This role focuses on optimizing real-time inference for state-of-the-art AI models, which power large consumer-facing applications. The ideal candidate will have a strong background in systems programming and a proven ability to take models from research to production, ensuring reliability and performance. The company values engineers who can navigate ambiguity, drive impact, and deeply understand the underlying logic of their work, fostering a collaborative and fast-paced environment.

$270k - $500k

Mountain View, California, USA hybrid FullTime
KubernetesPythonGo +2 more

Software Engineer, Security

5mo ago
thinkingmachines

thinkingmachines

Thinking Machines Lab is seeking a Software Engineer focused on security to ensure their AI products are secure by default while enabling rapid product iteration. This role involves embedding with product and research teams to integrate security into the design and development process, as well as building tools and automation to maintain system safety at scale. The company is dedicated to advancing collaborative general intelligence and empowering users with AI tools tailored to their needs.

$350k - $475k

San Francisco onsite
OpenAIMistralPython +2 more

Software Engineer, Systems Generalist

5mo ago
thinkingmachines

thinkingmachines

Thinking Machines Lab is seeking generalist infrastructure and systems engineers to build the core systems powering their foundation models and support internal research and product development teams. This high-impact role involves architecting and scaling critical infrastructure across the full technical stack, solving complex distributed systems problems, and building robust, scalable platforms. You will work directly with researchers to accelerate experiments, improve infrastructure efficiency, and enable key insights across models, products, and data assets.

$350k - $475k

San Francisco onsite
OpenAIMistralKubernetes +5 more

Software Engineer, Codex Core Agents

5mo ago
OpenAI

OpenAI

We are seeking engineers to build the core infrastructure that powers Codex agents in production. This role involves designing and operating systems for sandboxed execution, orchestration, stateful workflows, and reliable, efficient operation at scale. You will work at the intersection of distributed systems, developer tooling, and AI, building primitives that enhance the speed, safety, reliability, and ease of use of Codex for the organization.

San Francisco onsite FullTime
OpenAIRustAI Agents

Software Engineer, Site Reliability

6mo ago
H

Hebbia

Hebbia is seeking a Site Reliability Engineer who approaches the role with a software engineering mindset. You will be responsible for the entire lifecycle of critical production systems, focusing on design, development, and enhancement rather than just operation. This involves writing production-quality code to ensure platform reliability at scale, collaborating with product engineering teams to integrate reliability into architectural decisions from the outset, and building essential internal tooling for all engineers. The role emphasizes coding, including instrumenting services, optimizing performance, developing deployment platforms, and translating incident learnings into permanent architectural improvements.

$160k - $350k

NYC onsite FullTime
AWSPythonGo +2 more

Software Engineer - GPU Networking & Distributed Systems

6mo ago
B

Baseten

Baseten is building the global operating system for distributed, heterogeneous AI hardware, powering mission-critical inference for leading AI companies. As LLM and multi-modal workloads scale, the network becomes a critical component. This role focuses on leading GPU Networking efforts, making RDMA a first-class building block and optimizing distributed inference. You will architect the software fabric that unifies thousands of GPUs, co-optimizing communication and computation for advanced AI workloads.

San Francisco hybrid FullTime
KubernetesPythonGo +3 more

Systems Engineering Manager

7mo ago
M

Modal

Modal is building the next-generation infrastructure layer for AI, enabling instant GPU access, sub-second container starts, and native storage for customers like Lovable, Ramp, Cognition, DoorDash, and Suno. We are seeking an Engineering Manager to lead a team of experienced engineers focused on our serverless GPU platform. This is a hands-on leadership role where you will balance technical contributions with people management, setting direction, removing obstacles, and cultivating a strong engineering culture while addressing complex challenges in distributed computing, large-scale data handling, and performance optimization.

New York onsite FullTime
RustJavaC#

Software Engineer, ML Infrastructure

7mo ago
c

cursor

The ML Infrastructure team builds large-scale compute, storage, and software infrastructure to support the company's work building the world's best agentic coding model. This role works closely with ML researchers and engineers to enable their work through improvements to our training framework, systems reliability/performance, and developer experience. We are looking for strong engineers interested in building high-performance infrastructure and the software to support it.

San Francisco onsite FullTime
KubernetesPythonTypeScript +2 more

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.