Rust Jobs
179 open roles mentioning Rust
Software Engineer, GPT Infrastructure
OpenAI
We are seeking a software engineer to help build the platform that qualifies and optimizes inference workloads across heterogeneous compute environments. You will develop both OpenAI-hosted services and secure partner-side software for running long-lived optimization workflows. These workflows generate candidate kernels, runtime configurations, and serving-stack changes; compile and execute them on target hardware; verify their correctness; measure their performance; and use the results to guide further optimization. You will work across model architecture, distributed execution, compilers, runtimes, networking, and accelerator systems. A central part of the role is turning research prototypes and one-off hardware bring-up efforts into reliable, reusable infrastructure with clear contracts, reproducible results, strong observability, and well-defined security boundaries.
Member of Technical Staff (Software Engineer, Computer)
Perplexity AI
Perplexity is seeking builders to join its highly leveraged engineering team to create innovative products that accelerate human productivity. As an engineer, you will build products and systems that define how AI empowers humans to think, decide, and act. You will ship fast and work across product, infrastructure, and frontier-level advancements in technology, focusing on craftsmanship, ownership, entrepreneurship, scholarship, and partnership. This role is for individuals who identify problems, design solutions, and ship them, acting with urgency and a founder's mindset to deliver for users.
Software Engineer, Applied AI
Mercor
As a Software Engineer on Applied AI, you will build, deploy, and operate systems that bridge frontier AI research and data delivery. This is a high-ownership, deeply technical role where you will tackle ill-defined problems, prototype rapidly with researchers and customers, and transition systems from early experiments to reliable, scalable production. You will manage projects end-to-end, including requirements gathering, creating data pipelines, and enhancing model-adjacent infrastructure, while collaborating with frontier AI labs and internal teams to deliver impactful applied AI solutions.
Backend Engineer, Chanakya
sarvam
Sarvam is building India's full-stack sovereign AI platform, focusing on research, models, infrastructure, and applications to make AI work for India. As a Backend Engineer, you will design and implement the core system infrastructure powering Sarvam's 'atoms', including MCP servers, on-prem deployment tooling, document ingestion pipelines, agentic backends, and API layers. These systems operate in constrained environments, and the architecture decisions you make will have a lasting impact. You will ship real systems and be accountable for their performance in the field.
Senior Full-Stack Engineer
Nabla
Nabla is seeking a Senior Full-Stack Engineer to join a cross-functional squad and contribute end-to-end to a specific part of the product. This role involves taking a leading role in building and scaling Nabla’s AI-powered healthcare platform, working across the stack from back-end services to web, desktop, and mobile applications. You will collaborate closely with Product, Design, and Machine Learning teams to deliver high-impact features that directly serve healthcare professionals, with the goal of restoring the human connection at the heart of healthcare by streamlining clinical documentation.
Member of Technical Staff (AI Inference Engineer)
Perplexity AI
We are seeking an engineer to join our team responsible for building and running the inference engine behind Perplexity's queries. This role involves deploying dozens of model architectures at scale, managing tight latency and cost budgets, and working with a stack including Rust, Python, CUDA, and CuTe DSL. You will contribute to supporting new models, migrating GPU kernels, developing a Rust-native serving runtime, optimizing performance, and enhancing reliability and observability.
Member of Technical Staff (AI Inference Engineer)
Perplexity AI
We are seeking an AI Inference Engineer to join our dynamic team. This role is central to Perplexity's operations, as you will build and manage the inference engine that powers every query. You will deploy a variety of model architectures at scale, focusing on meeting stringent latency and cost requirements. Our technology stack includes Rust, Python, CUDA, and the CuTe DSL.
Internship - Search Machine Learning Engineer
Perplexity AI
Perplexity is seeking a Search Machine Learning Engineer Intern to contribute to the development of next-generation search technologies, specifically focusing on retrieval and ranking. This internship offers a hands-on opportunity to collaborate with experienced engineers, enhance search quality, experiment with novel models, and implement features that directly influence user search and information discovery experiences. The program is designed for a 12-24 week full-time engagement, conducted in person at our London office.
$12k - $24k
Staff / Principal Machine Learning Engineer, Serving - Switzerland
inworld
Inworld is seeking a Staff/Principal Machine Learning Engineer to join their research lab focused on building top-ranked real-time voice models. These models power large consumer-facing AI applications across various sectors. The role involves optimizing real-time inference, developing best-in-class APIs and products, and contributing to the research and development of state-of-the-art models. The ideal candidate is a fast learner who thrives in ambiguity and can demonstrate a strong portfolio of built, broken, and understood systems. This position emphasizes impact, shipping stable code, and a deep understanding of the underlying logic behind engineering decisions.
Staff / Principal Machine Learning Engineer, Serving - UK
inworld
Inworld is seeking a Staff/Principal Machine Learning Engineer to join their research lab focused on building top-ranked real-time voice models. These models power large consumer-facing AI applications across various sectors, reaching hundreds of millions of end-users. The role involves optimizing real-time inference, developing state-of-the-art models, and creating best-in-class APIs and products. The ideal candidate thrives in ambiguity, learns quickly, and can demonstrate a strong track record of building, breaking, and understanding complex systems. This position emphasizes impact, shipping stable and reliable systems, and a deep understanding of the underlying logic behind engineering decisions.
£140k - £200k
Senior / Lead Machine Learning Engineer, Serving - Serbia
inworld
Inworld is a research lab building top-ranked real-time voice models used in large consumer-facing AI applications. We are seeking a Senior/Lead Machine Learning Engineer to optimize real-time inference and create best-in-class APIs and products. This role involves tackling complex, ambiguous problems and driving impact through shipped, stable systems. We value engineers who are comfortable with ambiguity, have a bias for action, and obsess over performance, latency, and reliability as core product features.
Senior / Lead Machine Learning Engineer, Serving - Germany
inworld
Inworld is seeking a Senior / Lead Machine Learning Engineer to join our team in Germany. We are a research lab building top-ranked real-time voice models used in large-scale AI applications across various sectors. Our work involves research, development, optimizing real-time inference, and creating best-in-class APIs and products. We are looking for fast learners with a strong ability to build, break, and understand systems, who thrive in ambiguity and can demonstrate their past work. This role emphasizes full-cycle ownership, taking models from research to production-ready serving.
Staff / Principal Machine Learning Engineer, Serving - USA
inworld
Inworld is seeking a Staff/Principal Machine Learning Engineer to join their team in Mountain View, USA. This role focuses on optimizing real-time inference for state-of-the-art AI models, which power large consumer-facing applications. The ideal candidate will have a strong background in systems programming and a proven ability to take models from research to production, ensuring reliability and performance. The company values engineers who can navigate ambiguity, drive impact, and deeply understand the underlying logic of their work, fostering a collaborative and fast-paced environment.
$270k - $500k
Software Engineer, Security
thinkingmachines
Thinking Machines Lab is seeking a Software Engineer focused on security to ensure their AI products are secure by default while enabling rapid product iteration. This role involves embedding with product and research teams to integrate security into the design and development process, as well as building tools and automation to maintain system safety at scale. The company is dedicated to advancing collaborative general intelligence and empowering users with AI tools tailored to their needs.
$350k - $475k
Software Engineer, Systems Generalist
thinkingmachines
Thinking Machines Lab is seeking generalist infrastructure and systems engineers to build the core systems powering their foundation models and support internal research and product development teams. This high-impact role involves architecting and scaling critical infrastructure across the full technical stack, solving complex distributed systems problems, and building robust, scalable platforms. You will work directly with researchers to accelerate experiments, improve infrastructure efficiency, and enable key insights across models, products, and data assets.
$350k - $475k
Software Engineer, Codex Core Agents
OpenAI
We are seeking engineers to build the core infrastructure that powers Codex agents in production. This role involves designing and operating systems for sandboxed execution, orchestration, stateful workflows, and reliable, efficient operation at scale. You will work at the intersection of distributed systems, developer tooling, and AI, building primitives that enhance the speed, safety, reliability, and ease of use of Codex for the organization.
Software Engineer, Site Reliability
Hebbia
Hebbia is seeking a Site Reliability Engineer who approaches the role with a software engineering mindset. You will be responsible for the entire lifecycle of critical production systems, focusing on design, development, and enhancement rather than just operation. This involves writing production-quality code to ensure platform reliability at scale, collaborating with product engineering teams to integrate reliability into architectural decisions from the outset, and building essential internal tooling for all engineers. The role emphasizes coding, including instrumenting services, optimizing performance, developing deployment platforms, and translating incident learnings into permanent architectural improvements.
$160k - $350k
Software Engineer - GPU Networking & Distributed Systems
Baseten
Baseten is building the global operating system for distributed, heterogeneous AI hardware, powering mission-critical inference for leading AI companies. As LLM and multi-modal workloads scale, the network becomes a critical component. This role focuses on leading GPU Networking efforts, making RDMA a first-class building block and optimizing distributed inference. You will architect the software fabric that unifies thousands of GPUs, co-optimizing communication and computation for advanced AI workloads.
Systems Engineering Manager
Modal
Modal is building the next-generation infrastructure layer for AI, enabling instant GPU access, sub-second container starts, and native storage for customers like Lovable, Ramp, Cognition, DoorDash, and Suno. We are seeking an Engineering Manager to lead a team of experienced engineers focused on our serverless GPU platform. This is a hands-on leadership role where you will balance technical contributions with people management, setting direction, removing obstacles, and cultivating a strong engineering culture while addressing complex challenges in distributed computing, large-scale data handling, and performance optimization.
Software Engineer, ML Infrastructure
cursor
The ML Infrastructure team builds large-scale compute, storage, and software infrastructure to support the company's work building the world's best agentic coding model. This role works closely with ML researchers and engineers to enable their work through improvements to our training framework, systems reliability/performance, and developer experience. We are looking for strong engineers interested in building high-performance infrastructure and the software to support it.