PyTorch Jobs

223 open roles mentioning PyTorch

Machine Learning Engineer, Reinforcement Learning

5mo ago
S

Skild

Skild AI is building the world's first general-purpose robotic intelligence that is robust and adapts to unseen scenarios without failing. We are seeking a Machine Learning Engineer to design and implement cutting-edge reinforcement learning algorithms for robotic applications. This role involves conducting experiments, optimizing models for real-world robotic environments, and collaborating closely with our robotics, research, and engineering teams. Your work will directly contribute to the development of intelligent, adaptable robots capable of autonomous learning and complex task performance.

Pittsburgh, San Mateo onsite
PythonC#PyTorch +4 more

Research Engineer, Post-training & Deployment

5mo ago
S

Skild

Skild AI is seeking a Research Engineer to join our post-training team. In this role, you will be responsible for enhancing Skild's foundation models and deploying them onto robots in real-world scenarios. You will collaborate with customers and utilize deployment data to ensure reliable robot behavior, focusing on safety, efficiency, and robustness under operational constraints. This position bridges the gap between strong lab performance and dependable customer deployments, defining the standard for scaling autonomous systems into global infrastructure.

San Mateo, CA onsite
PythonPyTorchTensorFlow +3 more

Member of Technical Staff (AI Infrastructure Engineer)

5mo ago
P

Perplexity AI

We are seeking an AI Infrastructure Engineer to join our expanding team. In this role, you will collaborate closely with our Inference and Research teams to construct, deploy, and enhance our extensive AI training and inference clusters. Your work will involve managing Kubernetes and Slurm environments, optimizing distributed training for large language models, and developing robust orchestration systems. This position offers the opportunity to significantly impact the performance and scalability of our AI infrastructure.

London hybrid FullTime
AWSKubernetesPython +5 more

Member of Technical Staff (AI Infrastructure Engineer)

5mo ago
P

Perplexity AI

We are seeking an AI Infrastructure Engineer to join our expanding team. In this role, you will collaborate closely with our Inference and Research teams to construct, deploy, and enhance our large-scale AI training and inference clusters. Our work involves Kubernetes, Slurm, Python, C++, PyTorch, and primarily operates on AWS.

San Francisco onsite FullTime
AWSKubernetesPython +5 more

Member of Technical Staff (AI Inference Engineer)

5mo ago
P

Perplexity AI

We are seeking an engineer to join our team responsible for building and running the inference engine behind Perplexity's queries. This role involves deploying dozens of model architectures at scale, managing tight latency and cost budgets, and working with a stack including Rust, Python, CUDA, and CuTe DSL. You will contribute to supporting new models, migrating GPU kernels, developing a Rust-native serving runtime, optimizing performance, and enhancing reliability and observability.

San Francisco onsite FullTime
KubernetesPythonRust +4 more

Member of Technical Staff (AI Inference Engineer)

5mo ago
P

Perplexity AI

We are seeking an AI Inference Engineer to join our dynamic team. This role is central to Perplexity's operations, as you will build and manage the inference engine that powers every query. You will deploy a variety of model architectures at scale, focusing on meeting stringent latency and cost requirements. Our technology stack includes Rust, Python, CUDA, and the CuTe DSL.

London onsite FullTime
KubernetesPythonRust +4 more

Internship - Search Machine Learning Engineer

5mo ago
P

Perplexity AI

Perplexity is seeking a Search Machine Learning Engineer Intern to contribute to the development of next-generation search technologies, specifically focusing on retrieval and ranking. This internship offers a hands-on opportunity to collaborate with experienced engineers, enhance search quality, experiment with novel models, and implement features that directly influence user search and information discovery experiences. The program is designed for a 12-24 week full-time engagement, conducted in person at our London office.

$12k - $24k

London onsite FullTime
PythonRustRAG +4 more

Software Engineer, Security

5mo ago
thinkingmachines

thinkingmachines

Thinking Machines Lab is seeking a Software Engineer focused on security to ensure their AI products are secure by default while enabling rapid product iteration. This role involves embedding with product and research teams to integrate security into the design and development process, as well as building tools and automation to maintain system safety at scale. The company is dedicated to advancing collaborative general intelligence and empowering users with AI tools tailored to their needs.

$350k - $475k

San Francisco onsite
OpenAIMistralPython +2 more

Software Engineer, Systems Generalist

5mo ago
thinkingmachines

thinkingmachines

Thinking Machines Lab is seeking generalist infrastructure and systems engineers to build the core systems powering their foundation models and support internal research and product development teams. This high-impact role involves architecting and scaling critical infrastructure across the full technical stack, solving complex distributed systems problems, and building robust, scalable platforms. You will work directly with researchers to accelerate experiments, improve infrastructure efficiency, and enable key insights across models, products, and data assets.

$350k - $475k

San Francisco onsite
OpenAIMistralKubernetes +5 more

Software Engineer, Workload Enablement

5mo ago
OpenAI

OpenAI

OpenAI is seeking a Software Engineer to join the Scaling team, which builds the architectural and engineering backbone for the company's infrastructure. This role focuses on enabling production workloads and end-to-end testing on new platforms. You will be responsible for creating test harnesses, developing platform stress benchmarks, and porting existing AI model training and inference workloads to new systems and hardware. A key aspect of this position involves analyzing performance, identifying bottlenecks, and characterizing the behavior of new compute, communication, storage, and control plane systems, including their failure modes.

San Francisco hybrid FullTime
OpenAIKubernetesPython +3 more

Applied AI, Technical Lead, Forward Deployed AI Engineer - Montreal

5mo ago
Mistral AI

Mistral AI

Mistral AI is seeking a Technical Lead, Applied AI to drive the technical strategy, execution, and delivery of complex AI solutions for enterprise customers. This role involves leading project teams of Applied AI Engineers to ensure successful deployment of Mistral AI products and the development of high-impact, scalable AI use cases. You will serve as the primary technical point of contact for strategic customers, guiding them through the entire lifecycle from pre-sales to post-implementation, and collaborating with research, product, and engineering teams to shape future offerings. The position bridges cutting-edge AI research with real-world enterprise applications, ensuring solutions are robust, scalable, and aligned with customer needs and Mistral's vision.

Montreal remote Full-time
MistralLangChainAWS +10 more

Senior Member of Technical Staff, Safety and Security for Agents

5mo ago
Cohere

Cohere

Cohere is seeking a Senior Member of Technical Staff to join the Safety and Security for Agents team. In this role, you will significantly contribute to the development of safer, fairer, more trustworthy, and more secure Large Language Models (LLMs). Your work will focus on data generation, post-training algorithms, and evaluation methods to ensure the safety of next-generation models that interact with external resources and take actions. You will collaborate closely with machine learning teams, data annotation teams, and product and policy teams, requiring a blend of machine learning expertise, ethical AI principles, experimental design, and data management skills. This position offers significant autonomy and decision-making power within a small team, with the opportunity to shape the future of LLMs for societal benefit.

London hybrid FullTime
CoherePythonPyTorch +3 more

Applied AI, Forward Deployed Machine Learning Engineer - Montreal

5mo ago
Mistral AI

Mistral AI

Mistral AI is seeking an Applied AI Engineer to join their team and facilitate customer adoption of their AI products. This role involves collaborating closely with customers to address complex technical challenges, from pre-sale to post-implementation, ensuring solutions meet and exceed client expectations. You will manage daily customer relations, act as a key resource for externalizing research into production, and work on state-of-the-art Generative AI applications across various industries. This position offers the opportunity to contribute to a pioneering company shaping the future of AI and make a meaningful impact.

Montreal onsite Full-time
MistralLangChainPython +9 more

Senior Staff Software Engineer, AI Model Lifecycle

5mo ago
c

crusoe

Crusoe is seeking a Senior Staff Software Engineer for the AI Model Lifecycle team to build a comprehensive managed platform for the entire application development lifecycle, with a specific focus on leveraging Machine Learning models, including Large Language Models (LLMs). This role is crucial in accelerating the abundance of energy and intelligence by powering the world's most ambitious AI workloads. You will join a team building the future of AI infrastructure, solving the bottleneck of power for AI compute with an energy-first approach. We are looking for problem-solving, opportunity-finding teammates with a sense of urgency who thrive on building the path forward.

$238k - $318k

San Francisco, CA - US onsite FullTime
PythonGoFine-Tuning +2 more

Generative AI Inference Engineer

5mo ago
Stability AI

Stability AI

We are seeking passionate Machine Learning Engineers to join our Inference team, focusing on the creative applications of generative AI models. The ideal candidate will have substantial experience developing and running inference for multi-modal models. A deep understanding of diffusion model architectures and familiarity with workflow tools like ComfyUI are a big plus. You will be expected to leverage and push the boundaries of state-of-the-art inference optimization techniques for multi-modal generative models. This role offers the opportunity to work alongside top researchers and engineers, utilizing cutting-edge high-performance computing resources to make a significant impact in the rapidly evolving field of generative AI.

United States remote
AWSAzureDocker +10 more

Senior Software Engineer, AI Model Lifecycle

5mo ago
c

crusoe

Crusoe is seeking a Senior Software Engineer for their AI Model Lifecycle team to build a comprehensive managed platform for the entire application development lifecycle, with a specific focus on leveraging Machine Learning models, including Large Language Models (LLMs). This role is crucial in accelerating the abundance of energy and intelligence by powering ambitious AI workloads. The ideal candidate will be a problem-solving, opportunity-finding teammate with a sense of urgency, who believes in the scale of Crusoe's ambition and thrives on a path not fully paved.

$172k - $231k

San Francisco, CA - US onsite FullTime
PythonGoFine-Tuning +2 more

Research Scientist – Controlled 3D Generation

5mo ago
Stability AI

Stability AI

We are seeking a Research Scientist passionate about 3D generation, flow matching, and diffusion models. You will help advance the frontier of controllable 3D content creation by building models that generate consistent, editable, and physically grounded 3D assets and scenes. This role involves conducting cutting-edge research, designing and implementing scalable training pipelines, and developing techniques for conditioning and control. You will analyze model behavior, collaborate with cross-disciplinary teams to translate research into production-ready systems, and publish results at top-tier venues.

Remote remote
PyTorchJAXCUDA +7 more

Machine Learning Engineer, Integrity

6mo ago
OpenAI

OpenAI

As a Machine Learning Engineer in OpenAI's Integrity team, you will work on state-of-the-art models and classifiers, experiment with new architectures and approaches, and advance our capabilities in content and user understanding. This role offers the chance to transform research breakthroughs into tangible solutions that enhance the trust and safety of our platform, with a focus on training LLMs and building ML models. You will be instrumental in designing and deploying advanced machine learning models to solve real-world problems, bringing OpenAI's research from concept to implementation and creating AI-driven applications with direct impact.

San Francisco onsite FullTime
OpenAIFine-TuningPyTorch +3 more

Member of Technical Staff, Software Engineer

6mo ago
f

fireworks ai

Fireworks is seeking a Member of Technical Staff, Software Engineer to join their team. This role involves building the core backend systems that power Fireworks' platform for specialized intelligence, enabling companies to build, train, and serve AI models tailored to their own data, workflows, and products. You will own major product surfaces from architecture to production, improving reliability, performance, and developer experience. This is platform engineering with product impact, where your systems will directly shape how customers build on top of AI. You will work closely with product, frontend, infra, and GTM to ship end-to-end features, and use AI tooling aggressively to automate tasks.

San Mateo hybrid FullTime
Fine-TuningPyTorchModel Serving +1 more

Data Infrastructure Engineer

7mo ago
H

Heyge

At HeyGen, we are at the forefront of developing applications powered by our cutting-edge AI research. As a Data Infrastructure Engineer, you will lead the development of fundamental data systems and infrastructure. These systems are essential for powering our innovative applications, including Avatar IV, Photo Avatar, Instant Avatar, Interactive Avatar, and Video Translation. Your role will be crucial in enhancing the efficiency and scalability of these systems, which are vital to HeyGen's success.

Los Angeles, Palo Alto, San Francisco, Toronto onsite
PythonPyTorchTensorFlow

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.