PyTorch Jobs

223 open roles mentioning PyTorch

Applied AI, Technical Lead, Forward Deployed AI Engineer - Abu Dhabi

4mo ago
Mistral AI

Mistral AI

Mistral AI is seeking a Technical Lead, Applied AI to drive the technical strategy, execution, and delivery of complex AI solutions for enterprise customers. This role involves leading project teams of Applied AI Engineers to ensure successful deployment of Mistral AI products and the development of high-impact, scalable AI use cases. You will serve as the primary technical point of contact for strategic customers, guiding them through the entire lifecycle from pre-sales to post-implementation, while collaborating with research, product, and engineering teams to shape future offerings. The position bridges cutting-edge AI research with real-world enterprise applications, ensuring solutions are robust, scalable, and aligned with customer needs and Mistral's technological vision.

Abu Dhabi onsite Full-time
MistralLangChainAWS +10 more

Research, Pre-Training Data

4mo ago
thinkingmachines

thinkingmachines

Thinking Machines Lab is seeking pre-training researchers to join their mission of advancing collaborative general intelligence. This role is central to developing the next generation of AI models by blending research with large-scale data engineering. You will be responsible for assembling pre-training datasets and data systems, designing and implementing methods for sourcing, curating, and analyzing data for quality and performance. The position involves working with automated pipelines and human-in-the-loop processes, contributing both scientific insights and production-grade code. It's an ideal opportunity for individuals passionate about the intersection of data, machine learning, and systems, and who are eager to shape the future of AI.

$350k - $475k

San Francisco onsite
OpenAIMistralPython +5 more

Research, Post-Training Data

4mo ago
thinkingmachines

thinkingmachines

Thinking Machines Lab is seeking researchers to bridge the gap between raw AI intelligence and useful, safe, and collaborative systems. This role focuses on post-training data research, combining human insight and machine learning techniques to capture and steer model behavior based on human preferences. You will be responsible for translating research ideas into actionable data through labeling and collection campaigns, understanding data quality science, and developing metrics to measure the impact of data and training interventions. The position also involves exploring new paradigms for human-AI interaction and scalable oversight, blending research, data operations, and technical implementation to advance human-centered AI systems. This role requires both fundamental research and practical engineering, making it ideal for individuals who enjoy deep theoretical exploration and hands-on experimentation.

$350k - $475k

San Francisco onsite
OpenAIMistralPython +5 more

Research Engineer, Infrastructure, Numerics

4mo ago
thinkingmachines

thinkingmachines

Thinking Machines Lab is seeking an infrastructure research engineer to design and build core systems for efficient large-scale model training, with a specific focus on numerics. This role involves enhancing the numerical foundations of their distributed training stack, optimizing precision formats, kernel optimizations, and communication frameworks to ensure stable, scalable, and fast training of trillion-parameter models. The ideal candidate will bridge research and systems engineering, possessing a strong understanding of both optimization mathematics and distributed compute realities.

$350k - $475k

San Francisco onsite
OpenAIMistralPyTorch +3 more

Research Engineer, Infrastructure, Kernels

4mo ago
thinkingmachines

thinkingmachines

Thinking Machines Lab is seeking an infrastructure research engineer to design, optimize, and maintain the compute foundations for large-scale language model training. This role involves developing high-performance ML kernels, enabling efficient low-precision arithmetic, and improving the distributed compute stack. You will work closely with researchers and systems architects, bridging algorithmic design with hardware efficiency, prototyping new kernel implementations, and defining numerical and parallelism strategies for scaling AI systems.

$350k - $475k

San Francisco onsite
OpenAIMistralPyTorch +2 more

Infrastructure Engineer, Security

4mo ago
thinkingmachines

thinkingmachines

Thinking Machines Lab is seeking an infrastructure engineer to lead and enhance the security infrastructure for their foundation models. This role involves working across compute, storage, networking, and data platforms to ensure systems are secure, reliable, and scalable. The engineer will define security controls, architecture, and tooling, integrating security by default into the platform. Collaboration with research and product teams will be key to enabling rapid progress while maintaining robust protection for models, data, and environments.

$200k - $475k

San Francisco onsite
OpenAIMistralKubernetes +4 more

Research Internship (Winter 2027)

4mo ago
Cohere

Cohere

Cohere is seeking a Research Intern to collaborate with researchers and tools on designing and implementing novel research ideas and shipping state-of-the-art models to production. Interns will have the opportunity to work on various teams covering base model training, retrieval augmented generation, data and evaluation, safety, and finetuning, or any research area relating to LLMs. This role offers a chance to broaden research connections while gaining deep experience in a growing AI startup.

Canada remote Intern
CoherePythonC# +5 more

Senior Front-End (React) Engineer

4mo ago
N

Nabla

Nabla is seeking a Senior Front-End Engineer to join a cross-functional squad focused on developing high-impact user experiences across web, desktop, browser extension, and mobile interfaces. This role involves setting UI quality and performance standards, evolving the design system, architecture, and tooling. You will lead the development of complex front-end features, drive technical decisions, partner with Design to create intuitive experiences, and collaborate with Product, ML, and Back-End teams to ship full-stack features. The position also includes mentoring other engineers and contributing to team best practices. Nabla is an AI company dedicated to improving healthcare by streamlining clinical documentation and workflows, backed by significant funding and led by experienced AI engineers.

Paris office onsite FullTime
KubernetesPythonTypeScript +5 more

Associate Product Manager

4mo ago
f

fireworks ai

This is a rare opportunity for an early-career product manager to work at the frontier of AI infrastructure, building tools used daily by developers and AI teams at the world's most ambitious companies. Inspired by Google’s APM program, APMs at Fireworks will receive mentorship from experienced PMs while also getting rotations across both Fireworks’ product areas and tasks. This role is designed to help you develop the foundational skills of a start-up product leader: rigorous thinking, user empathy, technical depth, and cross-functional leadership.

San Mateo hybrid FullTime
Fine-TuningPyTorchModel Serving +1 more

Applied Scientist / Research Engineer, AI4Engineering - EMEA

4mo ago
Mistral AI

Mistral AI

Mistral AI is seeking an Applied Scientist with deep expertise in engineering sciences to work at the forefront of AI-accelerated simulation. This role involves collaborating with industrial customers and internal research teams to build and deploy AI Physics Models, complementing our existing Large Language Models (LLMs). You will be involved in the entire process, from curating high-fidelity simulation datasets and training/evaluating models to delivering production-grade AI solutions directly to engineering teams. The target domains include computational fluid dynamics, structural mechanics, semiconductor design, multi-physics modeling, and digital twins. Working cross-functionally, you will ensure our models meet stringent engineering standards beyond just benchmark metrics.

Paris hybrid Full-time
MistralPythonRAG +9 more

Software Engineer - Voice AI (Inference Runtime)

4mo ago
B

Baseten

Baseten is seeking a highly impactful individual to lead the Voice AI product area, owning the end-to-end development and implementation of their in-house inference stack for Voice AI models. This role involves partnering closely with various engineering teams to push the boundaries of Voice AI, making a significant impact on industries like productivity, customer service, and education. You will be responsible for bringing state-of-the-art open-source models into production, focusing on optimizing model serving for latency, throughput, and GPU efficiency, and building large-scale, real-time infrastructure for multi-model voice agents.

San Francisco hybrid FullTime
DockerKubernetesPython +5 more

Applied AI, Forward Deployed Machine Learning Engineer, Critical and Sovereign Institutions, EMEA

4mo ago
Mistral AI

Mistral AI

Mistral AI is seeking an Applied AI Engineer to join its specialized unit focused on delivering high-impact, secure AI solutions for critical and sovereign institutions. This team works closely with clients in highly regulated environments to design, deploy, and maintain AI systems that meet stringent standards for reliability, security, and operational excellence. The role involves technical design, implementation, and deployment of AI solutions tailored to the unique needs of critical infrastructure and sovereign institutions, contributing directly to projects with significant societal and operational impact.

Paris onsite Full-time
MistralLangChainAWS +11 more

Clinician Scientist

4mo ago
Abridge

Abridge

Abridge is seeking Clinician Scientists to advance the development of its AI-powered clinical tools. This role requires a blend of deep clinical expertise and a background in AI, focusing on shaping AI-driven tools to ensure accuracy and quality for clinicians and patients. You will collaborate closely with engineers, researchers, product managers, and fellow clinicians to refine AI models, validate outputs, and create new functionalities that enhance documentation workflows.

SF Office hybrid FullTime
PythonPrompt EngineeringPyTorch +3 more

Member of Engineering (Post-training)

4mo ago
p

poolside

Poolside is building a company to create Artificial General Intelligence, aiming to accelerate software development through agentic systems, coding assistants, and frontier models. This role is part of the Applied Research team, focused on transforming pre-trained Large Language Models (LLMs) into well-aligned and highly capable AI systems specifically for coding and software development. You will be involved in building data pipelines and environments for agentic use cases, researching and implementing post-training algorithms, and designing experiments to test hypotheses, with access to significant GPU resources.

Remote (EMEA/East Coast) remote FullTime
PythonFine-TuningPyTorch +4 more

Member of Technical Staff - Post Training

4mo ago
B

Black Forest Labs

We are seeking a Member of Technical Staff specializing in Post Training to own the end-to-end post-training pipeline for our multimodal generative models. This role is crucial for transforming foundation models into polished products, encompassing data strategy, reward modeling, preference optimization, distillation, and safety tuning across image, editing, and video modalities. You will be instrumental in driving significant improvements in model quality, developing the infrastructure that accelerates research team iteration, and advancing the state-of-the-art in aligning generative models with human intent. This is a Staff/Senior individual contributor position for someone with proven experience shipping post-training for a frontier model.

€130k - €340k

Freiburg (Germany) onsite FullTime
Fine-TuningPyTorchRLHF

Member of Technical Staff - Pretraining

4mo ago
B

Black Forest Labs

We are seeking a Member of Technical Staff focused on pretraining to join our research team. This role is central to developing the next generation of multimodal foundation models for image, video, and audio. You will have a significant impact on shaping training objectives, architectures, data strategies, and systems, with your research directly influencing products used by millions. This is a Staff/Senior individual contributor position for someone with proven experience leading frontier pretraining efforts.

€130k - €340k

Freiburg (Germany) onsite FullTime
PythonPyTorchDistributed Training

Researcher, Vision

4mo ago
s

sarvam

Sarvam is building the bedrock of Sovereign AI for India, developing a full-stack sovereign AI platform across research, models, infrastructure, and applications. The company partners with leading enterprises and public institutions, backed by prominent venture capital firms and collaborating with India's top brands. This role involves working across the full lifecycle of vision-language model (VLM) development, including data, training, evaluation, and production. We seek researchers who are comfortable with the evolving nature of the field and can take a leading role.

Bengaluru onsite FullTime
PyTorchNLPRLHF +1 more

Senior Full-Stack Engineer

5mo ago
N

Nabla

Nabla is seeking a Senior Full-Stack Engineer to join a cross-functional squad and contribute end-to-end to a specific part of the product. This role involves taking a leading role in building and scaling Nabla’s AI-powered healthcare platform, working across the stack from back-end services to web, desktop, and mobile applications. You will collaborate closely with Product, Design, and Machine Learning teams to deliver high-impact features that directly serve healthcare professionals, with the goal of restoring the human connection at the heart of healthcare by streamlining clinical documentation.

Paris office onsite FullTime
KubernetesPythonTypeScript +5 more

Applied AI, Forward Deployed Machine Learning Engineer - Palo Alto

5mo ago
Mistral AI

Mistral AI

Mistral AI is seeking an Applied AI Engineer to join their team and facilitate the adoption of their products among customers. This role involves collaborating closely with clients from pre-sale to post-implementation, ensuring their solutions meet and exceed expectations. You will manage daily customer relations, acting as a key resource for externalizing research into production settings and driving the successful deployment of Mistral AI products. The position offers the opportunity to work on state-of-the-art Generative AI applications across various industries and contribute to a pioneering company shaping the future of AI.

Palo Alto onsite Full-time
MistralLangChainPython +9 more

Machine Learning Engineer

5mo ago
S

Skild

Skild AI is seeking a Machine Learning Engineer to design and implement cutting-edge reinforcement learning algorithms for robotic applications. This role involves conducting experiments, optimizing models for real-world robotic environments, and collaborating with robotics, research, and engineering teams. The work will directly contribute to the development of intelligent, adaptable robots capable of autonomous learning and complex task performance.

San Mateo, Pittsburgh onsite
PythonC#PyTorch +4 more

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.