PyTorch Jobs

109 open roles mentioning PyTorch

Research, Post-Training Data

2mo ago
thinkingmachines

thinkingmachines

Thinking Machines Lab is seeking researchers to bridge the gap between raw AI intelligence and useful, safe, and collaborative systems. This role focuses on post-training data research, combining human insight and machine learning techniques to capture and steer model behavior based on human preferences. You will be responsible for translating research ideas into actionable data through labeling and collection campaigns, understanding data quality science, and developing metrics to measure the impact of data and training interventions. The position also involves exploring new paradigms for human-AI interaction and scalable oversight, blending research, data operations, and technical implementation to advance human-centered AI systems. This role requires both fundamental research and practical engineering, making it ideal for individuals who enjoy deep theoretical exploration and hands-on experimentation.

$350k - $475k

San Francisco onsite
OpenAIMistralPython +5 more

Research Engineer, Infrastructure, Numerics

2mo ago
thinkingmachines

thinkingmachines

Thinking Machines Lab is seeking an infrastructure research engineer to design and build core systems for efficient large-scale model training, with a specific focus on numerics. This role involves enhancing the numerical foundations of their distributed training stack, optimizing precision formats, kernel optimizations, and communication frameworks to ensure stable, scalable, and fast training of trillion-parameter models. The ideal candidate will bridge research and systems engineering, possessing a strong understanding of both optimization mathematics and distributed compute realities.

$350k - $475k

San Francisco onsite
OpenAIMistralPyTorch +3 more

Research Engineer, Infrastructure, Kernels

2mo ago
thinkingmachines

thinkingmachines

Thinking Machines Lab is seeking an infrastructure research engineer to design, optimize, and maintain the compute foundations for large-scale language model training. This role involves developing high-performance ML kernels, enabling efficient low-precision arithmetic, and improving the distributed compute stack. You will work closely with researchers and systems architects, bridging algorithmic design with hardware efficiency, prototyping new kernel implementations, and defining numerical and parallelism strategies for scaling AI systems.

$350k - $475k

San Francisco onsite
OpenAIMistralPyTorch +2 more

Infrastructure Engineer, Security

2mo ago
thinkingmachines

thinkingmachines

Thinking Machines Lab is seeking an infrastructure engineer to lead and enhance the security infrastructure for their foundation models. This role involves working across compute, storage, networking, and data platforms to ensure systems are secure, reliable, and scalable. The engineer will define security controls, architecture, and tooling, integrating security by default into the platform. Collaboration with research and product teams will be key to enabling rapid progress while maintaining robust protection for models, data, and environments.

$200k - $475k

San Francisco onsite
OpenAIMistralKubernetes +4 more

Research Internship (Fall, Winter 2026)

3mo ago
Cohere

Cohere

Cohere is seeking a Research Intern to collaborate with researchers and tools on designing and implementing novel research ideas and shipping state-of-the-art models to production. Interns will have the opportunity to work on various teams covering base model training, retrieval augmented generation, data and evaluation, safety, and finetuning, or any research area relating to LLMs. This role offers a chance to broaden research connections while gaining deep experience in a growing AI startup.

Canada remote FullTime
CoherePythonC# +5 more

Applied Scientist / Research Engineer, AI4Engineering - EMEA

3mo ago
Mistral AI

Mistral AI

Mistral AI is seeking an Applied Scientist with deep expertise in engineering sciences to work at the forefront of AI-accelerated simulation. This role involves collaborating with industrial customers and internal research teams to build and deploy AI Physics Models, complementing our existing Large Language Models (LLMs). You will be involved in the entire process, from curating high-fidelity simulation datasets and training/evaluating models to delivering production-grade AI solutions directly to engineering teams. The target domains include computational fluid dynamics, structural mechanics, semiconductor design, multi-physics modeling, and digital twins. Working cross-functionally, you will ensure our models meet stringent engineering standards beyond just benchmark metrics.

Paris hybrid Full-time
MistralPythonRAG +9 more

Applied AI, Forward Deployed Machine Learning Engineer, Critical and Sovereign Institutions, EMEA

3mo ago
Mistral AI

Mistral AI

Mistral AI is seeking an Applied AI Engineer to join its specialized unit focused on delivering high-impact, secure AI solutions for critical and sovereign institutions. This team works closely with clients in highly regulated environments to design, deploy, and maintain AI systems that meet stringent standards for reliability, security, and operational excellence. The role involves technical design, implementation, and deployment of AI solutions tailored to the unique needs of critical infrastructure and sovereign institutions, contributing directly to projects with significant societal and operational impact.

Paris onsite Full-time
MistralLangChainAWS +11 more

Clinician Scientist

3mo ago
Abridge

Abridge

Abridge is seeking Clinician Scientists to advance the development of its AI-powered clinical tools. This role requires a blend of deep clinical expertise and a background in AI, focusing on shaping AI-driven tools to ensure accuracy and quality for clinicians and patients. You will collaborate closely with engineers, researchers, product managers, and fellow clinicians to refine AI models, validate outputs, and create new functionalities that enhance documentation workflows.

SF Office hybrid FullTime
PythonPrompt EngineeringPyTorch +3 more

Applied AI, Forward Deployed Machine Learning Engineer - Palo Alto

3mo ago
Mistral AI

Mistral AI

Mistral AI is seeking an Applied AI Engineer to join their team and facilitate the adoption of their products among customers. This role involves collaborating closely with clients from pre-sale to post-implementation, ensuring their solutions meet and exceed expectations. You will manage daily customer relations, acting as a key resource for externalizing research into production settings and driving the successful deployment of Mistral AI products. The position offers the opportunity to work on state-of-the-art Generative AI applications across various industries and contribute to a pioneering company shaping the future of AI.

Palo Alto onsite Full-time
MistralLangChainPython +9 more

Software Engineer, Security

3mo ago
thinkingmachines

thinkingmachines

Thinking Machines Lab is seeking a Software Engineer focused on security to ensure their AI products are secure by default while enabling rapid product iteration. This role involves embedding with product and research teams to integrate security into the design and development process, as well as building tools and automation to maintain system safety at scale. The company is dedicated to advancing collaborative general intelligence and empowering users with AI tools tailored to their needs.

$350k - $475k

San Francisco onsite
OpenAIMistralPython +2 more

Software Engineer, Systems Generalist

3mo ago
thinkingmachines

thinkingmachines

Thinking Machines Lab is seeking generalist infrastructure and systems engineers to build the core systems powering their foundation models and support internal research and product development teams. This high-impact role involves architecting and scaling critical infrastructure across the full technical stack, solving complex distributed systems problems, and building robust, scalable platforms. You will work directly with researchers to accelerate experiments, improve infrastructure efficiency, and enable key insights across models, products, and data assets.

$350k - $475k

San Francisco onsite
OpenAIMistralKubernetes +5 more

Applied AI, Technical Lead, Forward Deployed AI Engineer - Montreal

4mo ago
Mistral AI

Mistral AI

Mistral AI is seeking a Technical Lead, Applied AI to drive the technical strategy, execution, and delivery of complex AI solutions for enterprise customers. This role involves leading project teams of Applied AI Engineers to ensure successful deployment of Mistral AI products and the development of high-impact, scalable AI use cases. You will serve as the primary technical point of contact for strategic customers, guiding them through the entire lifecycle from pre-sales to post-implementation, and collaborating with research, product, and engineering teams to shape future offerings. The position bridges cutting-edge AI research with real-world enterprise applications, ensuring solutions are robust, scalable, and aligned with customer needs and Mistral's vision.

Montreal remote Full-time
MistralLangChainAWS +10 more

Senior Member of Technical Staff, Safety and Security for Agents

4mo ago
Cohere

Cohere

Cohere is seeking a Senior Member of Technical Staff to join the Safety and Security for Agents team. In this role, you will significantly contribute to the development of safer, fairer, more trustworthy, and more secure Large Language Models (LLMs). Your work will focus on data generation, post-training algorithms, and evaluation methods to ensure the safety of next-generation models that interact with external resources and take actions. You will collaborate closely with machine learning teams, data annotation teams, and product and policy teams, requiring a blend of machine learning expertise, ethical AI principles, experimental design, and data management skills. This position offers significant autonomy and decision-making power within a small team, with the opportunity to shape the future of LLMs for societal benefit.

London hybrid FullTime
CoherePythonPyTorch +3 more

Applied AI, Forward Deployed Machine Learning Engineer - Montreal

4mo ago
Mistral AI

Mistral AI

Mistral AI is seeking an Applied AI Engineer to join their team and facilitate customer adoption of their AI products. This role involves collaborating closely with customers to address complex technical challenges, from pre-sale to post-implementation, ensuring solutions meet and exceed client expectations. You will manage daily customer relations, act as a key resource for externalizing research into production, and work on state-of-the-art Generative AI applications across various industries. This position offers the opportunity to contribute to a pioneering company shaping the future of AI and make a meaningful impact.

Montreal onsite Full-time
MistralLangChainPython +9 more

Generative AI Inference Engineer

4mo ago
Stability AI

Stability AI

We are seeking passionate Machine Learning Engineers to join our Inference team, focusing on the creative applications of generative AI models. The ideal candidate will have substantial experience developing and running inference for multi-modal models. A deep understanding of diffusion model architectures and familiarity with workflow tools like ComfyUI are a big plus. You will be expected to leverage and push the boundaries of state-of-the-art inference optimization techniques for multi-modal generative models. This role offers the opportunity to work alongside top researchers and engineers, utilizing cutting-edge high-performance computing resources to make a significant impact in the rapidly evolving field of generative AI.

United States remote
AWSAzureDocker +10 more

Research Scientist – Controlled 3D Generation

4mo ago
Stability AI

Stability AI

We are seeking a Research Scientist passionate about 3D generation, flow matching, and diffusion models. You will help advance the frontier of controllable 3D content creation by building models that generate consistent, editable, and physically grounded 3D assets and scenes. This role involves conducting cutting-edge research, designing and implementing scalable training pipelines, and developing techniques for conditioning and control. You will analyze model behavior, collaborate with cross-disciplinary teams to translate research into production-ready systems, and publish results at top-tier venues.

Remote remote
PyTorchJAXCUDA +7 more

Research Engineer

5mo ago
Cohere

Cohere

Cohere Labs is seeking Research Engineers to join their dedicated research arm, focused on pushing machine learning forward through open, collaborative research and hands-on experimentation. This is a highly practical role where you will work closely with scientists and engineers to implement new methods, run large-scale experiments, and help shape the infrastructure supporting our research programs. You will be responsible for building experiments, debugging models, scaling training pipelines, and turning research ideas into working systems. We value curiosity, strong fundamentals, and a willingness to learn quickly in a fast-moving research environment, with a focus on practical impact.

Toronto remote FullTime
CohereFine-TuningPyTorch +4 more

Multimodal Generative AI Researcher

6mo ago
Stability AI

Stability AI

We are seeking a Research Scientist with deep expertise in training and fine-tuning large Vision-Language and Language Models (VLMs / LLMs) for downstream multimodal tasks. You will be instrumental in advancing models that reason across vision, language, and 3D, translating research breakthroughs into scalable engineering solutions. This role involves designing and fine-tuning large-scale VLMs/LLMs and hybrid architectures for complex tasks like visual reasoning, retrieval, 3D understanding, and embodied interaction.

Remote remote
PyTorchNLPComputer Vision +7 more

Research Engineer, Machine Learning

6mo ago
Mistral AI

Mistral AI

Mistral AI is democratizing AI through high-performance, optimized, open-source models and solutions. We are a dynamic, collaborative team passionate about AI's potential to transform society, with a diverse workforce driving innovation. As a Research Engineer – ML track, you will build and optimize large-scale learning systems powering our open-weight models. You will work hand-in-hand with Research Scientists, either enhancing the shared training framework and data pipelines or embedding within a research squad to turn fresh ideas into scalable code.

Palo Alto remote Full-time
MistralPythonPyTorch +9 more

Applied AI, Evaluation Engineer

6mo ago
Mistral AI

Mistral AI

Mistral AI is seeking an Evaluation Engineer to join our customer-facing Applied AI team. This role is crucial for ensuring our AI solutions are production-ready by designing methodologies, building infrastructure, and defining evaluation standards across various industries and use cases. You will bridge the gap between research and customer needs, creating evaluation frameworks that measure LLM performance in real-world scenarios, moving beyond standard benchmarks to address domain-specific risks and requirements. This position offers a unique opportunity to impact the deployment of cutting-edge AI by directly contributing to its measurable success for enterprise clients.

Paris remote Full-time
OpenAIMistralPython +6 more

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.