Reinforcement Learning Jobs

93 open roles mentioning Reinforcement Learning

Senior Software Engineer - Research Platform, Consumer Devices

2mo ago
OpenAI

OpenAI

We are seeking a Software Engineer to join our Applied Research team within the Consumer Devices group. In this role, you will build tools and services that enable AI research, evaluation, and data generation workflows. Your work will involve transforming ambiguous design questions into working research systems, collaborating closely with researchers, designers, and engineers. The goal is to simplify the creation, execution, and trustworthiness of these workflows, ensuring research artifacts have a clear lifecycle, runs are reproducible and observable, and results provide valuable evidence for product and model-training decisions. You will help maintain reliable and reusable underlying systems.

San Francisco hybrid FullTime
OpenAIFine-TuningReinforcement Learning

Research Engineer / Research Scientist -Personal AGI, Proactivity

2mo ago
OpenAI

OpenAI

The Proactivity Research team at OpenAI is focused on making AI models proactive and useful for users. This role involves researching and developing improvements to model personalization and agentic capabilities, working on areas like reinforcement learning, dataset creation, and post-training methods. The team collaborates closely with other research and product teams to build a personalized, collaborative, and proactive assistant. We are looking for individuals with strong ML engineering skills and research experience, particularly with novel and capable models, who are passionate about product-driven research.

San Francisco hybrid FullTime
OpenAIReinforcement Learning

Senior AI Product Manager, Healthcare Agents

2mo ago
Scale AI

Scale AI

Scale is seeking an AI Product Manager to lead the Healthcare vertical within the Agents Data & Reinforcement Learning Environments team. This role involves owning the development of realistic RL environments for training and evaluating AI agents in healthcare software and workflows, as well as defining the "data as a product" strategy that supports them. The ideal candidate will possess deep understanding of the Healthcare industry and its workflows, combined with insight into AI research and current agent capabilities. You will translate this expertise into environments and datasets that enable AI agents to perform real healthcare tasks, serving as the domain expert for Scale's key customers and researchers. A strong entrepreneurial and go-to-market mindset is essential for success in this position.

$206k - $257k

San Francisco, CA; Seattle, WA; New York, NY remote
GoReinforcement LearningAI +8 more

Senior AI Product Manager, Finance Agents

2mo ago
Scale AI

Scale AI

Scale is seeking an AI Product Manager to lead the Finance vertical within the Agents Data & Reinforcement Learning Environments team. This role involves owning the development of realistic RL environments for training and evaluating AI agents in financial workflows, as well as defining the "data as a product" strategy that supports them. You will leverage your deep understanding of the Finance industry and AI capabilities to identify valuable financial tasks for AI modeling, determine data sourcing and structuring strategies, and translate domain expertise into a defensible product. The ideal candidate will possess a strong entrepreneurial and go-to-market mindset, coupled with the ability to pair finance industry experience with an understanding of AI research and current agent capabilities in financial workflows.

$205k - $257k

San Francisco, CA; Seattle, WA; New York, NY onsite
GoReinforcement LearningAI +5 more

Forward Deployed Robotics Engineer

2mo ago
B

Black Forest Labs

Black Forest Labs is seeking a Forward Deployed Robotics Engineer to integrate their advanced generative models, including action/VLA models, into the robotics and physical-AI systems of their key customers. This role involves embedding with clients, implementing real-world integrations, and providing direct feedback to the product and research teams. The ideal candidate is passionate about bridging the gap between cutting-edge AI research and practical robotic applications, with a proven ability to deliver functional systems in research, open-source, or production environments. This is a high-ownership position focused on customer success and driving the adoption of foundational AI technologies.

€120k - €180k

Freiburg (Germany) onsite FullTime
Fine-TuningReinforcement Learning

Research Intern RL & Post-Training Systems, Turbo (Fall 2026)

2mo ago
Together AI

Together AI

The Turbo Research team focuses on making post-training and reinforcement learning for large language models efficient, scalable, and reliable. This work intersects RL algorithms, inference systems, and large-scale experimentation, where inference costs significantly impact training efficiency and the practicality of learning algorithms. As a research intern, you will investigate RL and post-training methods whose performance and scalability are closely tied to inference behavior, co-designing algorithms and systems. Projects aim to enable new experimental regimes, including larger models, longer rollouts, and more complex evaluations, by re-evaluating the interaction between inference, scheduling, and training.

San Francisco remote
PythonC#NLP +7 more

Software Engineer, Identity

2mo ago
Scale AI

Scale AI

Scale is seeking a Software Engineer, Identity to join our Platform Engineering team. In this role, you will be instrumental in designing and developing core platforms and software systems, with a specific focus on identity, access management, authorization, and authentication. You will gain broad exposure to the cutting edge of the AI industry as Scale supports enterprises, startups, and governments. This position offers the opportunity to contribute to the foundational elements of products that power advanced LLMs and generative models, playing a crucial role in how humanity interacts with AI.

$216k - $270k

San Francisco, CA; New York, NY remote
AWSAzurePython +14 more

AI research scientist

2mo ago
W

Writer

AI research at WRITER focuses on building the scientific foundation for ambitious enterprise AI deployments. As a staff AI research scientist, you will drive a high-impact research agenda centered on large language models, agentic reasoning, and system-level capabilities essential for enterprise-scale AI. This role offers a unique opportunity to advance the field while directly contributing to products used by hundreds of thousands daily. You will work on post-training, planning, multi-step reasoning, and agentic workflows, directly shaping the future of enterprise AI performance and scalability. The role provides resources, infrastructure, and cross-functional support to pursue and implement ambitious ideas rapidly.

San Francisco, CA hybrid FullTime
PythonAI AgentsFine-Tuning +5 more

Research Intern, Model Shaping (Fall 2026)

3mo ago
Together AI

Together AI

As a Research Intern in the Model Shaping team, you will work on advanced post-training methods, new techniques for efficient neural network training, and robust evaluation of foundation model capabilities. The Model Shaping team at Together AI focuses on tailoring open foundation models for downstream applications, building services for machine learning developers, and developing new methods for efficient model training and evaluation. This role offers the opportunity to contribute to cutting-edge research and potentially influence open-source projects.

San Francisco hybrid
PyTorchNLPReinforcement Learning +6 more

Technical Program Manager, Engineering

3mo ago
Scale AI

Scale AI

Scale is at the forefront of the AI revolution, developing data engines and technologies that power the world's leading LLMs. This role focuses on leading critical programs within the Platform and Security Engineering teams, overseeing the design and development of core data storage systems and security initiatives. You will drive company-wide programs, improve processes, and ensure alignment with industry standards, gaining exposure to the cutting edge of AI adoption across various sectors. The work is crucial for making AI models safe, aligned, and useful through human evaluation and reinforcement learning.

$181k - $226k

San Francisco, CA; New York, NY onsite
AWSFine-TuningSQL +2 more

Senior/Staff Machine Learning Research Engineer, General Agents, Enterprise GenAI

3mo ago
Scale AI

Scale AI

Scale AI is seeking a Senior/Staff Machine Learning Engineer for its General Agents team. This role is crucial in designing, building, and deploying production-ready AI agents to address high-impact enterprise challenges. You will be involved in the entire agent lifecycle, from conceptualization and system design to evaluation, deployment, and ongoing iteration. The position requires bridging cutting-edge agentic techniques with the practical demands of real-world customer environments, focusing on creating scalable, reliable, and generalizable agent systems.

$265k - $331k

San Francisco, CA; New York, NY onsite
OpenAIPythonAI Agents +10 more

Software Engineer, Platform

3mo ago
Scale AI

Scale AI

Scale is at the forefront of the AI revolution, building the Generative AI Data Engine and other products that power the world's most advanced LLMs. The Platform Engineering team is foundational to these efforts, responsible for designing and developing shared platforms, architecting core cloud infrastructure, and redefining software development processes. This role offers exposure to the cutting edge of AI development across various sectors, from startups to governments. You will drive the design and implementation of critical platforms, collaborate with cross-functional teams, and proactively improve engineering practices. This is an opportunity to shape the future of AI infrastructure and contribute to some of the most important work in how humanity interacts with AI.

$216k - $270k

San Francisco, CA; New York, NY remote
AWSDockerKubernetes +11 more

Technical Lead Manager, Physical AI

3mo ago
Scale AI

Scale AI

Scale AI is seeking a Technical Lead Manager for its Physical AI team, focusing on the development of general AI that can reason and act in the physical world. This role bridges cutting-edge Machine Learning research with physical robot deployment, leading a team of Research Engineers while remaining a hands-on technical contributor. The primary focus is on developing and evaluating Large-Scale Foundation Models, such as VLAs and World models, to enable robots and autonomous vehicles to generalize across diverse tasks and environments. The team leverages Scale's extensive data infrastructure to help build Foundation Models for Physical AI, aiming to redefine the future of automation.

$249k - $311k

San Francisco, CA remote
Fine-TuningPyTorchReinforcement Learning +9 more

Staff Software Engineer, Data Platform

3mo ago
Scale AI

Scale AI

Scale is at the forefront of the AI revolution, developing data engines and technologies that power the world's leading LLMs and generative models. This role is on the Platform Engineering team, responsible for the foundational data infrastructure that supports these cutting-edge AI products. You will lead the design and development of core data storage, streaming, caching, and indexing platforms, gaining exposure to the rapidly evolving AI landscape across various industries. The work involves driving architecture, implementation, and reliability of these critical systems, collaborating with stakeholders, and mentoring junior engineers.

$252k - $315k

San Francisco, CA; New York, NY remote
KubernetesPythonFine-Tuning +10 more

Forward Deployed Engineer, GenAI

3mo ago
Scale AI

Scale AI

Scale AI is seeking a Forward Deployed Engineer, GenAI to join their Data Engine team. This role is at the forefront of providing critical data infrastructure that powers advanced AI models, directly influencing how humanity interacts with AI. You will work with the world's leading AI companies and government agencies to solve their most complex AI data-related problems, contributing to the advancement of AI by delivering critical data solutions for leading AI innovators and government agencies. You will interact daily with technical customers, understand their unique challenges, and translate them into impactful solutions, while also designing, building, and deploying features across the entire stack. This position offers a unique opportunity to lead critical projects, shape engineering culture, and accelerate career growth in the rapidly evolving field of Generative AI.

$179k - $224k

San Francisco, CA; New York, NY hybrid
Reinforcement LearningRLHFAI +8 more

Senior Software Engineer, GenAI

3mo ago
Scale AI

Scale AI

Scale AI is seeking a Senior Software Engineer to join our Generative AI Data Engine team. This role is crucial in accelerating the development of AI applications by powering the world's most advanced LLMs and generative models. You will work on high-impact datasets, optimize contributor onboarding, and ensure data integrity through advanced trust, safety, and security measures. This position operates at the intersection of ML, operations, and analytics to deliver high-quality data at scale, contributing to the future of human-AI interaction.

$216k - $270k

San Francisco, CA; New York, NY hybrid
PythonTypeScriptReinforcement Learning +6 more

Software Engineer, RL Training Infra

3mo ago
OpenAI

OpenAI

This role focuses on keeping our frontier RL training runs fast, reliable, and unblocked. You will work across engineering and infrastructure problems as they emerge, from scaling and orchestration issues to inference bottlenecks, numerical problems, and hardware failures, as well as supporting large horizontal integrations in the big run, like multi-agent capabilities or memory. This is a role for a strong generalist who quickly learns anything needed for the task, has high attention to detail, debugs deeply, and is motivated by fixing the highest-impact problem in front of the team.

San Francisco hybrid FullTime
OpenAIGoReinforcement Learning

Machine Learning Research Scientist, Post-Training

4mo ago
Scale AI

Scale AI

Scale works with leading AI labs to accelerate progress in GenAI research, focusing on optimizing data curation and evaluation to enhance LLM capabilities in text and multimodal modalities. This role involves developing novel methods to improve the alignment and generalization of large-scale generative models, collaborating with researchers and engineers on best practices in data-driven AI development, and providing technical and strategic input to foundation model labs for the next generation of AI models.

$252k - $315k

San Francisco, CA; Seattle, WA; New York, NY onsite
LLMFine-TuningReinforcement Learning +7 more

GenAI Strategic Projects Lead, Public Sector

4mo ago
Scale AI

Scale AI

Scale AI is seeking a Strategic Projects Lead for its Public Sector team to own high-impact projects focused on Generative AI and Large Language Models. This role involves working across operations, engineering, and customer engagement to produce high-quality training and test data for LLMs, particularly for Public Sector customers. You will be instrumental in building Generative AI data-labeling pipelines, creating operational processes for an expert data workforce, and developing novel technology-driven approaches to enhance data quality. This is a unique opportunity to contribute at the intersection of AI and national security, partnering with internal ML experts and external stakeholders to ensure data supports mission-critical AI applications.

$170k - $212k

Washington, DC onsite
Prompt EngineeringFine-TuningReinforcement Learning +9 more

Software Engineer, Kernel Performance & AI Tooling

4mo ago
OpenAI

OpenAI

OpenAI's Hardware organization is developing AI-native silicon and system-level solutions for advanced AI workloads. We are seeking a systems-minded engineer to advance our kernel development, performance engineering, and hardware-software co-design capabilities, with a focus on AI-assisted workflows and tooling. This role sits at the intersection of kernel optimization, developer tooling, observability, and research infrastructure, aiming to improve how production kernels are built and optimized, and how future hardware-software systems are designed and evaluated. The ideal candidate is excited by low-level performance work and sees AI and automation as powerful tools for accelerating engineering velocity, helping to define the future of kernel engineering in the era of AI-assisted development.

San Francisco hybrid FullTime
OpenAIReinforcement Learning

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.