Reinforcement Learning Jobs
93 open roles mentioning Reinforcement Learning
Senior Software Engineer - Research Platform, Consumer Devices
OpenAI
We are seeking a Software Engineer to join our Applied Research team within the Consumer Devices group. In this role, you will build tools and services that enable AI research, evaluation, and data generation workflows. Your work will involve transforming ambiguous design questions into working research systems, collaborating closely with researchers, designers, and engineers. The goal is to simplify the creation, execution, and trustworthiness of these workflows, ensuring research artifacts have a clear lifecycle, runs are reproducible and observable, and results provide valuable evidence for product and model-training decisions. You will help maintain reliable and reusable underlying systems.
Research Engineer / Research Scientist -Personal AGI, Proactivity
OpenAI
The Proactivity Research team at OpenAI is focused on making AI models proactive and useful for users. This role involves researching and developing improvements to model personalization and agentic capabilities, working on areas like reinforcement learning, dataset creation, and post-training methods. The team collaborates closely with other research and product teams to build a personalized, collaborative, and proactive assistant. We are looking for individuals with strong ML engineering skills and research experience, particularly with novel and capable models, who are passionate about product-driven research.
Senior AI Product Manager, Healthcare Agents
Scale AI
Scale is seeking an AI Product Manager to lead the Healthcare vertical within the Agents Data & Reinforcement Learning Environments team. This role involves owning the development of realistic RL environments for training and evaluating AI agents in healthcare software and workflows, as well as defining the "data as a product" strategy that supports them. The ideal candidate will possess deep understanding of the Healthcare industry and its workflows, combined with insight into AI research and current agent capabilities. You will translate this expertise into environments and datasets that enable AI agents to perform real healthcare tasks, serving as the domain expert for Scale's key customers and researchers. A strong entrepreneurial and go-to-market mindset is essential for success in this position.
$206k - $257k
Senior AI Product Manager, Finance Agents
Scale AI
Scale is seeking an AI Product Manager to lead the Finance vertical within the Agents Data & Reinforcement Learning Environments team. This role involves owning the development of realistic RL environments for training and evaluating AI agents in financial workflows, as well as defining the "data as a product" strategy that supports them. You will leverage your deep understanding of the Finance industry and AI capabilities to identify valuable financial tasks for AI modeling, determine data sourcing and structuring strategies, and translate domain expertise into a defensible product. The ideal candidate will possess a strong entrepreneurial and go-to-market mindset, coupled with the ability to pair finance industry experience with an understanding of AI research and current agent capabilities in financial workflows.
$205k - $257k
Forward Deployed Robotics Engineer
Black Forest Labs
Black Forest Labs is seeking a Forward Deployed Robotics Engineer to integrate their advanced generative models, including action/VLA models, into the robotics and physical-AI systems of their key customers. This role involves embedding with clients, implementing real-world integrations, and providing direct feedback to the product and research teams. The ideal candidate is passionate about bridging the gap between cutting-edge AI research and practical robotic applications, with a proven ability to deliver functional systems in research, open-source, or production environments. This is a high-ownership position focused on customer success and driving the adoption of foundational AI technologies.
€120k - €180k
Research Intern RL & Post-Training Systems, Turbo (Fall 2026)
Together AI
The Turbo Research team focuses on making post-training and reinforcement learning for large language models efficient, scalable, and reliable. This work intersects RL algorithms, inference systems, and large-scale experimentation, where inference costs significantly impact training efficiency and the practicality of learning algorithms. As a research intern, you will investigate RL and post-training methods whose performance and scalability are closely tied to inference behavior, co-designing algorithms and systems. Projects aim to enable new experimental regimes, including larger models, longer rollouts, and more complex evaluations, by re-evaluating the interaction between inference, scheduling, and training.
Software Engineer, Identity
Scale AI
Scale is seeking a Software Engineer, Identity to join our Platform Engineering team. In this role, you will be instrumental in designing and developing core platforms and software systems, with a specific focus on identity, access management, authorization, and authentication. You will gain broad exposure to the cutting edge of the AI industry as Scale supports enterprises, startups, and governments. This position offers the opportunity to contribute to the foundational elements of products that power advanced LLMs and generative models, playing a crucial role in how humanity interacts with AI.
$216k - $270k
AI research scientist
Writer
AI research at WRITER focuses on building the scientific foundation for ambitious enterprise AI deployments. As a staff AI research scientist, you will drive a high-impact research agenda centered on large language models, agentic reasoning, and system-level capabilities essential for enterprise-scale AI. This role offers a unique opportunity to advance the field while directly contributing to products used by hundreds of thousands daily. You will work on post-training, planning, multi-step reasoning, and agentic workflows, directly shaping the future of enterprise AI performance and scalability. The role provides resources, infrastructure, and cross-functional support to pursue and implement ambitious ideas rapidly.
Research Intern, Model Shaping (Fall 2026)
Together AI
As a Research Intern in the Model Shaping team, you will work on advanced post-training methods, new techniques for efficient neural network training, and robust evaluation of foundation model capabilities. The Model Shaping team at Together AI focuses on tailoring open foundation models for downstream applications, building services for machine learning developers, and developing new methods for efficient model training and evaluation. This role offers the opportunity to contribute to cutting-edge research and potentially influence open-source projects.
Technical Program Manager, Engineering
Scale AI
Scale is at the forefront of the AI revolution, developing data engines and technologies that power the world's leading LLMs. This role focuses on leading critical programs within the Platform and Security Engineering teams, overseeing the design and development of core data storage systems and security initiatives. You will drive company-wide programs, improve processes, and ensure alignment with industry standards, gaining exposure to the cutting edge of AI adoption across various sectors. The work is crucial for making AI models safe, aligned, and useful through human evaluation and reinforcement learning.
$181k - $226k
Senior/Staff Machine Learning Research Engineer, General Agents, Enterprise GenAI
Scale AI
Scale AI is seeking a Senior/Staff Machine Learning Engineer for its General Agents team. This role is crucial in designing, building, and deploying production-ready AI agents to address high-impact enterprise challenges. You will be involved in the entire agent lifecycle, from conceptualization and system design to evaluation, deployment, and ongoing iteration. The position requires bridging cutting-edge agentic techniques with the practical demands of real-world customer environments, focusing on creating scalable, reliable, and generalizable agent systems.
$265k - $331k
Software Engineer, Platform
Scale AI
Scale is at the forefront of the AI revolution, building the Generative AI Data Engine and other products that power the world's most advanced LLMs. The Platform Engineering team is foundational to these efforts, responsible for designing and developing shared platforms, architecting core cloud infrastructure, and redefining software development processes. This role offers exposure to the cutting edge of AI development across various sectors, from startups to governments. You will drive the design and implementation of critical platforms, collaborate with cross-functional teams, and proactively improve engineering practices. This is an opportunity to shape the future of AI infrastructure and contribute to some of the most important work in how humanity interacts with AI.
$216k - $270k
Technical Lead Manager, Physical AI
Scale AI
Scale AI is seeking a Technical Lead Manager for its Physical AI team, focusing on the development of general AI that can reason and act in the physical world. This role bridges cutting-edge Machine Learning research with physical robot deployment, leading a team of Research Engineers while remaining a hands-on technical contributor. The primary focus is on developing and evaluating Large-Scale Foundation Models, such as VLAs and World models, to enable robots and autonomous vehicles to generalize across diverse tasks and environments. The team leverages Scale's extensive data infrastructure to help build Foundation Models for Physical AI, aiming to redefine the future of automation.
$249k - $311k
Staff Software Engineer, Data Platform
Scale AI
Scale is at the forefront of the AI revolution, developing data engines and technologies that power the world's leading LLMs and generative models. This role is on the Platform Engineering team, responsible for the foundational data infrastructure that supports these cutting-edge AI products. You will lead the design and development of core data storage, streaming, caching, and indexing platforms, gaining exposure to the rapidly evolving AI landscape across various industries. The work involves driving architecture, implementation, and reliability of these critical systems, collaborating with stakeholders, and mentoring junior engineers.
$252k - $315k
Forward Deployed Engineer, GenAI
Scale AI
Scale AI is seeking a Forward Deployed Engineer, GenAI to join their Data Engine team. This role is at the forefront of providing critical data infrastructure that powers advanced AI models, directly influencing how humanity interacts with AI. You will work with the world's leading AI companies and government agencies to solve their most complex AI data-related problems, contributing to the advancement of AI by delivering critical data solutions for leading AI innovators and government agencies. You will interact daily with technical customers, understand their unique challenges, and translate them into impactful solutions, while also designing, building, and deploying features across the entire stack. This position offers a unique opportunity to lead critical projects, shape engineering culture, and accelerate career growth in the rapidly evolving field of Generative AI.
$179k - $224k
Senior Software Engineer, GenAI
Scale AI
Scale AI is seeking a Senior Software Engineer to join our Generative AI Data Engine team. This role is crucial in accelerating the development of AI applications by powering the world's most advanced LLMs and generative models. You will work on high-impact datasets, optimize contributor onboarding, and ensure data integrity through advanced trust, safety, and security measures. This position operates at the intersection of ML, operations, and analytics to deliver high-quality data at scale, contributing to the future of human-AI interaction.
$216k - $270k
Software Engineer, RL Training Infra
OpenAI
This role focuses on keeping our frontier RL training runs fast, reliable, and unblocked. You will work across engineering and infrastructure problems as they emerge, from scaling and orchestration issues to inference bottlenecks, numerical problems, and hardware failures, as well as supporting large horizontal integrations in the big run, like multi-agent capabilities or memory. This is a role for a strong generalist who quickly learns anything needed for the task, has high attention to detail, debugs deeply, and is motivated by fixing the highest-impact problem in front of the team.
Machine Learning Research Scientist, Post-Training
Scale AI
Scale works with leading AI labs to accelerate progress in GenAI research, focusing on optimizing data curation and evaluation to enhance LLM capabilities in text and multimodal modalities. This role involves developing novel methods to improve the alignment and generalization of large-scale generative models, collaborating with researchers and engineers on best practices in data-driven AI development, and providing technical and strategic input to foundation model labs for the next generation of AI models.
$252k - $315k
GenAI Strategic Projects Lead, Public Sector
Scale AI
Scale AI is seeking a Strategic Projects Lead for its Public Sector team to own high-impact projects focused on Generative AI and Large Language Models. This role involves working across operations, engineering, and customer engagement to produce high-quality training and test data for LLMs, particularly for Public Sector customers. You will be instrumental in building Generative AI data-labeling pipelines, creating operational processes for an expert data workforce, and developing novel technology-driven approaches to enhance data quality. This is a unique opportunity to contribute at the intersection of AI and national security, partnering with internal ML experts and external stakeholders to ensure data supports mission-critical AI applications.
$170k - $212k
Software Engineer, Kernel Performance & AI Tooling
OpenAI
OpenAI's Hardware organization is developing AI-native silicon and system-level solutions for advanced AI workloads. We are seeking a systems-minded engineer to advance our kernel development, performance engineering, and hardware-software co-design capabilities, with a focus on AI-assisted workflows and tooling. This role sits at the intersection of kernel optimization, developer tooling, observability, and research infrastructure, aiming to improve how production kernels are built and optimized, and how future hardware-software systems are designed and evaluated. The ideal candidate is excited by low-level performance work and sees AI and automation as powerful tools for accelerating engineering velocity, helping to define the future of kernel engineering in the era of AI-assisted development.