Reinforcement Learning Jobs

93 open roles mentioning Reinforcement Learning

Machine Learning Engineer

5mo ago
S

Skild

Skild AI is seeking a Machine Learning Engineer to design and implement cutting-edge reinforcement learning algorithms for robotic applications. This role involves conducting experiments, optimizing models for real-world robotic environments, and collaborating with robotics, research, and engineering teams. The work will directly contribute to the development of intelligent, adaptable robots capable of autonomous learning and complex task performance.

San Mateo, Pittsburgh onsite
PythonC#PyTorch +4 more

Machine Learning Engineer, Reinforcement Learning

5mo ago
S

Skild

Skild AI is building the world's first general-purpose robotic intelligence that is robust and adapts to unseen scenarios without failing. We are seeking a Machine Learning Engineer to design and implement cutting-edge reinforcement learning algorithms for robotic applications. This role involves conducting experiments, optimizing models for real-world robotic environments, and collaborating closely with our robotics, research, and engineering teams. Your work will directly contribute to the development of intelligent, adaptable robots capable of autonomous learning and complex task performance.

Pittsburgh, San Mateo onsite
PythonC#PyTorch +4 more

Member of Technical Staff - Mid-Training Infra

5mo ago
Reflection ai

Reflection ai

Reflection is a research lab dedicated to making intelligence open and accessible. We build open models that empower individuals to control their intelligence and shape the future of AI. As a Member of Technical Staff focused on Mid-Training Infrastructure, you will be instrumental in designing, building, and operating large-scale GPU infrastructure crucial for high-throughput model inference and mid-training workloads. This role involves developing systems that support synthetic data generation and reinforcement learning pipelines at scale, as well as building high-performance inference platforms capable of serving and evaluating models across thousands of GPUs.

San Francisco, CA onsite FullTime
Reinforcement LearningModel Serving

Senior Staff Software Engineer, AI Model Lifecycle

5mo ago
c

crusoe

Crusoe is seeking a Senior Staff Software Engineer for the AI Model Lifecycle team to build a comprehensive managed platform for the entire application development lifecycle, with a specific focus on leveraging Machine Learning models, including Large Language Models (LLMs). This role is crucial in accelerating the abundance of energy and intelligence by powering the world's most ambitious AI workloads. You will join a team building the future of AI infrastructure, solving the bottleneck of power for AI compute with an energy-first approach. We are looking for problem-solving, opportunity-finding teammates with a sense of urgency who thrive on building the path forward.

$238k - $318k

San Francisco, CA - US onsite FullTime
PythonGoFine-Tuning +2 more

Senior Software Engineer, AI Model Lifecycle

5mo ago
c

crusoe

Crusoe is seeking a Senior Software Engineer for their AI Model Lifecycle team to build a comprehensive managed platform for the entire application development lifecycle, with a specific focus on leveraging Machine Learning models, including Large Language Models (LLMs). This role is crucial in accelerating the abundance of energy and intelligence by powering ambitious AI workloads. The ideal candidate will be a problem-solving, opportunity-finding teammate with a sense of urgency, who believes in the scale of Crusoe's ambition and thrives on a path not fully paved.

$172k - $231k

San Francisco, CA - US onsite FullTime
PythonGoFine-Tuning +2 more

Research Scientist, Multimodal Alignment, Safety, and Fairness

6mo ago
Google DeepMind

Google DeepMind

Google DeepMind's Frontier AI unit is seeking experienced Research Scientists to join a multimodal safety research effort. This role focuses on interdisciplinary sociotechnical modeling and requires a passion for understanding AI-society interactions, a strong awareness of AI alignment and safety, and a drive to develop novel ideas, methods, interfaces, and tools. You will contribute to advancing the state of the art in AI research and Google DeepMind's mission towards Artificial General Intelligence (AGI), with a focus on leading new breakthrough research directions in areas like AI behavior exploration, assessment, and steering, particularly for subjective and creative tasks. The work involves tackling fundamental research questions to improve alignment objectives, assess adherence to desired behaviors, and enable AI agents to monitor real-world social context and evolve system behaviors over long time-horizons. You will develop new paradigms for human+AI rating that are adaptive and context-aware, driving breakthroughs within Google DeepMind, Google products, and the broader AI alignment community.

Kirkland, Washington, US; Mountain View, California, US; New York City, New York, US onsite
PythonAI AgentsFine-Tuning +5 more

Research Scientist, Gemini Safety

6mo ago
Google DeepMind

Google DeepMind

The Gemini Safety team at Google DeepMind is responsible for the safety and fairness of the latest Gemini models. As a Research Scientist/Engineer, you will apply and develop cutting-edge data and algorithmic solutions to advance these user-facing models. This is a fast-paced, highly collaborative role within a supportive team dedicated to pushing the boundaries of AI for public benefit and scientific discovery, with safety and ethics as the highest priorities.

Mountain View, California, US onsite
Fine-TuningReinforcement LearningJAX +1 more

Software Engineer - Training Product

7mo ago
B

Baseten

Baseten is seeking a customer-obsessed software engineer to join their team and contribute to the development of mission-critical AI inference platforms. In this role, you will own features from conception to launch, working across the entire technology stack from API and UI down to the infrastructure layer. You will have the opportunity to fine-tune models, gain a deep understanding of user workflows, and collaborate closely with research engineers to build cutting-edge experiences that accelerate model development and address real-world pain points. If you are excited about diving deep into AI model training and building impactful products, this is the role for you.

San Francisco hybrid FullTime
KubernetesFine-TuningPyTorch +3 more

Member of Technical Staff - Safety

8mo ago
Reflection ai

Reflection ai

Reflection is a research lab dedicated to making intelligence open and accessible. We build open models that empower users to control their intelligence and shape the future of AI. As a Member of Technical Staff - Safety, you will be instrumental in ensuring the safety and reliability of our AI models. This role involves owning the red-teaming and adversarial evaluation pipeline, translating safety findings into concrete guardrails, and validating that every release meets our risk thresholds before deployment. You will develop scalable, automated safety benchmarks and research state-of-the-art jailbreaking techniques and defenses to proactively address potential vulnerabilities.

San Francisco, CA onsite FullTime
Reinforcement LearningRLHF

Product Manager, Forge

9mo ago
Mistral AI

Mistral AI

Mistral AI is seeking a talented and experienced Product Manager to define and execute the strategy for Forge, a product that empowers customers to build, fine-tune, and deploy custom AI models at scale. Forge transforms cutting-edge research into enterprise-ready capabilities by supporting model fine-tuning, reinforcement learning, and post-training workflows. This role operates at the intersection of research and product, enabling customers to train specialized models for real-world business value. You will collaborate closely with applied AI scientists and research engineers to translate frontier techniques into scalable and reliable solutions, shaping a 0-1 product with significant business impact and defining the future of how organizations train and deploy AI models.

Paris hybrid Full-time
MistralKubernetesGo +9 more

Research Engineer

10mo ago
G

Gamma

As a Research Engineer at Gamma, you will build models for visual communication, focusing on teaching them to reason about spatial composition, hierarchy, and visual language. This role is at the intersection of research rigor and product impact, offering the opportunity to develop evaluation frameworks and training data for a field with limited existing resources. You will fine-tune vision-language models to ensure Gamma's 100M+ users receive exceptional design quality with every generation. Success in this role requires a combination of deep expertise in VLMs and multimodal modeling, a research-oriented mindset, comfort with ambiguity, and a keen eye for visual and design quality.

$180k - $340k

San Francisco onsite FullTime
Fine-TuningReinforcement Learning

Forward Deployed Engineer, Lead - LLM Post-training

10mo ago
Reflection ai

Reflection ai

Reflection is a research lab dedicated to making intelligence open and accessible. We are seeking an exceptional technical leader to build and scale our post-training and evaluation capabilities within the Applied AI team. This role involves taking our open-weight models and adapting them for specific customer domains, tasks, and constraints. You will own the end-to-end technical strategy for model customization, from synthetic data generation and reward modeling through training and production deployment, working directly with customers and research teams.

New York, NY onsite FullTime
Reinforcement LearningDistributed Training

Applied Machine Learning Engineer

1y ago
f

fireworks ai

As an Applied Machine Learning Engineer, you will serve as a vital bridge between cutting-edge AI research and practical, real-world applications. Your work will focus on developing, fine-tuning, and operationalizing machine learning models that drive business value and enhance user experiences. This is a hands-on engineering role that combines deep technical expertise with a strong customer focus to deliver scalable AI solutions.

San Mateo hybrid FullTime
PythonFine-TuningPyTorch +4 more

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.