RLHF Jobs
68 open roles mentioning RLHF
Staff Software Engineer, Data Platform
Scale AI
Scale is at the forefront of the AI revolution, developing data engines and technologies that power the world's leading LLMs and generative models. This role is on the Platform Engineering team, responsible for the foundational data infrastructure that supports these cutting-edge AI products. You will lead the design and development of core data storage, streaming, caching, and indexing platforms, gaining exposure to the rapidly evolving AI landscape across various industries. The work involves driving architecture, implementation, and reliability of these critical systems, collaborating with stakeholders, and mentoring junior engineers.
$252k - $315k
Machine Learning Research Engineer, Agents - Enterprise GenAI
Scale AI
Scale is seeking a Machine Learning Research Engineer focused on Agents for its Enterprise GenAI team. This role will be instrumental in accelerating the development of AI applications by working on state-of-the-art post-training algorithms for complex enterprise agents. You will apply proprietary Agent RL Training + Building algorithms to real-world enterprise datasets and benchmarks, aiming to create best-in-class Agents that achieve state-of-the-art results. If you are passionate about shaping the future of GenAI, this is an opportunity to contribute to cutting-edge research and development.
$265k - $331k
Staff Machine Learning Research Engineer, Agent Post-training - Enterprise GenAI
Scale AI
Scale is seeking a Staff Agent Post-Training ML Research Engineer to join our Enterprise ML Research Lab. This role will focus on building out our next-generation Agent RL training platform, integrating cutting-edge research to train best-in-class Agents for real enterprise use-cases. You will contribute to the development of AI applications that are becoming vital across all sectors, from cybersecurity LLMs to foundation healthtech search models, shaping the future of the modern GenAI movement.
$265k - $331k
Forward Deployed Engineer, GenAI
Scale AI
Scale AI is seeking a Forward Deployed Engineer, GenAI to join their Data Engine team. This role is at the forefront of providing critical data infrastructure that powers advanced AI models, directly influencing how humanity interacts with AI. You will work with the world's leading AI companies and government agencies to solve their most complex AI data-related problems, contributing to the advancement of AI by delivering critical data solutions for leading AI innovators and government agencies. You will interact daily with technical customers, understand their unique challenges, and translate them into impactful solutions, while also designing, building, and deploying features across the entire stack. This position offers a unique opportunity to lead critical projects, shape engineering culture, and accelerate career growth in the rapidly evolving field of Generative AI.
$179k - $224k
Machine Learning Systems Research Engineer, Agent Post-training - Enterprise GenAI
Scale AI
Scale is seeking a Machine Learning Systems Research Engineer to join their Enterprise ML Research Lab. This role will focus on building algorithms for a next-generation Agent RL training platform, supporting large-scale training, and integrating state-of-the-art technologies to optimize ML systems. You will collaborate with other ML researchers and engineers who apply these algorithms to client use cases, including AI cybersecurity firewalls and healthtech search models. If you are passionate about shaping the future of AI, this is an exciting opportunity to contribute to cutting-edge advancements in enterprise GenAI.
$265k - $331k
Senior Software Engineer, GenAI
Scale AI
Scale AI is seeking a Senior Software Engineer to join our Generative AI Data Engine team. This role is crucial in accelerating the development of AI applications by powering the world's most advanced LLMs and generative models. You will work on high-impact datasets, optimize contributor onboarding, and ensure data integrity through advanced trust, safety, and security measures. This position operates at the intersection of ML, operations, and analytics to deliver high-quality data at scale, contributing to the future of human-AI interaction.
$216k - $270k
Tech Lead Manager- MLRE, ML Systems
Scale AI
Scale's LLM post-training platform team builds our internal distributed framework for large language model training, powering MLEs, researchers, data scientists, and operators for fast and automatic training and evaluation of LLMs. This platform also serves as the underlying training framework for the data quality evaluation pipeline. You will work closely with Scale’s ML teams and researchers to build the foundation platform which supports all our ML research and development works, optimizing it to enable next generation LLM training, inference, and data curation. If you are excited about shaping the future AI via fundamental innovations, we would love to hear from you!
$265k - $331k
Principal Applied Research Engineer
synthesia.io
Synthesia is seeking a Principal Research Engineer to lead the technical direction for offline video generation. This role involves owning the end-to-end process, from pre-training to post-training, and resolving the complexities that arise at scale. You will partner with research leadership to define long-term strategy, tackle challenging technical problems, and accelerate the delivery of research into production. The ideal candidate has hands-on experience training large generative models from scratch and is driven by a passion for pushing the boundaries of video generation and ensuring that work reaches users.
ML Researcher, Foundational Models
sarvam
Sarvam is building the bedrock of Sovereign AI for India, developing a full-stack AI platform focused on making AI genuinely work for India. This role is for a researcher who will tackle open-ended questions about the architecture, optimization, data composition, and training dynamics of our next generation of foundational models. You will have direct access to large compute resources and a tight feedback loop with engineers, driving research from initial hunches to production-ready decisions. This is a hands-on role requiring independent research design, execution at scale, and the ability to translate findings into concrete proposals for production training runs.
GenAI Strategic Projects Lead, Public Sector
Scale AI
Scale AI is seeking a Strategic Projects Lead for its Public Sector team to own high-impact projects focused on Generative AI and Large Language Models. This role involves working across operations, engineering, and customer engagement to produce high-quality training and test data for LLMs, particularly for Public Sector customers. You will be instrumental in building Generative AI data-labeling pipelines, creating operational processes for an expert data workforce, and developing novel technology-driven approaches to enhance data quality. This is a unique opportunity to contribute at the intersection of AI and national security, partnering with internal ML experts and external stakeholders to ensure data supports mission-critical AI applications.
$170k - $212k
Machine Learning Research Scientist, Post-Training
Scale AI
Scale works with leading AI labs to accelerate progress in GenAI research, focusing on optimizing data curation and evaluation to enhance LLM capabilities in text and multimodal modalities. This role involves developing novel methods to improve the alignment and generalization of large-scale generative models, collaborating with researchers and engineers on best practices in data-driven AI development, and providing technical and strategic input to foundation model labs for the next generation of AI models.
$252k - $315k
ML Research Engineer, ML Systems
Scale AI
Scale's ML platform (RLXF) team builds our internal distributed framework for large language model training and inference. This platform powers MLEs, researchers, data scientists, and operators for fast and automatic training and evaluation of LLMs, as well as data quality evaluation. You will work closely across Scale’s ML teams and researchers to build the foundation platform that supports all our ML research and development, optimizing it to enable the next generation of LLM training, inference, and data curation. If you are excited about shaping the future of AI via fundamental innovations, we would love to hear from you!
$190k - $237k
Research, Post-Training Data
thinkingmachines
Thinking Machines Lab is seeking researchers to bridge the gap between raw AI intelligence and useful, safe, and collaborative systems. This role focuses on post-training data research, combining human insight and machine learning techniques to capture and steer model behavior based on human preferences. You will be responsible for translating research ideas into actionable data through labeling and collection campaigns, understanding data quality science, and developing metrics to measure the impact of data and training interventions. The position also involves exploring new paradigms for human-AI interaction and scalable oversight, blending research, data operations, and technical implementation to advance human-centered AI systems. This role requires both fundamental research and practical engineering, making it ideal for individuals who enjoy deep theoretical exploration and hands-on experimentation.
$350k - $475k
Code Data Annotation Quality Specialist
Mistral AI
Mistral is seeking highly motivated Data Quality Specialists to join our Human Data Annotation team. This hybrid role involves reviewing and auditing code annotations against rubrics to ensure high-quality data for AI model training and evaluation. You will also be responsible for building, maintaining, and troubleshooting the internal tooling that annotators use daily. This position offers the opportunity to collaborate closely with annotators, technical program managers, and engineering stakeholders, contributing to the refinement of guidelines and processes that shape data production.
Forward Deployed Engineer - LLM Post-training
Reflection ai
Reflection is a research lab dedicated to making intelligence open and accessible. We build open-weight models that empower users to control their AI and shape its future. As a core member of the Applied AI team, you will drive model fine-tuning and evaluations for enterprise customers. This role involves adapting our open-weight models for specific customer domains, tasks, and constraints, working hands-on with customer data, running fine-tuning workflows, building evaluation harnesses, and deploying adapted models to production. You will collaborate directly with customers to understand their needs and with research teams to advance the possibilities of AI.
Member of Technical Staff - Post Training
Black Forest Labs
We are seeking a Member of Technical Staff specializing in Post Training to own the end-to-end post-training pipeline for our multimodal generative models. This role is crucial for transforming foundation models into polished products, encompassing data strategy, reward modeling, preference optimization, distillation, and safety tuning across image, editing, and video modalities. You will be instrumental in driving significant improvements in model quality, developing the infrastructure that accelerates research team iteration, and advancing the state-of-the-art in aligning generative models with human intent. This is a Staff/Senior individual contributor position for someone with proven experience shipping post-training for a frontier model.
€130k - €340k
Researcher, Vision
sarvam
Sarvam is building the bedrock of Sovereign AI for India, developing a full-stack sovereign AI platform across research, models, infrastructure, and applications. The company partners with leading enterprises and public institutions, backed by prominent venture capital firms and collaborating with India's top brands. This role involves working across the full lifecycle of vision-language model (VLM) development, including data, training, evaluation, and production. We seek researchers who are comfortable with the evolving nature of the field and can take a leading role.
Data Scientist - Evaluations, Chanakya
sarvam
Sarvam is building India's full-stack sovereign AI platform, focusing on making AI genuinely work for India across research, models, infrastructure, and applications. This intellectually demanding role anchors the evaluations function for the AI vertical. You will design, build, and maintain evaluation frameworks to measure model and system quality in operational contexts, focusing on domain-specific, high-stakes use cases where accuracy is critical. You will collaborate closely with MLOps Engineers, Product Managers, and the deployment team to ensure deployed AI systems meet stringent quality standards.
Forward Deployed Engineer - ML
Modal
We are seeking Forward Deployed ML Engineers to work at the intersection of deep technical work and direct customer impact. As an ML FDE, you will partner with leading AI companies and foundation model labs to help them achieve state-of-the-art performance on their most demanding workloads, including LLM serving, model training (SFT, RLHF), audio pipelines, and scientific computing. You will help teams reach outcomes that many engineers cannot achieve on their own. The FDE team comprises world-class software engineers, computational scientists, ML engineers, and former founders, and we are looking for individuals with strong engineering fundamentals, deep curiosity across the AI stack, and energy for tackling hard problems directly with customers.
Forward Deployed Engineer - ML
Modal
We are seeking Forward Deployed ML Engineers to work at the intersection of deep technical work and direct customer impact. As an ML FDE, you will partner with leading AI companies and foundation model labs to help them achieve state-of-the-art performance on their most demanding workloads, including LLM serving, model training (SFT, RLHF), audio pipelines, and scientific computing. You will help teams reach outcomes that many engineers cannot achieve on their own. The FDE team comprises world-class software engineers, computational scientists, ML engineers, and former founders, and we are looking for individuals with strong engineering fundamentals, deep curiosity across the AI stack, and energy for tackling hard problems directly with customers.