RLHF Jobs
68 open roles mentioning RLHF
Member of Technical Staff - Safety
Reflection ai
Reflection is a research lab dedicated to making intelligence open and accessible. We build open models that empower users to control their intelligence and shape the future of AI. As a Member of Technical Staff - Safety, you will be instrumental in ensuring the safety and reliability of our AI models. This role involves owning the red-teaming and adversarial evaluation pipeline, translating safety findings into concrete guardrails, and validating that every release meets our risk thresholds before deployment. You will develop scalable, automated safety benchmarks and research state-of-the-art jailbreaking techniques and defenses to proactively address potential vulnerabilities.
Product Manager, Forge
Mistral AI
Mistral AI is seeking a talented and experienced Product Manager to define and execute the strategy for Forge, a product that empowers customers to build, fine-tune, and deploy custom AI models at scale. Forge transforms cutting-edge research into enterprise-ready capabilities by supporting model fine-tuning, reinforcement learning, and post-training workflows. This role operates at the intersection of research and product, enabling customers to train specialized models for real-world business value. You will collaborate closely with applied AI scientists and research engineers to translate frontier techniques into scalable and reliable solutions, shaping a 0-1 product with significant business impact and defining the future of how organizations train and deploy AI models.
Researcher, Post Training
cartesia
Cartesia is seeking a Researcher for their Post-Training team to develop methods and systems that make multimodal models adaptive, aligned, and grounded in human intent. This role involves working at the intersection of machine learning research, alignment, and infrastructure, focusing on preference optimization, model evaluation, and feedback-driven learning. You will explore how feedback signals can guide models to reason more effectively across modalities and build the infrastructure to measure and improve these behaviors at scale. Your work will directly influence how Cartesia's foundation models learn, improve, and connect with people.
Research Engineer, Codex
OpenAI
The Codex Research team at OpenAI is responsible for developing frontier AI agents, including the models behind Codex, ChatGPT, and other advanced products. This role focuses on improving the capabilities, reliability, and product fit of these agentic models. You will have the opportunity to own research directions, build critical training infrastructure, create evaluation methods, or drive capabilities from initial concept through experimentation and launch. This is a broad, high-agency role for individuals who can tackle ambiguous problems across research, engineering, data, evals, and product, and who are excited about building AI that can act in the world, write code, use tools, and collaborate with users and other agents.
Applied Machine Learning Engineer
fireworks ai
As an Applied Machine Learning Engineer, you will serve as a vital bridge between cutting-edge AI research and practical, real-world applications. Your work will focus on developing, fine-tuning, and operationalizing machine learning models that drive business value and enhance user experiences. This is a hands-on engineering role that combines deep technical expertise with a strong customer focus to deliver scalable AI solutions.
Research Engineer, Frontier Evals & Environments
OpenAI
OpenAI is seeking a Research Engineer for Frontier Evals & Environments to help build north star model environments that drive progress towards safe AGI/ASI. This role will directly guide the research programs of ambitious training runs, with past open-sourced evaluations including GDPval, SWE-bench Verified, MLE-bench, PaperBench, and SWE-Lancer. You will work with researchers, engineers, product teams, infrastructure teams, and safety/alignment partners to define model run objectives, measure outcomes, and ship improvements into widely used products. This is a high-agency role for individuals passionate about seeing their work directly impact frontier models and steering them towards beneficial outcomes.
Researcher, Safety Oversight
OpenAI
OpenAI is seeking a senior researcher passionate about AI safety to join the Safety Oversight Research team. This role will be instrumental in setting research directions to ensure safe AGI and will focus on identifying and mitigating misuse and misalignment in AI systems. You will play a critical part in defining the future of safe AI systems at OpenAI, significantly impacting our mission to build and deploy safe AGI. The Safety Systems team is dedicated to ensuring our advanced models can be safely deployed for societal benefit, driving our commitment to AI safety and fostering trust and transparency.
Researcher, Robustness & Safety Training
OpenAI
OpenAI is seeking a senior researcher passionate about AI safety to set research directions and conduct projects aimed at making AI systems safer, more aligned, and robust against adversarial or malicious use cases. This role is critical in shaping the future of safe AI systems at OpenAI and significantly impacting the mission to build and deploy safe AGI. The Safety Systems team focuses on ensuring AI models can be safely deployed to benefit society, advancing capabilities for robust and safe AI behavior, and addressing challenges as AI becomes more powerful and widely used.