evaluation Jobs

2 open roles mentioning evaluation

Product Manager, Gen AI

2mo ago
Scale AI

Scale AI

Scale AI is seeking Product Managers to join its GenAI organization, focusing on building the data infrastructure that powers advanced AI. These roles involve shaping systems, tooling, and experiences for a two-sided marketplace connecting AI labs and enterprises with a global network of contributors. You will work on high-impact, technically complex problems at the frontier of AI, owning product areas end-to-end from strategy to execution. The positions require deep cross-functional collaboration with engineering, design, data science, operations, and other stakeholders in a fast-paced, growth-stage environment.

New York, NY; San Francisco, CA remote
AIMachine LearningGenAI +5 more

Model Behavior Architect- Safety

10mo ago
Mistral AI

Mistral AI

As a Model Behavior Architect, you will be at the forefront of defining and measuring LLM behavior. We are seeking individuals with a background in engineering, machine learning, and large language models, who possess expertise in model evaluation, policy writing, and creating evaluation pipelines for complex tasks. Your role will involve close collaboration with our Science team to establish standards for Reasoning, Audio, Alignment, Tools, and other frontier initiatives. This is an opportunity to engage with cutting-edge, open-ended research challenges and translate your insights into superior models.

Paris remote Full-time
MistralLLMsMachine Learning +3 more

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.