Fine-Tuning Jobs

107 open roles mentioning Fine-Tuning

Software Engineer, Robotics

2mo ago
Scale AI

Scale AI

Scale's Robotics business unit is focused on solving the data challenges in Physical AI for Robotics, Autonomous Vehicles, and Computer Vision. In this role, you will be a key contributor to building production systems for robotics data collection, model training pipelines, and evaluation infrastructure. You will have the opportunity to own significant parts of our robotics platform, collaborate directly with cutting-edge robotics and AV customers, and influence the future of embodied AI systems.

Argentina; Uruguay remote
AWSDockerKubernetes +9 more

GenAI Strategic Projects Lead, Public Sector

2mo ago
Scale AI

Scale AI

Scale AI is seeking a Strategic Projects Lead for its Public Sector team to own high-impact projects focused on Generative AI and Large Language Models. This role involves working across operations, engineering, and customer engagement to produce high-quality training and test data for LLMs, particularly for Public Sector customers. You will be instrumental in building Generative AI data-labeling pipelines, creating operational processes for an expert data workforce, and developing novel technology-driven approaches to enhance data quality. This is a unique opportunity to contribute at the intersection of AI and national security, partnering with internal ML experts and external stakeholders to ensure data supports mission-critical AI applications.

$170k - $212k

Washington, DC onsite
Prompt EngineeringFine-TuningReinforcement Learning +9 more

Machine Learning Research Scientist, Reasoning

2mo ago
Scale AI

Scale AI

This role operates at the forefront of AI research and real-world implementation, with a strong focus on reasoning within large language models (LLMs). You will play a key role in shaping Scale’s data strategy by identifying the most effective data sources and methodologies for improving LLM reasoning. Success in this role requires a deep understanding of LLMs, planning algorithms, and novel approaches to agentic reasoning, as well as creativity in tackling challenges related to data generation, model interaction, and evaluation. You will contribute to impactful research on language model reasoning, collaborate with external researchers, and work closely with engineering teams to bring state-of-the-art advancements into scalable, real-world solutions.

$252k - $315k

San Francisco, CA; Seattle, WA; New York, NY onsite
LangGraphAWSFine-Tuning +10 more

Senior / Staff Machine Learning Research Scientist, Agents

2mo ago
Scale AI

Scale AI

This role is at the intersection of cutting-edge AI research and practical application, with a focus on studying the data types essential for building state-of-the-art agents, such as browser and SWE agents. The ideal candidate will explore the data landscape needed to advance intelligent, adaptable AI agents, guiding the data strategy at Scale to drive innovation. This position requires not only expertise in LLM agents and planning algorithms but also creativity in addressing novel challenges related to data, interaction, and evaluation. You will contribute to impactful research publications on agents, collaborate with customer researchers, and work alongside the engineering team to translate these advancements into real-world, scalable solutions.

$302k - $378k

San Francisco, CA; Seattle, WA; New York, NY onsite
LangGraphAWSAI Agents +8 more

Machine Learning Research Scientist, Post-Training

2mo ago
Scale AI

Scale AI

Scale works with leading AI labs to accelerate progress in GenAI research, focusing on optimizing data curation and evaluation to enhance LLM capabilities in text and multimodal modalities. This role involves developing novel methods to improve the alignment and generalization of large-scale generative models, collaborating with researchers and engineers on best practices in data-driven AI development, and providing technical and strategic input to foundation model labs for the next generation of AI models.

$252k - $315k

San Francisco, CA; Seattle, WA; New York, NY onsite
LLMFine-TuningReinforcement Learning +7 more

Research Internship (Fall, Winter 2026)

3mo ago
Cohere

Cohere

Cohere is seeking a Research Intern to collaborate with researchers and tools on designing and implementing novel research ideas and shipping state-of-the-art models to production. Interns will have the opportunity to work on various teams covering base model training, retrieval augmented generation, data and evaluation, safety, and finetuning, or any research area relating to LLMs. This role offers a chance to broaden research connections while gaining deep experience in a growing AI startup.

Canada remote FullTime
CoherePythonC# +5 more

Staff Software Engineer, Anti-Abuse & Security

3mo ago
Replit

Replit

The Anti-Abuse team at Replit is responsible for defending the platform from exploitation by detecting and preventing various forms of abuse, including phishing, cryptomining, LLM token farming, and malicious use of the platform. This role involves adversarial work, requiring the development of detection systems, heuristics, and automated responses to stay ahead of constantly adapting attackers. A unique aspect of this role is the AI-native nature of Replit's platform, presenting opportunities to build guardrails for AI-generated code, detect prompt injection attacks at scale, and leverage LLMs for defense. The position offers hands-on experience applying AI to security problems in a production environment with real attackers, with end-to-end ownership from identifying abuse patterns to shipping scalable solutions.

Foster City, CA hybrid FullTime
KubernetesPythonTypeScript +4 more

Forward Deployed Engineer - LLM Post-training

3mo ago
Reflection ai

Reflection ai

Reflection is a research lab dedicated to making intelligence open and accessible. We build open-weight models that empower users to control their AI and shape its future. As a core member of the Applied AI team, you will drive model fine-tuning and evaluations for enterprise customers. This role involves adapting our open-weight models for specific customer domains, tasks, and constraints, working hands-on with customer data, running fine-tuning workflows, building evaluation harnesses, and deploying adapted models to production. You will collaborate directly with customers to understand their needs and with research teams to advance the possibilities of AI.

San Francisco, CA onsite FullTime
PythonFine-TuningRLHF

Research Engineer, Data Infrastructure

3mo ago
Mistral AI

Mistral AI

Mistral AI is seeking a Research Engineer focused on Data Infrastructure to build and operate the next generation of our data systems. This role involves designing and scaling massive compute fleets and storage systems for high performance and scalability. You will contribute to a future of decoupled control and data planes, scaling big data compute and storage platforms while ensuring secure and governed data access for MLOps and research. The position requires full lifecycle ownership, from architecting migrations away from legacy orchestrators to implementing production-grade pipelines and participating in on-call rotations for critical training jobs.

Palo Alto remote Full-time
MistralKubernetesPython +5 more

Applied AI, Forward Deployed Machine Learning Engineer - Palo Alto

3mo ago
Mistral AI

Mistral AI

Mistral AI is seeking an Applied AI Engineer to join their team and facilitate the adoption of their products among customers. This role involves collaborating closely with clients from pre-sale to post-implementation, ensuring their solutions meet and exceed expectations. You will manage daily customer relations, acting as a key resource for externalizing research into production settings and driving the successful deployment of Mistral AI products. The position offers the opportunity to work on state-of-the-art Generative AI applications across various industries and contribute to a pioneering company shaping the future of AI.

Palo Alto onsite Full-time
MistralLangChainPython +9 more

Research Engineer, Data Infrastructure

3mo ago
Mistral AI

Mistral AI

Mistral AI is seeking a Research Engineer focused on Data Infrastructure to architect and build the backbone of our frontier model training and fine-tuning ecosystem. This role involves designing and scaling massive compute fleets and storage systems for high performance and scalability, with a vision towards exabyte-scale architecture. You will contribute to a strategic transition from legacy scheduling to modern orchestration, implementing sophisticated multi-cluster orchestration and cloud-bursting capabilities. The position requires full lifecycle ownership, from architecting migrations away from legacy orchestrators to implementing production-grade pipelines and participating in on-call rotations for critical training jobs.

Paris hybrid Full-time
MistralKubernetesPython +5 more

Applied AI, Forward Deployed Machine Learning Engineer - Montreal

4mo ago
Mistral AI

Mistral AI

Mistral AI is seeking an Applied AI Engineer to join their team and facilitate customer adoption of their AI products. This role involves collaborating closely with customers to address complex technical challenges, from pre-sale to post-implementation, ensuring solutions meet and exceed client expectations. You will manage daily customer relations, act as a key resource for externalizing research into production, and work on state-of-the-art Generative AI applications across various industries. This position offers the opportunity to contribute to a pioneering company shaping the future of AI and make a meaningful impact.

Montreal onsite Full-time
MistralLangChainPython +9 more

Solutions Architect - Public Sector

4mo ago
Cohere

Cohere

Cohere is seeking a Solutions Architect to play a significant role in growing the company's Defence and National Security business. This dynamic role requires a blend of strategic thinking and hands-on execution, focusing on building customer demos and proof-of-concepts that highlight the business value of Cohere's AI platform. As the technical relationship owner, you will collaborate with stakeholders to understand their objectives and translate them into technical solutions. You will also serve as the voice of the customer, liaising between clients and the product team, providing guidance on best practices, identifying platform improvements, and cultivating technical champions within customer organizations to drive adoption and gather feedback.

$265k - $315k

Ottawa hybrid FullTime
CohereAWSAzure +5 more

Research Scientist, Multimodal Alignment, Safety, and Fairness

4mo ago
Google DeepMind

Google DeepMind

Google DeepMind's Frontier AI unit is seeking experienced Research Scientists to join a multimodal safety research effort. This role focuses on interdisciplinary sociotechnical modeling and requires a passion for understanding AI-society interactions, a strong awareness of AI alignment and safety, and a drive to develop novel ideas, methods, interfaces, and tools. You will contribute to advancing the state of the art in AI research and Google DeepMind's mission towards Artificial General Intelligence (AGI), with a focus on leading new breakthrough research directions in areas like AI behavior exploration, assessment, and steering, particularly for subjective and creative tasks. The work involves tackling fundamental research questions to improve alignment objectives, assess adherence to desired behaviors, and enable AI agents to monitor real-world social context and evolve system behaviors over long time-horizons. You will develop new paradigms for human+AI rating that are adaptive and context-aware, driving breakthroughs within Google DeepMind, Google products, and the broader AI alignment community.

Kirkland, Washington, US; Mountain View, California, US; New York City, New York, US onsite
PythonAI AgentsFine-Tuning +5 more

Research Scientist, Gemini Safety

4mo ago
Google DeepMind

Google DeepMind

The Gemini Safety team at Google DeepMind is responsible for the safety and fairness of the latest Gemini models. As a Research Scientist/Engineer, you will apply and develop cutting-edge data and algorithmic solutions to advance these user-facing models. This is a fast-paced, highly collaborative role within a supportive team dedicated to pushing the boundaries of AI for public benefit and scientific discovery, with safety and ethics as the highest priorities.

Mountain View, California, US onsite
Fine-TuningReinforcement LearningJAX +1 more

Research Engineer

5mo ago
Cohere

Cohere

Cohere Labs is seeking Research Engineers to join their dedicated research arm, focused on pushing machine learning forward through open, collaborative research and hands-on experimentation. This is a highly practical role where you will work closely with scientists and engineers to implement new methods, run large-scale experiments, and help shape the infrastructure supporting our research programs. You will be responsible for building experiments, debugging models, scaling training pipelines, and turning research ideas into working systems. We value curiosity, strong fundamentals, and a willingness to learn quickly in a fast-moving research environment, with a focus on practical impact.

Toronto remote FullTime
CohereFine-TuningPyTorch +4 more

Applied Legal Researcher

5mo ago
Harvey

Harvey

Harvey is transforming how legal and professional services operate by combining frontier agentic AI, an enterprise-grade platform, and deep domain expertise. We are looking for a legal researcher with a strong understanding of how law firms, financial institutions, large corporations, and other organizations deliver professional services and execute complex legal and knowledge work tasks. This role is crucial for accelerating customer workflows with AI systems, requiring a deep understanding of their needs and how AI can be applied.

$180k - $220k

New York onsite FullTime
Prompt EngineeringFine-Tuning

Applied Legal Researcher

5mo ago
Harvey

Harvey

Harvey is transforming how legal and professional services operate by combining frontier agentic AI, an enterprise-grade platform, and deep domain expertise. We are looking for a legal researcher with a strong understanding of how law firms, financial institutions, large corporations, and other organizations deliver professional services and execute complex legal and knowledge work tasks. This role is crucial for accelerating customer workflows with AI systems, requiring a deep understanding of their needs and how AI can be applied.

$180k - $220k

San Francisco onsite FullTime
Prompt EngineeringFine-Tuning

Senior Software Engineer, Anti-Abuse & Security

6mo ago
Replit

Replit

The Anti-Abuse team at Replit is responsible for defending the platform from exploitation by detecting and shutting down various forms of abuse, including phishing, cryptomining, and LLM token farming. This role involves building advanced detection systems, heuristics, and automated responses to stay ahead of constantly adapting attackers. A unique aspect of this position is the AI-native nature of Replit's platform, offering hands-on experience in applying AI to security problems, such as building guardrails for AI-generated code, detecting prompt injection attacks, and using LLMs defensively. You will own problems end-to-end, from identifying abuse patterns to shipping scalable solutions.

Foster City, CA hybrid FullTime
KubernetesPythonTypeScript +4 more

Forward Deployed Engineer - AI Engineer

7mo ago
Reflection ai

Reflection ai

Reflection is a research lab dedicated to making intelligence open and accessible. We are seeking a core member for our Applied AI team to lead Forward Deployed Engineering efforts with enterprise customers. This role involves translating advanced AI research into high-impact, real-world applications, owning the technical strategy and delivery of agentic systems from discovery to production launch.

New York, NY onsite FullTime
DockerKubernetesPython +3 more

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.