Fine-Tuning Jobs
107 open roles mentioning Fine-Tuning
Software Engineer, Robotics
Scale AI
Scale's Robotics business unit is focused on solving the data challenges in Physical AI for Robotics, Autonomous Vehicles, and Computer Vision. In this role, you will be a key contributor to building production systems for robotics data collection, model training pipelines, and evaluation infrastructure. You will have the opportunity to own significant parts of our robotics platform, collaborate directly with cutting-edge robotics and AV customers, and influence the future of embodied AI systems.
GenAI Strategic Projects Lead, Public Sector
Scale AI
Scale AI is seeking a Strategic Projects Lead for its Public Sector team to own high-impact projects focused on Generative AI and Large Language Models. This role involves working across operations, engineering, and customer engagement to produce high-quality training and test data for LLMs, particularly for Public Sector customers. You will be instrumental in building Generative AI data-labeling pipelines, creating operational processes for an expert data workforce, and developing novel technology-driven approaches to enhance data quality. This is a unique opportunity to contribute at the intersection of AI and national security, partnering with internal ML experts and external stakeholders to ensure data supports mission-critical AI applications.
$170k - $212k
Machine Learning Research Scientist, Reasoning
Scale AI
This role operates at the forefront of AI research and real-world implementation, with a strong focus on reasoning within large language models (LLMs). You will play a key role in shaping Scale’s data strategy by identifying the most effective data sources and methodologies for improving LLM reasoning. Success in this role requires a deep understanding of LLMs, planning algorithms, and novel approaches to agentic reasoning, as well as creativity in tackling challenges related to data generation, model interaction, and evaluation. You will contribute to impactful research on language model reasoning, collaborate with external researchers, and work closely with engineering teams to bring state-of-the-art advancements into scalable, real-world solutions.
$252k - $315k
Senior / Staff Machine Learning Research Scientist, Agents
Scale AI
This role is at the intersection of cutting-edge AI research and practical application, with a focus on studying the data types essential for building state-of-the-art agents, such as browser and SWE agents. The ideal candidate will explore the data landscape needed to advance intelligent, adaptable AI agents, guiding the data strategy at Scale to drive innovation. This position requires not only expertise in LLM agents and planning algorithms but also creativity in addressing novel challenges related to data, interaction, and evaluation. You will contribute to impactful research publications on agents, collaborate with customer researchers, and work alongside the engineering team to translate these advancements into real-world, scalable solutions.
$302k - $378k
Machine Learning Research Scientist, Post-Training
Scale AI
Scale works with leading AI labs to accelerate progress in GenAI research, focusing on optimizing data curation and evaluation to enhance LLM capabilities in text and multimodal modalities. This role involves developing novel methods to improve the alignment and generalization of large-scale generative models, collaborating with researchers and engineers on best practices in data-driven AI development, and providing technical and strategic input to foundation model labs for the next generation of AI models.
$252k - $315k
Research Internship (Fall, Winter 2026)
Cohere
Cohere is seeking a Research Intern to collaborate with researchers and tools on designing and implementing novel research ideas and shipping state-of-the-art models to production. Interns will have the opportunity to work on various teams covering base model training, retrieval augmented generation, data and evaluation, safety, and finetuning, or any research area relating to LLMs. This role offers a chance to broaden research connections while gaining deep experience in a growing AI startup.
Staff Software Engineer, Anti-Abuse & Security
Replit
The Anti-Abuse team at Replit is responsible for defending the platform from exploitation by detecting and preventing various forms of abuse, including phishing, cryptomining, LLM token farming, and malicious use of the platform. This role involves adversarial work, requiring the development of detection systems, heuristics, and automated responses to stay ahead of constantly adapting attackers. A unique aspect of this role is the AI-native nature of Replit's platform, presenting opportunities to build guardrails for AI-generated code, detect prompt injection attacks at scale, and leverage LLMs for defense. The position offers hands-on experience applying AI to security problems in a production environment with real attackers, with end-to-end ownership from identifying abuse patterns to shipping scalable solutions.
Forward Deployed Engineer - LLM Post-training
Reflection ai
Reflection is a research lab dedicated to making intelligence open and accessible. We build open-weight models that empower users to control their AI and shape its future. As a core member of the Applied AI team, you will drive model fine-tuning and evaluations for enterprise customers. This role involves adapting our open-weight models for specific customer domains, tasks, and constraints, working hands-on with customer data, running fine-tuning workflows, building evaluation harnesses, and deploying adapted models to production. You will collaborate directly with customers to understand their needs and with research teams to advance the possibilities of AI.
Research Engineer, Data Infrastructure
Mistral AI
Mistral AI is seeking a Research Engineer focused on Data Infrastructure to build and operate the next generation of our data systems. This role involves designing and scaling massive compute fleets and storage systems for high performance and scalability. You will contribute to a future of decoupled control and data planes, scaling big data compute and storage platforms while ensuring secure and governed data access for MLOps and research. The position requires full lifecycle ownership, from architecting migrations away from legacy orchestrators to implementing production-grade pipelines and participating in on-call rotations for critical training jobs.
Applied AI, Forward Deployed Machine Learning Engineer - Palo Alto
Mistral AI
Mistral AI is seeking an Applied AI Engineer to join their team and facilitate the adoption of their products among customers. This role involves collaborating closely with clients from pre-sale to post-implementation, ensuring their solutions meet and exceed expectations. You will manage daily customer relations, acting as a key resource for externalizing research into production settings and driving the successful deployment of Mistral AI products. The position offers the opportunity to work on state-of-the-art Generative AI applications across various industries and contribute to a pioneering company shaping the future of AI.
Research Engineer, Data Infrastructure
Mistral AI
Mistral AI is seeking a Research Engineer focused on Data Infrastructure to architect and build the backbone of our frontier model training and fine-tuning ecosystem. This role involves designing and scaling massive compute fleets and storage systems for high performance and scalability, with a vision towards exabyte-scale architecture. You will contribute to a strategic transition from legacy scheduling to modern orchestration, implementing sophisticated multi-cluster orchestration and cloud-bursting capabilities. The position requires full lifecycle ownership, from architecting migrations away from legacy orchestrators to implementing production-grade pipelines and participating in on-call rotations for critical training jobs.
Applied AI, Forward Deployed Machine Learning Engineer - Montreal
Mistral AI
Mistral AI is seeking an Applied AI Engineer to join their team and facilitate customer adoption of their AI products. This role involves collaborating closely with customers to address complex technical challenges, from pre-sale to post-implementation, ensuring solutions meet and exceed client expectations. You will manage daily customer relations, act as a key resource for externalizing research into production, and work on state-of-the-art Generative AI applications across various industries. This position offers the opportunity to contribute to a pioneering company shaping the future of AI and make a meaningful impact.
Solutions Architect - Public Sector
Cohere
Cohere is seeking a Solutions Architect to play a significant role in growing the company's Defence and National Security business. This dynamic role requires a blend of strategic thinking and hands-on execution, focusing on building customer demos and proof-of-concepts that highlight the business value of Cohere's AI platform. As the technical relationship owner, you will collaborate with stakeholders to understand their objectives and translate them into technical solutions. You will also serve as the voice of the customer, liaising between clients and the product team, providing guidance on best practices, identifying platform improvements, and cultivating technical champions within customer organizations to drive adoption and gather feedback.
$265k - $315k
Research Scientist, Multimodal Alignment, Safety, and Fairness
Google DeepMind
Google DeepMind's Frontier AI unit is seeking experienced Research Scientists to join a multimodal safety research effort. This role focuses on interdisciplinary sociotechnical modeling and requires a passion for understanding AI-society interactions, a strong awareness of AI alignment and safety, and a drive to develop novel ideas, methods, interfaces, and tools. You will contribute to advancing the state of the art in AI research and Google DeepMind's mission towards Artificial General Intelligence (AGI), with a focus on leading new breakthrough research directions in areas like AI behavior exploration, assessment, and steering, particularly for subjective and creative tasks. The work involves tackling fundamental research questions to improve alignment objectives, assess adherence to desired behaviors, and enable AI agents to monitor real-world social context and evolve system behaviors over long time-horizons. You will develop new paradigms for human+AI rating that are adaptive and context-aware, driving breakthroughs within Google DeepMind, Google products, and the broader AI alignment community.
Research Scientist, Gemini Safety
Google DeepMind
The Gemini Safety team at Google DeepMind is responsible for the safety and fairness of the latest Gemini models. As a Research Scientist/Engineer, you will apply and develop cutting-edge data and algorithmic solutions to advance these user-facing models. This is a fast-paced, highly collaborative role within a supportive team dedicated to pushing the boundaries of AI for public benefit and scientific discovery, with safety and ethics as the highest priorities.
Research Engineer
Cohere
Cohere Labs is seeking Research Engineers to join their dedicated research arm, focused on pushing machine learning forward through open, collaborative research and hands-on experimentation. This is a highly practical role where you will work closely with scientists and engineers to implement new methods, run large-scale experiments, and help shape the infrastructure supporting our research programs. You will be responsible for building experiments, debugging models, scaling training pipelines, and turning research ideas into working systems. We value curiosity, strong fundamentals, and a willingness to learn quickly in a fast-moving research environment, with a focus on practical impact.
Applied Legal Researcher
Harvey
Harvey is transforming how legal and professional services operate by combining frontier agentic AI, an enterprise-grade platform, and deep domain expertise. We are looking for a legal researcher with a strong understanding of how law firms, financial institutions, large corporations, and other organizations deliver professional services and execute complex legal and knowledge work tasks. This role is crucial for accelerating customer workflows with AI systems, requiring a deep understanding of their needs and how AI can be applied.
$180k - $220k
Applied Legal Researcher
Harvey
Harvey is transforming how legal and professional services operate by combining frontier agentic AI, an enterprise-grade platform, and deep domain expertise. We are looking for a legal researcher with a strong understanding of how law firms, financial institutions, large corporations, and other organizations deliver professional services and execute complex legal and knowledge work tasks. This role is crucial for accelerating customer workflows with AI systems, requiring a deep understanding of their needs and how AI can be applied.
$180k - $220k
Senior Software Engineer, Anti-Abuse & Security
Replit
The Anti-Abuse team at Replit is responsible for defending the platform from exploitation by detecting and shutting down various forms of abuse, including phishing, cryptomining, and LLM token farming. This role involves building advanced detection systems, heuristics, and automated responses to stay ahead of constantly adapting attackers. A unique aspect of this position is the AI-native nature of Replit's platform, offering hands-on experience in applying AI to security problems, such as building guardrails for AI-generated code, detecting prompt injection attacks, and using LLMs defensively. You will own problems end-to-end, from identifying abuse patterns to shipping scalable solutions.
Forward Deployed Engineer - AI Engineer
Reflection ai
Reflection is a research lab dedicated to making intelligence open and accessible. We are seeking a core member for our Applied AI team to lead Forward Deployed Engineering efforts with enterprise customers. This role involves translating advanced AI research into high-impact, real-world applications, owning the technical strategy and delivery of agentic systems from discovery to production launch.