Fine-Tuning Jobs
175 open roles mentioning Fine-Tuning
Forward Deployed Product Manager, Enterprise
Scale AI
Scale is seeking a Forward Deployed Product Manager (FDPM) to drive the success of enterprise AI deployments. This role is embedded with customers, focusing on achieving real production outcomes and translating operational realities into actionable product insights. Unlike a traditional roadmap PM or solutions engineer, the FDPM owns product outcomes within a portfolio of enterprise accounts, building trust with senior stakeholders, guiding deployments to production, and identifying product versus execution bottlenecks. The ideal candidate can differentiate between stated customer needs, underlying problems, and optimal platform solutions, having previously shipped products into large organizations.
$206k - $300k
Software Engineer, Robotics
Scale AI
Scale's Robotics business unit is focused on solving the data bottleneck in Physical AI across Robotics, Autonomous Vehicles, and Computer Vision. In this role, you will be a key contributor building production systems for robotics data collection, model training pipelines, and evaluation infrastructure. You will have the opportunity to own critical parts of our robotics platform, work directly with cutting-edge robotics and AV customers, and shape the future of embodied AI systems.
Staff Software Engineer, Data Platform
Scale AI
Scale is at the forefront of the AI revolution, developing data engines and technologies that power the world's leading LLMs and generative models. This role is on the Platform Engineering team, responsible for the foundational data infrastructure that supports these cutting-edge AI products. You will lead the design and development of core data storage, streaming, caching, and indexing platforms, gaining exposure to the rapidly evolving AI landscape across various industries. The work involves driving architecture, implementation, and reliability of these critical systems, collaborating with stakeholders, and mentoring junior engineers.
$252k - $315k
Software Engineer, Robotics & Autonomous Systems
Scale AI
Scale's Robotics business unit is dedicated to solving the data bottleneck in Physical AI across Robotics, Autonomous Vehicles, and Computer Vision. In this role, you'll be a key contributor building production systems for robotics data collection, model training pipelines, and evaluation infrastructure. You'll have the opportunity to own critical parts of our robotics platform, work directly with cutting-edge robotics and AV customers, and shape the future of embodied AI systems.
$180k - $225k
Principal Applied Research Engineer
synthesia.io
Synthesia is seeking a Principal Research Engineer to lead the technical direction for offline video generation. This role involves owning the end-to-end process, from pre-training to post-training, and resolving the complexities that arise at scale. You will partner with research leadership to define long-term strategy, tackle challenging technical problems, and accelerate the delivery of research into production. The ideal candidate has hands-on experience training large generative models from scratch and is driven by a passion for pushing the boundaries of video generation and ensuring that work reaches users.
ML Researcher, Foundational Models
sarvam
Sarvam is building the bedrock of Sovereign AI for India, developing a full-stack AI platform focused on making AI genuinely work for India. This role is for a researcher who will tackle open-ended questions about the architecture, optimization, data composition, and training dynamics of our next generation of foundational models. You will have direct access to large compute resources and a tight feedback loop with engineers, driving research from initial hunches to production-ready decisions. This is a hands-on role requiring independent research design, execution at scale, and the ability to translate findings into concrete proposals for production training runs.
Staff Technical Program Manager, Managed Intelligence
crusoe
Crusoe is seeking a Staff Technical Program Manager to join their Managed Intelligence team. This role is crucial for connecting model engineering, IaaS, product, and data center operations to deliver a reliable and scalable inference platform for AI-native companies. You will own end-to-end program delivery, including multi-quarter roadmaps, model onboarding, inference optimization, and production readiness for new model versions. This is a unique opportunity to shape the TPM function within a rapidly growing, vertically integrated AI infrastructure company.
$200k - $240k
Solutions Architect - Defence and National Security
Cohere
Cohere is seeking a Solutions Architect to join our Defence and National Security team. This role requires a blend of strategic thinking and hands-on execution, focusing on building customer demos and proof-of-concepts that highlight the value of Cohere's AI platform. You will act as the technical relationship owner, collaborating with stakeholders to understand their business objectives and translate them into technical solutions. Your responsibilities will include serving as the voice of the customer, liaising between clients and our product team, guiding best practices, identifying platform improvements, and fostering technical champions within customer organizations to drive adoption and gather feedback.
Research Scientist, Foundation Model (Video Generation)
pika
Pika is pioneering the next generation of creative infrastructure with real-time, multimodal generation and intelligent agentic platforms. We are seeking accomplished Research Scientists in Foundation Models to advance our mission of making agentic, real-time generative technology accessible and transformative for millions of creators. As a key member of our research team, you will design and implement core technologies, develop new methodologies for large-scale multimodal pre-training/mid-training, and drive innovative approaches for foundational model architecture, collaborating closely with engineering and product teams to shape the future of real-time creative and agentic platforms at scale.
Forward Deployed Engineer, Lead - AI Engineer
Reflection ai
Reflection is a research lab dedicated to making intelligence open and accessible for everyone to use, customize, and build on. We are seeking an exceptional technical lead to build and scale our Forward Deployed Engineering function within the AI Solutions team. This role is critical in bridging our cutting-edge AI research with real-world enterprise deployments. As a Forward Deployed Engineer Lead, you will own the end-to-end technical strategy, execution, and delivery of complex agentic applications, from early pre-sales discovery through production deployment.
Machine Learning Research Scientist, Reasoning
Scale AI
This role operates at the forefront of AI research and real-world implementation, with a strong focus on reasoning within large language models (LLMs). You will play a key role in shaping Scale’s data strategy by identifying the most effective data sources and methodologies for improving LLM reasoning. Success in this role requires a deep understanding of LLMs, planning algorithms, and novel approaches to agentic reasoning, as well as creativity in tackling challenges related to data generation, model interaction, and evaluation. You will contribute to impactful research on language model reasoning, collaborate with external researchers, and work closely with engineering teams to bring state-of-the-art advancements into scalable, real-world solutions.
$252k - $315k
Senior / Staff Machine Learning Research Scientist, Agents
Scale AI
This role is at the intersection of cutting-edge AI research and practical application, with a focus on studying the data types essential for building state-of-the-art agents, such as browser and SWE agents. The ideal candidate will explore the data landscape needed to advance intelligent, adaptable AI agents, guiding the data strategy at Scale to drive innovation. This position requires not only expertise in LLM agents and planning algorithms but also creativity in addressing novel challenges related to data, interaction, and evaluation. You will contribute to impactful research publications on agents, collaborate with customer researchers, and work alongside the engineering team to translate these advancements into real-world, scalable solutions.
$302k - $378k
Machine Learning Research Scientist, Post-Training
Scale AI
Scale works with leading AI labs to accelerate progress in GenAI research, focusing on optimizing data curation and evaluation to enhance LLM capabilities in text and multimodal modalities. This role involves developing novel methods to improve the alignment and generalization of large-scale generative models, collaborating with researchers and engineers on best practices in data-driven AI development, and providing technical and strategic input to foundation model labs for the next generation of AI models.
$252k - $315k
GenAI Strategic Projects Lead, Public Sector
Scale AI
Scale AI is seeking a Strategic Projects Lead for its Public Sector team to own high-impact projects focused on Generative AI and Large Language Models. This role involves working across operations, engineering, and customer engagement to produce high-quality training and test data for LLMs, particularly for Public Sector customers. You will be instrumental in building Generative AI data-labeling pipelines, creating operational processes for an expert data workforce, and developing novel technology-driven approaches to enhance data quality. This is a unique opportunity to contribute at the intersection of AI and national security, partnering with internal ML experts and external stakeholders to ensure data supports mission-critical AI applications.
$170k - $212k
Software Engineer, Robotics
Scale AI
Scale's Robotics business unit is focused on solving the data challenges in Physical AI for Robotics, Autonomous Vehicles, and Computer Vision. In this role, you will be a key contributor to building production systems for robotics data collection, model training pipelines, and evaluation infrastructure. You will have the opportunity to own significant parts of our robotics platform, collaborate directly with cutting-edge robotics and AV customers, and influence the future of embodied AI systems.
Research Internship (Winter 2027)
Cohere
Cohere is seeking a Research Intern to collaborate with researchers and tools on designing and implementing novel research ideas and shipping state-of-the-art models to production. Interns will have the opportunity to work on various teams covering base model training, retrieval augmented generation, data and evaluation, safety, and finetuning, or any research area relating to LLMs. This role offers a chance to broaden research connections while gaining deep experience in a growing AI startup.
Associate Product Manager
fireworks ai
This is a rare opportunity for an early-career product manager to work at the frontier of AI infrastructure, building tools used daily by developers and AI teams at the world's most ambitious companies. Inspired by Google’s APM program, APMs at Fireworks will receive mentorship from experienced PMs while also getting rotations across both Fireworks’ product areas and tasks. This role is designed to help you develop the foundational skills of a start-up product leader: rigorous thinking, user empathy, technical depth, and cross-functional leadership.
Forward Deployed Engineer - LLM Post-training
Reflection ai
Reflection is a research lab dedicated to making intelligence open and accessible. We build open-weight models that empower users to control their AI and shape its future. As a core member of the Applied AI team, you will drive model fine-tuning and evaluations for enterprise customers. This role involves adapting our open-weight models for specific customer domains, tasks, and constraints, working hands-on with customer data, running fine-tuning workflows, building evaluation harnesses, and deploying adapted models to production. You will collaborate directly with customers to understand their needs and with research teams to advance the possibilities of AI.
Research Engineer, Data Infrastructure
Mistral AI
Mistral AI is seeking a Research Engineer focused on Data Infrastructure to build and operate the next generation of our data systems. This role involves designing and scaling massive compute fleets and storage systems for high performance and scalability. You will contribute to a future of decoupled control and data planes, scaling big data compute and storage platforms while ensuring secure and governed data access for MLOps and research. The position requires full lifecycle ownership, from architecting migrations away from legacy orchestrators to implementing production-grade pipelines and participating in on-call rotations for critical training jobs.
Member of Engineering (Post-training)
poolside
Poolside is building a company to create Artificial General Intelligence, aiming to accelerate software development through agentic systems, coding assistants, and frontier models. This role is part of the Applied Research team, focused on transforming pre-trained Large Language Models (LLMs) into well-aligned and highly capable AI systems specifically for coding and software development. You will be involved in building data pipelines and environments for agentic use cases, researching and implementing post-training algorithms, and designing experiments to test hypotheses, with access to significant GPU resources.