Fine-Tuning Jobs
107 open roles mentioning Fine-Tuning
Anthropic Fellows Program
Anthropic
Anthropic is seeking candidates for its Fellows Program, designed to cultivate AI research and engineering talent. This program provides funding and mentorship to promising individuals, irrespective of prior experience, to work on empirical projects aligned with Anthropic's research priorities. The goal is to produce public outputs, such as paper submissions, with a strong track record of fellows achieving this in previous cohorts. The program offers a 4-month full-time research period, direct mentorship from Anthropic researchers, access to a shared workspace in Berkeley or London, and connections to the AI safety and security research community. Applications are reviewed on a rolling basis for cohorts starting in July 2026 and beyond.
$25k - $50k
Anthropic Fellows Program, ML Systems & Performance
Anthropic
Anthropic is seeking talented individuals for its Fellows Program, focusing on AI research and engineering to ensure AI systems are reliable, interpretable, and steerable. This program provides funding and mentorship to promising technical talent, regardless of prior experience, to work on empirical projects aligned with Anthropic's research priorities. The goal is to produce public outputs, such as paper submissions, contributing to the development of safe and beneficial AI for society. Fellows will engage in full-time research for four months, with direct mentorship from Anthropic researchers and access to a shared workspace in either Berkeley, California, or London, UK. The program also offers connections to the broader AI safety and security research community.
$25k - $50k
Anthropic Fellows Program, Reinforcement Learning
Anthropic
Anthropic is seeking talented individuals for its Fellows Program, focusing on AI research and engineering. This program provides funding and mentorship to promising technical talent, regardless of prior experience, to work on empirical projects aligned with Anthropic's research priorities. The goal is to produce public outputs, such as research papers, contributing to the development of reliable, interpretable, and steerable AI systems that are safe and beneficial for society. Fellows will engage in full-time research for four months, with opportunities for extension, and will be mentored by Anthropic researchers.
$25k - $50k
Engineering Manager, Research Productivity
Anthropic
Anthropic's Research Tools team builds systems that support large-scale, distributed finetuning runs and improve researcher productivity. As a manager, you will lead a team of machine learning and distributed systems experts to enhance the efficiency of these systems and tools, facilitate rapid iteration on model development and research, and continuously evolve the infrastructure to integrate new research advancements. This role is central to Anthropic's technical operations, requiring collaboration with research teams to integrate their innovations into the production finetuning pipeline, product teams for customer-oriented model improvements, and infrastructure teams to optimize training runs and data pipelines.
Full-Stack Software Engineer, Reinforcement Learning
Anthropic
As a Full-Stack Software Engineer in Reinforcement Learning (RL), you will be instrumental in building the platforms, tools, and interfaces essential for environment creation, data collection, and training observability. Your work will directly impact the quality of data used to train Anthropic's next-generation AI models. You will own product surfaces from end-to-end, encompassing backend services, APIs, and web UIs used by researchers, external vendors, and data labelers. The role emphasizes shipping polished, reliable products quickly, even when faced with ambiguous, high-stakes problems. This team operates at a rapid pace, focusing on judgment and taste to meet researcher needs, iterating on data collection strategies to distill expert knowledge into models within short feedback loops.
Product Manager, Compute Platform
Anthropic
Anthropic is building reliable, interpretable, and steerable AI systems to be safe and beneficial for users and society. As a Product Manager focused on Compute Platform, you will partner with Infrastructure, Compute Operations, Engineering, Finance & Strategy, and Research teams. Your role will be to build the scheduling, orchestration, and capacity management systems that power Anthropic’s compute infrastructure, which is essential for all model training, evaluation, and inference workloads.
Research Engineer, Domain Scaling
Anthropic
The Domain Scaling team aims to make Claude world-class at real-world knowledge work in domains like finance, healthcare, and legal. This role combines direct applied research with data sourcing (real-world and synthetic) to improve our models. You will own the end-to-end process of creating RL environments for new capabilities, which includes identifying high-value tasks, designing reward signals, managing vendor relationships, and measuring impact on model performance.
Research Engineer, Computer Use
Anthropic
The Computer Use team focuses on teaching Claude to see, use, and understand computer interfaces. As a Research Engineer on the team, you'll work on advancing our models' ability to reliably and safely operate real software. We're looking for someone who's genuinely excited about both the research and the product sides of computer use. Your work will translate directly into model improvements in our own and our customers' products. You can try Claude's computer use capabilities today through the Claude in Chrome extension and Claude Cowork.
Software Engineer, Platform, Tinker
thinkingmachines
Thinking Machines Lab is seeking a Software Engineer to own the platform systems that power Tinker, their fine-tuning API. This role involves developing and maintaining critical components such as billing and usage metering, permissions and access control, organization and team management, data export functionalities, and audit logging. You will collaborate closely with product, legal, and other cross-functional teams, ensuring that new features, pricing adjustments, and enterprise deals are seamlessly integrated into the platform. This is an opportunity to grow a rapidly expanding platform and contribute to the Tinker community.
$350k - $475k
Engineering Manager, UAE
Scale AI
Scale's Global Public Sector team is seeking a technically strong and client-oriented Engineering Manager to lead the development of AI applications for UAE government agencies. This role is based in the UAE and places you at the center of the nation's ambitious AI-first initiatives. You will be responsible for engineering delivery, embedding deeply with clients, and building Scale's presence as a leading AI partner in the region. The ideal candidate is comfortable in both government stakeholder meetings and technical design reviews, driving AI solutions from concept to production.
Applied AI, Machine Learning Engineer
Mistral AI
Mistral AI is seeking an Applied AI Engineer to join their team and facilitate the adoption of their AI products among customers. This role involves collaborating with clients to address complex technical challenges, from pre-sale to post-implementation, ensuring Mistral's solutions meet and exceed expectations. The engineer will manage daily customer relations, act as a key resource in externalizing research into production settings, and work on state-of-the-art Generative AI applications across various industries. This position offers the opportunity to contribute to a pioneering company shaping the future of AI and make a meaningful impact.
Senior/Staff Applied AI, Machine Learning Engineer
Mistral AI
Mistral AI is seeking an Applied AI Engineer to join our Applied AI Engineering team. This role is integral to driving the successful deployment of Mistral AI products and building complex enterprise use cases. You will work hand-in-hand with customers from pre-sale to post-implementation, ensuring our solutions meet and exceed client expectations. You will manage customer relations involving multiple stakeholders and function as a key resource in externalizing our research in production settings, facilitating the adoption of our products and collaborating to address complex technical challenges.
Engineering Manager, Saudi Arabia
Scale AI
Scale AI is seeking a technically strong and customer-oriented Engineering Manager to lead the development of AI applications for Saudi government clients. This role is pivotal in shaping Scale AI's presence in the Gulf region, focusing on ambitious public sector AI initiatives. The ideal candidate will be comfortable with engineering leadership, AI/ML delivery, and direct client engagement, bridging the gap between technical complexities and government stakeholders. This is an opportunity to be a founding member of a rapidly growing team, driving the transformation of government operations through cutting-edge AI technology in a fast-paced, high-impact environment.
Deployment Strategist
Scale AI
Scale is seeking a Defense Deployment Strategist to lead the end-to-end effort of deploying and scaling Agentic AI capabilities within the National Reconnaissance Office (NRO) and the National Geospatial-Intelligence Agency (NGA). This role is crucial for integrating AI into critical intelligence workflows, providing a strategic edge to the U.S. and its allies. You will be the primary point of contact at the intersection of AI and national intelligence, responsible for capturing opportunities, driving deployment, and fostering growth within these highly technical and mission-critical agencies. This is a hands-on, individual contributor role with significant leadership expectations, acting as the 'CEO of the Account.'
Software Engineer, Identity
Scale AI
Scale is seeking a Software Engineer, Identity to join our Platform Engineering team. In this role, you will be instrumental in designing and developing core platforms and software systems, with a specific focus on identity, access management, authorization, and authentication. You will gain broad exposure to the cutting edge of the AI industry as Scale supports enterprises, startups, and governments. This position offers the opportunity to contribute to the foundational elements of products that power advanced LLMs and generative models, playing a crucial role in how humanity interacts with AI.
$216k - $270k
AI research scientist
Writer
AI research at WRITER focuses on building the scientific foundation for ambitious enterprise AI deployments. As a staff AI research scientist, you will drive a high-impact research agenda centered on large language models, agentic reasoning, and system-level capabilities essential for enterprise-scale AI. This role offers a unique opportunity to advance the field while directly contributing to products used by hundreds of thousands daily. You will work on post-training, planning, multi-step reasoning, and agentic workflows, directly shaping the future of enterprise AI performance and scalability. The role provides resources, infrastructure, and cross-functional support to pursue and implement ambitious ideas rapidly.
Senior Machine Learning Engineer, Voice AI
Together AI
Together AI is building the best inference infrastructure for voice applications, powering production-grade, real-time voice agents and applications. We are seeking a Senior ML Engineer to lead the model serving layer for voice workloads. This role involves hands-on optimization of inference engines and models like Whisper, Parakeet, Orpheus, and Kokoro to achieve frontier-level latency and throughput. You will focus on unique voice inference challenges such as streaming audio, tokenization, and real-time latency budgets, shaping how voice models are served as the industry shifts towards end-to-end speech-to-speech systems. This is a foundational hire on a small, high-impact team.
$200k - $260k
Staff Machine Learning Engineer, Voice AI
Together AI
Together AI is building the best inference infrastructure for voice applications, powering production-grade, real-time voice agents and applications. We are seeking a Staff ML Engineer to lead the model serving layer for voice workloads. This role involves hands-on optimization of inference engines and models like Whisper, Parakeet, Orpheus, and Kokoro, focusing on pushing latency and throughput boundaries. You will address unique challenges in voice inference, such as streaming audio and real-time latency, and shape the future of how voice models are served as the industry shifts towards end-to-end speech-to-speech systems. This is a foundational hire on a small, high-impact team.
$220k - $280k
Research Engineer, Frontier Speculative Decoding
Together AI
Together AI is building the Inference Platform that powers the world's most advanced generative AI models. This role will serve as a critical bridge between cutting-edge research and real-world applications, focusing on translating internal model training research into production-ready deployments for customers. The work involves a deep commitment to data-centric development, meticulous hyperparameter tuning, and rigorous checkpoint evaluation. You will transform general-purpose models into highly performant, specialized tools by fine-tuning them on customer-specific data and internal datasets, working with dedicated GPU clusters rather than training foundation models from scratch.
$190k - $270k
Machine Learning Engineer
Together AI
Together AI is seeking an ML Engineer to develop systems and APIs for customer inference and fine-tuning of LLMs. The ideal candidate will have experience implementing runtime systems for large-scale AI/ML model inference, including the largest LLMs. This role involves designing and building production systems for reliability and performance at scale, partnering with cross-functional teams, and improving system efficiency and stability.
$160k - $220k