Fine-Tuning Jobs

107 open roles mentioning Fine-Tuning

Research Scientist, Takeoff Intel

3d ago
Anthropic

Anthropic

Anthropic is seeking a Research Scientist to focus on measuring and understanding recursive self-improvement in large models. The ideal candidate has hands-on research experience with large models, including pretraining, fine-tuning, reinforcement learning, evaluations, or agent scaffolds. You will leverage your understanding of the model development loop to identify key signals, design evaluations and models to measure them, and interpret the results to understand the pace of AI advancement. This role involves making research bets, owning outcomes, and communicating findings through written assessments for internal decision-makers and public reporting. Collaboration with various research and policy teams is expected.

San Francisco, CA onsite
AnthropicFine-TuningClaude

Research Engineer, Agents

6d ago
Anthropic

Anthropic

Agentic systems are a rapidly growing area in AI deployment, with applications in coding, research, customer support, and network security. Anthropic is building a team focused on enhancing Claude's capabilities as an agent, enabling it to handle more complex, long-horizon tasks and coordinate with other agents. This role involves tackling challenges in harness design, agent affordances, infrastructure, and fine-tuning to maximize agent performance. The team is seeking individuals who can demonstrate their skill in getting LLMs to perform complex tasks through projects involving agent design, quantitative experiments, benchmarking, synthetic data generation, or model fine-tuning.

Remote-Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NY remote
AnthropicFine-TuningClaude

Research Scientist/Engineer, Biological Safety

9d ago
Anthropic

Anthropic

Anthropic is building reliable, interpretable, and steerable AI systems to be safe and beneficial for users and society. As a Safeguards Biological Safety Research Scientist, you will apply your technical skills to design and develop safety systems that detect harmful AI behaviors and prevent misuse by sophisticated threat actors. This role is at the forefront of defining responsible AI safety in the biological domain, translating complex biosecurity concepts into technical safeguards and balancing AI's potential for life sciences research with preventing misuse.

San Francisco, CA onsite
AnthropicPythonFine-Tuning +1 more

Technical Program Manager, AI Delivery for Public Sector & Defence, UK

9d ago
Cohere

Cohere

Cohere is seeking an experienced Program Manager to join our customer-facing Program Management team, focusing on UK public sector and defence accounts. This role requires a deep understanding of technical program delivery, the complexities of frontier AI models, and the unique needs of government and defence organizations. You will act as the crucial link between Cohere's advanced AI capabilities and the specific requirements of UK Government and defence clients, navigating intricate procurement processes, stringent security demands, data sovereignty concerns, and regulatory compliance to ensure successful deployment and scaling of our AI solutions. You will collaborate with various internal teams, serving as the technical program lead and trusted advisor for our most sensitive and high-impact government engagements.

$12k - $18k

London hybrid FullTime
CohereAWSAzure +3 more

Technical Program Manager, AI Delivery for Public Sector & Defense, France

10d ago
Cohere

Cohere

Cohere is seeking an experienced Program Manager to join its customer-facing Program Management team, focusing on public sector and defense accounts in Europe, particularly in France. This specialized role requires an individual who understands technical program delivery, the complexity of frontier AI models, and the unique requirements of government and defense organizations. You will act as the critical bridge between Cohere's AI capabilities and the needs of Ministries, Agencies, and defense organizations across Europe, navigating complex procurement processes, security requirements, data sovereignty, and regulatory compliance to ensure successful deployment and scaling of AI solutions. You will collaborate with various internal teams, serving as the technical program lead and trusted advisor for sensitive government and defense engagements.

France onsite FullTime
CohereAWSAzure +3 more

Software Engineer, Full Stack, Tinker

13d ago
thinkingmachines

thinkingmachines

Thinking Machines Lab is seeking a full-stack engineer to develop and deploy the products and services that Tinker users engage with daily. This role involves working across frontend, backend, and infrastructure to build the Tinker console, developer tools, and other essential components for the platform. Tinker is a fine-tuning API that enables researchers and developers to customize frontier AI models using their own data and algorithms, managing the underlying infrastructure to provide flexibility and access to advanced capabilities.

$350k - $475k

San Francisco onsite
OpenAIMistralPython +5 more

Site Reliability Engineer (SRE)

13d ago
thinkingmachines

thinkingmachines

Thinking Machines Lab is seeking a Site Reliability Engineer (SRE) to ensure the end-to-end reliability of their Tinker platform. This role involves working closely with engineers and research teams to enhance the robustness and resilience of every system layer. The SRE will be instrumental in maintaining and improving the infrastructure that supports custom AI model fine-tuning, ensuring a seamless experience for researchers and developers.

$350k - $475k

San Francisco onsite
OpenAIMistralKubernetes +3 more

Research Engineer, Developer Experience, Tinker

15d ago
thinkingmachines

thinkingmachines

Thinking Machines Lab is seeking a Research Engineer focused on developer experience to build and enhance their Tinker platform. This role involves working hands-on with users to understand their challenges and translate them into product improvements. You will be responsible for creating and updating documentation, adding library features, prototyping integrations, and ensuring users can smoothly customize frontier AI models. This position acts as a crucial link between Tinker users and the internal research and infrastructure teams, surfacing user patterns to inform product and infrastructure priorities and sharing learnings through various channels.

$350k - $475k

San Francisco onsite
OpenAIMistralFine-Tuning +2 more

Research Engineer, Visual Knowledge Work

16d ago
Anthropic

Anthropic

We are seeking research engineers with a strong computer vision background to enhance the visual and spatial reasoning capabilities of our state-of-the-art Claude models. This role involves research, development, and evaluation, taking a full-stack approach across pretraining, RL, and runtime techniques. You will collaborate closely with the product organization to ensure that vision improvements directly impact Claude's performance on real-world tasks and address customer challenges.

New York City, NY; San Francisco, CA; Seattle, WA onsite
AnthropicFine-TuningClaude +3 more

Research Product Manager, Model Behaviors

16d ago
Anthropic

Anthropic

Anthropic's mission is to create reliable, interpretable, and steerable AI systems that are safe and beneficial for users and society. As a Product Manager for Model Behaviors, you will partner with the Alignment Finetuning team to define and shape Claude's character, behaviors, and reinforcement signals. This role directly influences how millions of people experience AI by systematically identifying high-priority behavioral improvements, coordinating across Research, Product, and Safeguards teams, and accelerating the ability to ship well-aligned models. The ideal candidate possesses deep user empathy and the judgment to navigate nuanced behavior questions.

San Francisco, CA | New York City, NY onsite
AnthropicGoFine-Tuning +1 more

Head of ANZ, Applied AI

16d ago
Anthropic

Anthropic

As the founding leader of Applied AI Solutions Architecture in ANZ, you will drive the adoption of frontier AI by enabling the deployment of Anthropic's products across Australian and New Zealand enterprises. You'll leverage your technical skills and consultative sales experience to drive positive AI transformation that addresses customer business needs and technical requirements, ensuring high reliability and safety. You will be responsible for building and leading the ANZ Applied AI team, establishing processes and best practices for technical pre-sales engagements, and representing Anthropic as a technical lead on important partnerships. In collaboration with Sales, Product, and Engineering teams, you'll help enterprise partners incorporate leading-edge AI systems into their products and platforms, using excellent communication skills to explain complex solutions to diverse audiences and identifying opportunities to innovate while maintaining safety standards.

Sydney, Australia onsite
AnthropicGoPrompt Engineering +2 more

Research Engineer, Code RL (Reinforcement Learning)

16d ago
Anthropic

Anthropic

We are seeking a Research Engineer for our Code RL team, focused on advancing AI models' capabilities in writing, editing, testing, debugging, and shipping real software. This role involves designing RL environments, coding tasks, and reward signals, as well as running training experiments on frontier models. You will diagnose model performance, improve pipeline speed and reliability, and contribute to areas like agentic coding behaviors, code correctness, and autonomous engineering. The position blends cutting-edge research with practical engineering to build high-quality, scalable AI systems.

San Francisco, CA | New York City, NY onsite
AnthropicPythonFine-Tuning +5 more

Research Engineer, Knowledge Foundations

16d ago
Anthropic

Anthropic

The Knowledge Work team at Anthropic builds the training environments and evaluations that empower Claude to excel in professional workflows, including searching, analyzing, and creating content across various tools and documents. As this work expands, the underlying systems require the same rigor as the research itself. In this role, you will design and execute experiments to enhance Claude's ability to search, retrieve, and reason over information at scale. Your responsibilities will encompass environment design, data curation, RL training, evaluation, and supporting infrastructure, requiring flexibility to address progress blockers. You will collaborate closely with researchers and other RL teams to implement capabilities that directly influence Claude's performance. We believe that true ownership and impact stem from both hardening existing environments and creating new ones, ensuring the quality of the entire stack that drives superhuman epistemics.

San Francisco, CA onsite
AnthropicPythonFine-Tuning +2 more

Research Engineer, Knowledge Team

16d ago
Anthropic

Anthropic

Anthropic is seeking Research Engineers to reimagine how Claude interacts with external data sources. This role involves designing novel architectures for organizing information and training language models to effectively utilize these architectures. The goal is to move beyond traditional data paradigms to accommodate the capabilities of Large Language Models (LLMs).

Remote-Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NY remote
AnthropicPythonRAG +3 more

Research Engineer, Production Model Post-Training

16d ago
Anthropic

Anthropic

Anthropic's production models undergo sophisticated post-training processes to enhance their capabilities, alignment, and safety. As a Research Engineer on our Post-Training team, you'll train our base models through the complete post-training stack to deliver the production Claude models that users interact with. You'll work at the intersection of cutting-edge research and production engineering, implementing, scaling, and improving post-training techniques like Constitutional AI, RLHF, and other alignment methodologies. Your work will directly impact the quality, safety, and capabilities of our production models. For this role, interviews are conducted in Python, and the position may require responding to incidents on short notice, including weekends.

San Francisco, CA | New York City, NY | Seattle, WA onsite
AnthropicPythonFine-Tuning +3 more

Research Engineer, Production Model Post-Training

16d ago
Anthropic

Anthropic

Anthropic's production models undergo sophisticated post-training processes to enhance their capabilities, alignment, and safety. As a Research Engineer on our Post-Training team, you'll train our base models through the complete post-training stack to deliver the production Claude models that users interact with. You'll work at the intersection of cutting-edge research and production engineering, implementing, scaling, and improving post-training techniques like Constitutional AI, RLHF, and other alignment methodologies. Your work will directly impact the quality, safety, and capabilities of our production models. Note: For this role, we conduct all interviews in Python. This role may require responding to incidents on short-notice, including on weekends.

Zürich, CH onsite
AnthropicPythonFine-Tuning +3 more

Research Engineer / Scientist, Alignment

16d ago
Anthropic

Anthropic

Anthropic is seeking a Research Engineer/Scientist for its Alignment Science team. This role involves designing and executing machine learning experiments to understand and steer the behavior of advanced AI systems, with a focus on AI safety and potential risks from future human-level AI. You will collaborate with other teams on exploratory research, contributing to Anthropic's mission of creating reliable, interpretable, and steerable AI systems that are helpful, honest, and harmless.

San Francisco, CA onsite
AnthropicKubernetesPython +4 more

[Expression of Interest] Research Engineer / Scientist, Alignment - London

16d ago
Anthropic

Anthropic

Anthropic is building reliable, interpretable, and steerable AI systems to ensure AI is safe and beneficial for society. As a Research Engineer on the Alignment Science team in London, you will design and execute machine learning experiments to understand and steer the behavior of advanced AI systems. You will focus on AI safety, particularly risks from future human-level AI systems, collaborating with teams like Interpretability and Frontier Red Team. The role involves exploratory research in areas such as AI Control and Alignment Stress-testing, aiming to make AI helpful, honest, and harmless.

London, UK onsite
AnthropicKubernetesPython +3 more

Research Engineer, Universes

16d ago
Anthropic

Anthropic

The Universes team within Research is responsible for training AI models to perform complex, difficult, long-horizon agentic tasks in ultra-realistic settings. We design and implement novel training environments that go far beyond what models can do today — environments where models learn to navigate ambiguity, handle interruptions, maintain context over extended interactions, and exercise judgment in open-ended scenarios. We're looking for Research Engineers to help us build the next generation of training environments for capable and safe agentic AI. This role blends research and engineering responsibilities, requiring you to both implement novel approaches and contribute to research direction. You'll work on fundamental research in reinforcement learning, designing training environments and methodologies that push the state of the art, and building evaluations that measure genuine capability.

Remote-Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NY remote
AnthropicGoFine-Tuning +1 more

Research Scientist, Life Sciences

16d ago
Anthropic

Anthropic

Anthropic is seeking an exceptional Research Scientist to join its Life Sciences team. This role focuses on making Claude a superhuman life sciences research assistant, operating at the intersection of machine learning, software engineering, and biology. You will directly improve model capabilities on scientific tasks through post-training, evaluation design, and RL environment development. As a core member, you will translate deep biological domain knowledge into model training objectives, benchmarks, and agentic workflows, helping establish Anthropic as a leader in AI-accelerated biology and shaping how frontier models reason about computational biology tasks. This is a unique opportunity to shape how frontier AI models learn biology, working alongside top AI researchers on problems crucial for human health and scientific understanding.

San Francisco, CA onsite
AnthropicPythonFine-Tuning +3 more

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.