Fine-Tuning Jobs

175 open roles mentioning Fine-Tuning

Member of Technical Staff (AI Researcher)

19d ago
P

Perplexity AI

Perplexity is seeking top-tier AI Research Scientists and Engineers to advance our AI products and capabilities, focusing on building the future of AI-powered search and agent experiences. You will contribute to SOTA experiences that handle hundreds of millions of queries and continue to scale rapidly. Depending on your interests and expertise, you can join one of three specialized teams: the Core Research Team focusing on foundational models, the Agent Products Team fine-tuning models for agent and product experiences, or the Comet Agent Team dedicated to developing and enhancing the Comet Agent product.

San Francisco onsite FullTime
PythonC#Fine-Tuning +3 more

Forward Deployed Engineer EMEA

19d ago
R

Runpod

Runpod is seeking a Forward Deployed Engineer to join their Revenue Team and support the growth of their AI Developer Cloud platform. This role focuses on providing technical insights, resolving customer challenges, and building trust to ensure a seamless and delightful Runpod experience. You will work closely with customers, sales, product, engineering, and support teams to enhance customer satisfaction and drive product improvements. The ideal candidate combines technical expertise with strong communication skills to address technical challenges, guide onboarding, and collaborate across functions.

$100k - $180k

Remote - EMEA remote FullTime
DockerPythonJavaScript +5 more

Product Manager, Model APIs

20d ago
s

sarvam

Sarvam is building India's full-stack sovereign AI platform, focusing on research, models, infrastructure, and applications to make AI work for India. They partner with leading enterprises and public institutions, backed by prominent venture capital firms. This role is for a Product Manager for Model APIs, sitting at the intersection of research and the market. You will own the API surface, including its shape, pricing, packaging, and production readiness. Additionally, you will influence model direction by prioritizing capabilities, languages, and domains, and defining what constitutes a shippable product. This is an opportunity to directly impact the model roadmap, not just its interface.

Bengaluru onsite FullTime
PythonFine-TuningSQL

Research Scientist (Generative Modeling)

20d ago
W

World-labs

World Labs is a frontier AI research and product company focused on spatial intelligence, advancing beyond large language models. Co-founded by leading researchers, the company is pioneering world models that perceive, generate, reason, and interact with virtual and physical worlds. Their flagship product, Marble, transforms various media into navigable 3D worlds, with applications in gaming, film, architecture, robotics, and immersive experiences. Backed by significant investment, World Labs is building a world-class team at the intersection of AI research and real-world deployment.

$250k - $325k

San Francisco onsite
PythonFine-TuningPyTorch +3 more

Product Manager - Chanakya

21d ago
s

sarvam

Sarvam AI is seeking a Product Manager to join its specialized Chanakya vertical, focused on national security and institutions of national importance. This is a unique opportunity to be at the forefront of building a next-generation defense decision-support product, taking it from conception to launch with real users. You will be instrumental in shaping the product's direction, working closely with senior institutional users and navigating complex environments. This role offers unparalleled learning across the full stack, from understanding customer needs to customizing AI solutions and potentially fine-tuning models. It's a leadership seat with a clear growth path, offering the chance to build and lead emerging functions within a high-impact, zero-to-one journey backed by Sarvam's capabilities.

Delhi onsite FullTime
GoFine-Tuning

Associate General Counsel, Product (Data Products, Machine Learning & Artificial Intelligence)

23d ago
Abridge

Abridge

Abridge is seeking an Associate General Counsel, Product with a focus on Data Products, Machine Learning & Artificial Intelligence. This role will serve as the primary legal advisor for teams developing Abridge's AI and machine learning models, products, and technologies. You will be instrumental in guiding product teams on cutting-edge AI developments, from ambient clinical documentation to emerging agentic and multimodal healthcare experiences. This is a unique opportunity to shape the legal landscape at the intersection of AI and healthcare for a category-defining company.

SF Office hybrid FullTime
GoFine-Tuning

Research Engineer, Computer Use

24d ago
Anthropic

Anthropic

The Computer Use team focuses on teaching Claude to see, use, and understand computer interfaces. As a Research Engineer on the team, you'll work on advancing our models' ability to reliably and safely operate real software. We're looking for someone who's genuinely excited about both the research and the product sides of computer use. Your work will translate directly into model improvements in our own and our customers' products. You can try Claude's computer use capabilities today through the Claude in Chrome extension and Claude Cowork.

San Francisco, CA | New York City, NY | Seattle, WA onsite
AnthropicPythonFine-Tuning +2 more

Research Engineer, Domain Scaling

24d ago
Anthropic

Anthropic

The Domain Scaling team aims to make Claude world-class at real-world knowledge work in domains like finance, healthcare, and legal. This role combines direct applied research with data sourcing (real-world and synthetic) to improve our models. You will own the end-to-end process of creating RL environments for new capabilities, which includes identifying high-value tasks, designing reward signals, managing vendor relationships, and measuring impact on model performance.

San Francisco, CA | New York City, NY | Seattle, WA onsite
AnthropicFine-TuningClaude +1 more

Research Engineer, Code RL (Reinforcement Learning)

24d ago
Anthropic

Anthropic

We are seeking a Research Engineer for our Code RL team, focused on advancing AI models' capabilities in writing, editing, testing, debugging, and shipping real software. This role involves designing RL environments, coding tasks, and reward signals, as well as running training experiments on frontier models. You will diagnose model performance, improve pipeline speed and reliability, and contribute to areas like agentic coding behaviors, code correctness, and autonomous engineering. The position blends cutting-edge research with practical engineering to build high-quality, scalable AI systems.

San Francisco, CA | New York City, NY onsite
AnthropicPythonFine-Tuning +5 more

Research Scientist, Life Sciences

24d ago
Anthropic

Anthropic

Anthropic is seeking an exceptional Research Scientist to join its Life Sciences team. This role focuses on making Claude a superhuman life sciences research assistant, operating at the intersection of machine learning, software engineering, and biology. You will directly improve model capabilities on scientific tasks through post-training, evaluation design, and RL environment development. As a core member, you will translate deep biological domain knowledge into model training objectives, benchmarks, and agentic workflows, helping establish Anthropic as a leader in AI-accelerated biology and shaping how frontier models reason about computational biology tasks. This is a unique opportunity to shape how frontier AI models learn biology, working alongside top AI researchers on problems crucial for human health and scientific understanding.

San Francisco, CA onsite
AnthropicPythonFine-Tuning +3 more

Engineering Manager, Research Productivity

24d ago
Anthropic

Anthropic

Anthropic's Research Tools team builds systems that support large-scale, distributed finetuning runs and improve researcher productivity. As a manager, you will lead a team of machine learning and distributed systems experts to enhance the efficiency of these systems and tools, facilitate rapid iteration on model development and research, and continuously evolve the infrastructure to integrate new research advancements. This role is central to Anthropic's technical operations, requiring collaboration with research teams to integrate their innovations into the production finetuning pipeline, product teams for customer-oriented model improvements, and infrastructure teams to optimize training runs and data pipelines.

San Francisco, CA | New York City, NY onsite
AnthropicFine-Tuning

Full-Stack Software Engineer, Reinforcement Learning

24d ago
Anthropic

Anthropic

As a Full-Stack Software Engineer in Reinforcement Learning (RL), you will be instrumental in building the platforms, tools, and interfaces essential for environment creation, data collection, and training observability. Your work will directly impact the quality of data used to train Anthropic's next-generation AI models. You will own product surfaces from end-to-end, encompassing backend services, APIs, and web UIs used by researchers, external vendors, and data labelers. The role emphasizes shipping polished, reliable products quickly, even when faced with ambiguous, high-stakes problems. This team operates at a rapid pace, focusing on judgment and taste to meet researcher needs, iterating on data collection strategies to distill expert knowledge into models within short feedback loops.

San Francisco, CA | New York City, NY onsite
AnthropicAWSDocker +5 more

Research Engineer, Visual Knowledge Work

24d ago
Anthropic

Anthropic

We are seeking research engineers with a strong computer vision background to enhance the visual and spatial reasoning capabilities of our state-of-the-art Claude models. This role involves research, development, and evaluation, taking a full-stack approach across pretraining, RL, and runtime techniques. You will collaborate closely with the product organization to ensure that vision improvements directly impact Claude's performance on real-world tasks and address customer challenges.

New York City, NY; San Francisco, CA; Seattle, WA onsite
AnthropicFine-TuningClaude +3 more

Research Scientist/Engineer, Biological Safety

24d ago
Anthropic

Anthropic

Anthropic is building reliable, interpretable, and steerable AI systems to be safe and beneficial for users and society. As a Safeguards Biological Safety Research Scientist, you will apply your technical skills to design and develop safety systems that detect harmful AI behaviors and prevent misuse by sophisticated threat actors. This role is at the forefront of defining responsible AI safety in the biological domain, translating complex biosecurity concepts into technical safeguards and balancing AI's potential for life sciences research with preventing misuse.

San Francisco, CA onsite
AnthropicPythonFine-Tuning +1 more

Research Engineer, Universes

24d ago
Anthropic

Anthropic

The Universes team within Research is responsible for training AI models to perform complex, difficult, long-horizon agentic tasks in ultra-realistic settings. We design and implement novel training environments that go far beyond what models can do today — environments where models learn to navigate ambiguity, handle interruptions, maintain context over extended interactions, and exercise judgment in open-ended scenarios. We're looking for Research Engineers to help us build the next generation of training environments for capable and safe agentic AI. This role blends research and engineering responsibilities, requiring you to both implement novel approaches and contribute to research direction. You'll work on fundamental research in reinforcement learning, designing training environments and methodologies that push the state of the art, and building evaluations that measure genuine capability.

Remote-Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NY remote
AnthropicGoFine-Tuning +1 more

Research Engineer, RL Engineering

24d ago
Anthropic

Anthropic

Anthropic's mission is to create reliable, interpretable, and steerable AI systems that are safe and beneficial for users and society. As an ML Systems Engineer on the Reinforcement Learning Engineering team, you will build and improve the critical algorithms and infrastructure that researchers use to train AI models like Claude. Your work will directly enable breakthroughs in AI capabilities and safety, focusing on enhancing the performance, robustness, and usability of these systems to accelerate research progress. You will support and empower the research team in their mission to build beneficial AI systems, specifically by building, maintaining, and improving the algorithms and systems used for finetuning production and research models with methods like RLHF.

San Francisco, CA | New York City, NY | Seattle, WA onsite
AnthropicPythonFine-Tuning +3 more

Research Engineer / Scientist, Alignment

24d ago
Anthropic

Anthropic

Anthropic is seeking a Research Engineer/Scientist for its Alignment Science team. This role involves designing and executing machine learning experiments to understand and steer the behavior of advanced AI systems, with a focus on AI safety and potential risks from future human-level AI. You will collaborate with other teams on exploratory research, contributing to Anthropic's mission of creating reliable, interpretable, and steerable AI systems that are helpful, honest, and harmless.

San Francisco, CA onsite
AnthropicKubernetesPython +4 more

Research Engineer, Production Model Post-Training

24d ago
Anthropic

Anthropic

Anthropic's production models undergo sophisticated post-training processes to enhance their capabilities, alignment, and safety. As a Research Engineer on our Post-Training team, you'll train our base models through the complete post-training stack to deliver the production Claude models that users interact with. You'll work at the intersection of cutting-edge research and production engineering, implementing, scaling, and improving post-training techniques like Constitutional AI, RLHF, and other alignment methodologies. Your work will directly impact the quality, safety, and capabilities of our production models. For this role, interviews are conducted in Python, and the position may require responding to incidents on short notice, including weekends.

San Francisco, CA | New York City, NY | Seattle, WA onsite
AnthropicPythonFine-Tuning +3 more

Research Engineer, Knowledge Team

24d ago
Anthropic

Anthropic

Anthropic is seeking Research Engineers to reimagine how Claude interacts with external data sources. This role involves designing novel architectures for organizing information and training language models to effectively utilize these architectures. The goal is to move beyond traditional data paradigms to accommodate the capabilities of Large Language Models (LLMs).

Remote-Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NY remote
AnthropicPythonRAG +3 more

Senior Product Counsel - Data Products, Machine Learning & Artificial Intelligence

24d ago
Abridge

Abridge

Abridge is seeking a Senior Product Counsel specializing in Data Products, Machine Learning, and Artificial Intelligence. This role will serve as the primary legal advisor for teams developing Abridge's AI and machine learning models, products, and technologies. You will be at the forefront of legal, AI, and healthcare, advising product teams on cutting-edge AI developments, from ambient clinical documentation to emerging agentic and multimodal healthcare experiences. This is a unique opportunity to shape how a category-defining company navigates novel legal terrain in a fast-paced, high-growth startup environment.

SF Office hybrid FullTime
GoFine-Tuning

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.