Fine-Tuning Jobs
175 open roles mentioning Fine-Tuning
Member of Technical Staff - VLM
Black Forest Labs
Black Forest Labs is seeking a Staff / Senior Individual Contributor to pioneer the integration of vision-language models (VLMs) directly into their advanced generative AI systems, like FLUX. This role focuses on developing novel VLM approaches and architectures, rather than just applying existing ones. You will explore how vision and language representations inform each other, how multimodal understanding enhances generation quality, and how to deploy these capabilities at scale. The ideal candidate will have a proven track record of pretraining or significantly advancing a VLM that has been deployed or publicly released.
€130k - €340k
Member of Technical Staff - Post Training
Black Forest Labs
We are seeking a Member of Technical Staff specializing in Post Training to own the end-to-end post-training pipeline for our multimodal generative models. This role is crucial for transforming foundation models into polished products, encompassing data strategy, reward modeling, preference optimization, distillation, and safety tuning across image, editing, and video modalities. You will be instrumental in driving significant improvements in model quality, developing the infrastructure that accelerates research team iteration, and advancing the state-of-the-art in aligning generative models with human intent. This is a Staff/Senior individual contributor position for someone with proven experience shipping post-training for a frontier model.
€130k - €340k
ML Ops Engineer, Chanakya
sarvam
Sarvam is building India's full-stack sovereign AI platform, focusing on research, models, infrastructure, and applications to make AI work for India. The MLOps Engineer will own the model lifecycle across all deployments, ensuring systems are always operational, accurate, and auditable. This role involves supporting field engineers and managing deployment infrastructure for new products, with an uncompromising standard for reliability, as model failures are considered operational risks.
Product Manager (Models)
sarvam
Sarvam is building India's full-stack sovereign AI platform, focusing on research, models, infrastructure, and applications to make AI work for India. We partner with leading enterprises and public institutions. We are seeking a Model Product Manager to lead initiatives for Automatic Speech Recognition (ASR), Text-to-Speech (TTS), and model evaluation pipelines. In this role, you will collaborate closely with research and engineering teams to define product roadmaps, manage training and evaluation processes, and deliver models that enhance product experience and reliability.
Data Scientist - Evaluations, Chanakya
sarvam
Sarvam is building India's full-stack sovereign AI platform, focusing on making AI genuinely work for India across research, models, infrastructure, and applications. This intellectually demanding role anchors the evaluations function for the AI vertical. You will design, build, and maintain evaluation frameworks to measure model and system quality in operational contexts, focusing on domain-specific, high-stakes use cases where accuracy is critical. You will collaborate closely with MLOps Engineers, Product Managers, and the deployment team to ensure deployed AI systems meet stringent quality standards.
Applied AI, Forward Deployed Machine Learning Engineer - Palo Alto
Mistral AI
Mistral AI is seeking an Applied AI Engineer to join their team and facilitate the adoption of their products among customers. This role involves collaborating closely with clients from pre-sale to post-implementation, ensuring their solutions meet and exceed expectations. You will manage daily customer relations, acting as a key resource for externalizing research into production settings and driving the successful deployment of Mistral AI products. The position offers the opportunity to work on state-of-the-art Generative AI applications across various industries and contribute to a pioneering company shaping the future of AI.
Research Engineer, Data Infrastructure
Mistral AI
Mistral AI is seeking a Research Engineer focused on Data Infrastructure to architect and build the backbone of our frontier model training and fine-tuning ecosystem. This role involves designing and scaling massive compute fleets and storage systems for high performance and scalability, with a vision towards exabyte-scale architecture. You will contribute to a strategic transition from legacy scheduling to modern orchestration, implementing sophisticated multi-cluster orchestration and cloud-bursting capabilities. The position requires full lifecycle ownership, from architecting migrations away from legacy orchestrators to implementing production-grade pipelines and participating in on-call rotations for critical training jobs.
Applied AI Engineer, Codex Core Agent
OpenAI
We are seeking applied AI engineers to transform Codex agents from impressive demonstrations into dependable tools. This role focuses on enhancing agent performance for real-world software engineering tasks, bridging the gap between research capabilities and practical usefulness. You will collaborate closely with research, infrastructure, and product teams to ensure agents are powerful, steerable, and reliable. The objective is to achieve measurable improvements in solve rates, usefulness, and economic value for users by refining model behavior and system integration.
Applied AI, Forward Deployed Machine Learning Engineer - Montreal
Mistral AI
Mistral AI is seeking an Applied AI Engineer to join their team and facilitate customer adoption of their AI products. This role involves collaborating closely with customers to address complex technical challenges, from pre-sale to post-implementation, ensuring solutions meet and exceed client expectations. You will manage daily customer relations, act as a key resource for externalizing research into production, and work on state-of-the-art Generative AI applications across various industries. This position offers the opportunity to contribute to a pioneering company shaping the future of AI and make a meaningful impact.
Senior Staff Software Engineer, AI Model Lifecycle
crusoe
Crusoe is seeking a Senior Staff Software Engineer for the AI Model Lifecycle team to build a comprehensive managed platform for the entire application development lifecycle, with a specific focus on leveraging Machine Learning models, including Large Language Models (LLMs). This role is crucial in accelerating the abundance of energy and intelligence by powering the world's most ambitious AI workloads. You will join a team building the future of AI infrastructure, solving the bottleneck of power for AI compute with an energy-first approach. We are looking for problem-solving, opportunity-finding teammates with a sense of urgency who thrive on building the path forward.
$238k - $318k
Senior Software Engineer, AI Model Lifecycle
crusoe
Crusoe is seeking a Senior Software Engineer for their AI Model Lifecycle team to build a comprehensive managed platform for the entire application development lifecycle, with a specific focus on leveraging Machine Learning models, including Large Language Models (LLMs). This role is crucial in accelerating the abundance of energy and intelligence by powering ambitious AI workloads. The ideal candidate will be a problem-solving, opportunity-finding teammate with a sense of urgency, who believes in the scale of Crusoe's ambition and thrives on a path not fully paved.
$172k - $231k
Solutions Architect - Public Sector
Cohere
Cohere is seeking a Solutions Architect to play a significant role in growing the company's Defence and National Security business. This dynamic role requires a blend of strategic thinking and hands-on execution, focusing on building customer demos and proof-of-concepts that highlight the business value of Cohere's AI platform. As the technical relationship owner, you will collaborate with stakeholders to understand their objectives and translate them into technical solutions. You will also serve as the voice of the customer, liaising between clients and the product team, providing guidance on best practices, identifying platform improvements, and cultivating technical champions within customer organizations to drive adoption and gather feedback.
Machine Learning Engineer, Integrity
OpenAI
As a Machine Learning Engineer in OpenAI's Integrity team, you will work on state-of-the-art models and classifiers, experiment with new architectures and approaches, and advance our capabilities in content and user understanding. This role offers the chance to transform research breakthroughs into tangible solutions that enhance the trust and safety of our platform, with a focus on training LLMs and building ML models. You will be instrumental in designing and deploying advanced machine learning models to solve real-world problems, bringing OpenAI's research from concept to implementation and creating AI-driven applications with direct impact.
Research Scientist, Multimodal Alignment, Safety, and Fairness
Google DeepMind
Google DeepMind's Frontier AI unit is seeking experienced Research Scientists to join a multimodal safety research effort. This role focuses on interdisciplinary sociotechnical modeling and requires a passion for understanding AI-society interactions, a strong awareness of AI alignment and safety, and a drive to develop novel ideas, methods, interfaces, and tools. You will contribute to advancing the state of the art in AI research and Google DeepMind's mission towards Artificial General Intelligence (AGI), with a focus on leading new breakthrough research directions in areas like AI behavior exploration, assessment, and steering, particularly for subjective and creative tasks. The work involves tackling fundamental research questions to improve alignment objectives, assess adherence to desired behaviors, and enable AI agents to monitor real-world social context and evolve system behaviors over long time-horizons. You will develop new paradigms for human+AI rating that are adaptive and context-aware, driving breakthroughs within Google DeepMind, Google products, and the broader AI alignment community.
Research Scientist, Gemini Safety
Google DeepMind
The Gemini Safety team at Google DeepMind is responsible for the safety and fairness of the latest Gemini models. As a Research Scientist/Engineer, you will apply and develop cutting-edge data and algorithmic solutions to advance these user-facing models. This is a fast-paced, highly collaborative role within a supportive team dedicated to pushing the boundaries of AI for public benefit and scientific discovery, with safety and ethics as the highest priorities.
Head of EMEA Partnerships
crusoe
Crusoe is seeking a Head of EMEA Partnerships to drive regional go-to-market acceleration for its AI infrastructure. This high-impact, individual contributor role involves translating global partnership strategies into tangible business outcomes across Europe and the Middle East. The position requires activating a vibrant ecosystem of startups, channel partners, and System Integrators to establish Crusoe Cloud as the leading AI infrastructure choice. The role involves collaborating with US-based leadership on global playbooks while taking full ownership of their tactical execution in key regional hubs. The ideal candidate is a metric-driven strategist adept at identifying and capitalizing on high-signal opportunities in a fast-paced environment.
Member of Technical Staff, Software Engineer
fireworks ai
Fireworks is seeking a Member of Technical Staff, Software Engineer to join their team. This role involves building the core backend systems that power Fireworks' platform for specialized intelligence, enabling companies to build, train, and serve AI models tailored to their own data, workflows, and products. You will own major product surfaces from architecture to production, improving reliability, performance, and developer experience. This is platform engineering with product impact, where your systems will directly shape how customers build on top of AI. You will work closely with product, frontend, infra, and GTM to ship end-to-end features, and use AI tooling aggressively to automate tasks.
Applied Legal Researcher
Harvey
Harvey is transforming how legal and professional services operate by combining frontier agentic AI, an enterprise-grade platform, and deep domain expertise. We are looking for a legal researcher with a strong understanding of how law firms, financial institutions, large corporations, and other organizations deliver professional services and execute complex legal and knowledge work tasks. This role is crucial for accelerating customer workflows with AI systems, requiring a deep understanding of their needs and how AI can be applied.
$180k - $220k
Applied Legal Researcher
Harvey
Harvey is transforming how legal and professional services operate by combining frontier agentic AI, an enterprise-grade platform, and deep domain expertise. We are looking for a legal researcher with a strong understanding of how law firms, financial institutions, large corporations, and other organizations deliver professional services and execute complex legal and knowledge work tasks. This role is crucial for accelerating customer workflows with AI systems, requiring a deep understanding of their needs and how AI can be applied.
$180k - $220k
Applied Researcher
Physical Intelligence
Physical Intelligence is building general-purpose AI for the physical world, developing foundation models and learning algorithms for robots and future physically-actuated devices. The Deployments team focuses on solving real-world problems by integrating with customer workflows, training models for dexterous tasks, and ensuring system reliability. This full-stack robotics role involves everything from customer-facing experiences to fine-tuning models for novel tasks, aiming to deliver the best solutions.