PyTorch Jobs
223 open roles mentioning PyTorch
Multimodal Generative AI Researcher
Stability AI
We are seeking a Research Scientist with deep expertise in training and fine-tuning large Vision-Language and Language Models (VLMs / LLMs) for downstream multimodal tasks. You will be instrumental in advancing models that reason across vision, language, and 3D, translating research breakthroughs into scalable engineering solutions. This role involves designing and fine-tuning large-scale VLMs/LLMs and hybrid architectures for complex tasks like visual reasoning, retrieval, 3D understanding, and embodied interaction.
Research Engineer, Machine Learning
Mistral AI
Mistral AI is democratizing AI through high-performance, optimized, open-source models and solutions. We are a dynamic, collaborative team passionate about AI's potential to transform society, with a diverse workforce driving innovation. As a Research Engineer – ML track, you will build and optimize large-scale learning systems powering our open-weight models. You will work hand-in-hand with Research Scientists, either enhancing the shared training framework and data pipelines or embedding within a research squad to turn fresh ideas into scalable code.
Software Engineer - Training Product
Baseten
Baseten is seeking a customer-obsessed software engineer to join their team and contribute to the development of mission-critical AI inference platforms. In this role, you will own features from conception to launch, working across the entire technology stack from API and UI down to the infrastructure layer. You will have the opportunity to fine-tune models, gain a deep understanding of user workflows, and collaborate closely with research engineers to build cutting-edge experiences that accelerate model development and address real-world pain points. If you are excited about diving deep into AI model training and building impactful products, this is the role for you.
Applied AI, Evaluation Engineer
Mistral AI
Mistral AI is seeking an Evaluation Engineer to join our customer-facing Applied AI team. This role is crucial for ensuring our AI solutions are production-ready by designing methodologies, building infrastructure, and defining evaluation standards across various industries and use cases. You will bridge the gap between research and customer needs, creating evaluation frameworks that measure LLM performance in real-world scenarios, moving beyond standard benchmarks to address domain-specific risks and requirements. This position offers a unique opportunity to impact the deployment of cutting-edge AI by directly contributing to its measurable success for enterprise clients.
MTS, Security
fireworks ai
Fireworks is seeking a Security Engineer to play a key role in designing, implementing, and operating security controls across AI infrastructure, AI platforms, and internal systems. This role is crucial for strengthening our security posture and supporting rapid growth, ensuring the confidentiality, integrity, and availability of data, models, and infrastructure as organizations increasingly rely on large language models and cloud-native AI services. You will be instrumental in building trust by embedding security across all layers of our technology stack.
Internship - Search Machine Learning Engineer
Perplexity AI
Perplexity is seeking a Search Machine Learning Engineer Intern to contribute to the development of next-generation search technologies, with a specific emphasis on retrieval and ranking. Interns will collaborate with seasoned engineers to enhance search quality, explore novel models, and implement features that directly influence user search and information discovery experiences. This internship program offers a duration of 12-24 weeks, operating on a full-time or part-time basis, and requires in-person attendance at our Belgrade office.
$12k - $24k
Staff Research Engineer - Interactive Avatars
synthesia.io
Synthesia is seeking a Staff Research Engineer to join their R&D Department, focusing on cutting-edge challenges in Generative AI, specifically avatar-centric interactive video diffusion models. This role offers the chance to work on applied research that directly impacts solutions used by thousands of businesses worldwide. You will be instrumental in shaping the future of AI video agents capable of human-like thinking, acting, and reacting. The team operates in a fast-paced environment, iterating frequently and building impactful models that ship to production.
Software Engineer - Model Performance Systems
Baseten
Baseten is seeking Software Engineers to join their team in a specialized, high-impact role at the intersection of high-performance computing (HPC) and Large Language Model (LLM) engineering. This position involves not only building automated performance monitoring and diagnostic tools for next-generation AI infrastructure but also defining the roadmap, driving key technical decisions, and taking full ownership of the future direction of this work. The role offers the opportunity to shape the platform that engineers use to deploy cutting-edge AI models into production.
Member of Technical Staff, Data Analysis and Evaluation
Cohere
Cohere is seeking a Member of Technical Staff in Data Analysis and Evaluation to ensure the quality, reliability, and performance of our large language models (LLMs). This role involves designing and conducting data collection tasks, assessing dataset quality, and analyzing model robustness and generalisability. You will collaborate with researchers, engineers, and data annotators to drive data-driven decisions and enhance AI system effectiveness. The position requires expertise in statistics, experimental design, and machine learning to ensure high-quality data and reliable model performance across diverse scenarios, contributing to Cohere's mission of advancing AI.
Senior ML Systems Engineer, Frameworks & Tooling
Cohere
Cohere is seeking a Senior ML Systems Engineer to join their team and build, maintain, and evolve the training framework that powers their frontier-scale language models. This role is ideal for someone passionate about large-scale training, distributed systems, and HPC infrastructure, offering the opportunity to design and maintain core components for fast, reliable, and scalable model training. You will also build tooling to connect research ideas to thousands of GPUs, working across the full stack of ML systems with significant autonomy and impact.
Applied AI, Technical Lead - Forward Deployed AI Engineer
Mistral AI
Mistral AI is seeking a Technical Lead, Applied AI to drive the technical strategy, execution, and delivery of complex AI solutions for enterprise customers. In this role, you will lead project teams of Applied AI Engineers, ensuring the successful deployment of Mistral AI products and the development of high-impact, scalable AI use cases. You will act as the primary technical point of contact for strategic customers, guiding them through the entire lifecycle from pre-sales to post-implementation, while collaborating closely with research, product, and engineering teams to shape the future of our offerings. As a Technical Lead, you will bridge the gap between cutting-edge AI research and real-world enterprise applications, ensuring our solutions are robust, scalable, and aligned with both customer needs and Mistral’s technological vision.
Audio Inference Engineer, Model Efficiency
Cohere
Cohere is seeking an Audio Inference Engineer focused on Model Efficiency to join a fast-growing team of researchers and engineers. The mission of this team is to build reliable machine learning systems and optimize audio inference serving efficiency using innovative techniques. As an engineer on this team, you will advance core audio model serving metrics, including latency, throughput, and quality by diving deep into systems, identifying bottlenecks, and delivering creative solutions for audio processing and streaming workloads. You will collaborate closely with both the training and serving infrastructure teams to ensure seamless integration between model development and deployment, with a special focus on real-time and streaming audio inference.
Member of Technical Staff, LLM Infrastructure
fireworks ai
As a Software Engineer on the AI Infrastructure team, you will help design the core systems that power Fireworks AI’s generative AI platform. You will build infrastructure and tools that ensure the reliability, performance, quality, and availability of our AI system. Your mission is to make Fireworks AI the most reliable and user-friendly generative AI platform in the world. You will partner closely with our cloud infrastructure, product, and performance teams to deliver infrastructure that bridges the gap between our customers and the ultra-performant proprietary Fireworks inference engine.
Applied AI, Forward Deployed Machine Learning Engineer - Morocco
Mistral AI
Mistral AI is seeking an Applied AI Engineer to join its customer-facing technical organization. This role involves facilitating customer adoption of Mistral AI products, addressing complex technical challenges, and deploying cutting-edge AI solutions that deliver measurable business impact. You will bridge the gap between AI research and real-world enterprise applications, ensuring solutions are robust, scalable, and aligned with customer needs and Mistral's technological vision. The team operates with a focus on outputs, direct communication, and a low-ego, high-standards environment, embracing an unstructured setting.
Member of Technical Staff, Evals & Post-Training Product
fireworks ai
Fireworks is seeking a Member of Technical Staff, Evals & Post-Training Product to define how developers improve models on the Fireworks platform. This role combines scalable system design, deep data science, and model quality. You will build the infrastructure and workflows connecting evaluation and post-training, taking our evaluation setup to the next stage by improving programmatic access and scale. You will work across backend systems, sandbox infrastructure, and user-facing surfaces to simplify the authoring of evaluations, understanding of results, and rapid iteration.
AI Scientist - Robotics
Mistral AI
Mistral is on a mission to democratize AI by producing frontier intelligence for everyone, developed in the open and built by a global team. We are a dynamic, collaborative group passionate about AI's potential to transform society, with a diverse workforce distributed across Europe, the USA, and Asia. We focus on developing models for enterprise and consumers that can change business operations and integrate into daily lives, while also releasing frontier models open-source. We are hiring experts in large language model training and distributed systems to shape the future of AI.
Applied AI, Technical Lead, Forward Deployed AI Engineer - EMEA
Mistral AI
Mistral AI is seeking a Technical Lead, Applied AI to drive the technical strategy, execution, and delivery of complex AI solutions for enterprise customers. This role involves leading project teams of Applied AI Engineers to ensure successful deployment of Mistral AI products and the development of high-impact, scalable AI use cases. You will serve as the primary technical point of contact for strategic customers, guiding them through the entire lifecycle from pre-sales to post-implementation, and collaborating with research, product, and engineering teams to shape future offerings. The position bridges cutting-edge AI research with real-world enterprise applications, ensuring solutions are robust, scalable, and aligned with customer needs and Mistral's vision.
AI Scientist - Warsaw
Mistral AI
Mistral AI is a pioneering company dedicated to democratizing AI through high-performance, optimized, open-source models, products, and solutions. We aim to simplify tasks, save time, and enhance learning and creativity by integrating cutting-edge AI into daily working life. Our comprehensive platform serves both enterprise and personal needs, featuring offerings like Le Chat, La Plateforme, Mistral Code, and Mistral Compute. We are a dynamic, collaborative, and diverse team passionate about AI's potential to transform society, driving innovation from our distributed teams across France, USA, UK, Germany, and Singapore. Join us to shape the future of AI and make a meaningful impact.
AI Scientist - Zurich
Mistral AI
Mistral AI is a pioneering company dedicated to democratizing AI through high-performance, optimized, open-source, and cutting-edge models, products, and solutions. Our comprehensive AI platform serves both enterprise and personal needs, offering tools like Le Chat, La Plateforme, Mistral Code, and Mistral Compute to bring frontier intelligence to end-users. We are a dynamic, collaborative, and diverse team passionate about AI's potential to transform society, thriving in competitive environments and committed to driving innovation. Join us to be part of shaping the future of AI and making a meaningful impact.
Member of Technical Staff - Pre-Training
Reflection ai
Reflection is a research lab dedicated to making intelligence open and accessible for everyone. We build open models that empower individuals to control their intelligence and shape the future of AI. As a Member of Technical Staff - Pre-Training, you will be instrumental in researching and building solutions across algorithms, scaling laws, data processing, optimizers, and model architecture. This role involves designing and executing scientific experiments to deepen our understanding of scaling large language models and data efficiency, while also implementing state-of-the-art deep learning methods. You will have the opportunity to lead small research projects independently and contribute to larger initiatives, optimizing training infrastructure for efficient scaling and working across the entire stack from low-level optimizations to high-level model design.