PyTorch Jobs
109 open roles mentioning PyTorch
Performance Engineer, GPU
Anthropic
Pioneering the next generation of AI requires breakthrough innovations in GPU performance and systems engineering. As a GPU Performance Engineer, you'll architect and implement the foundational systems that power Claude and push the frontiers of what's possible with large language models. You'll be responsible for maximizing GPU utilization and performance at unprecedented scale, developing cutting-edge optimizations that directly enable new model capabilities and dramatically improve inference efficiency. Working at the intersection of hardware and software, you'll implement state-of-the-art techniques from custom kernel development to distributed system architectures. Your work will span the entire stack—from low-level tensor core optimizations to orchestrating thousands of GPUs in perfect synchronization. Strong candidates will have a track record of delivering transformative GPU performance improvements in production ML systems and will be excited to shape the future of AI infrastructure alongside world-class researchers and engineers.
Research Intern (BS/MS/PhD)
pika
Pika is shaping the future of creative infrastructure with real-time, multimodal generation and intelligent agentic platforms. We are seeking passionate Research Interns at the BS, MS, or PhD level with significant experience working on Videogen to join our team and make an immediate impact. In this role, you will directly contribute to mid-training and post-training efforts for Videogen models and support advanced generative systems. This internship is ideal for students and researchers who have hands-on Videogen experience, are excited to work at the intersection of generative AI and multimedia, and are eager to make meaningful contributions alongside top AI talent.
Software Engineer, Platform, Tinker
thinkingmachines
Thinking Machines Lab is seeking a Software Engineer to own the platform systems that power Tinker, their fine-tuning API. This role involves developing and maintaining critical components such as billing and usage metering, permissions and access control, organization and team management, data export functionalities, and audit logging. You will collaborate closely with product, legal, and other cross-functional teams, ensuring that new features, pricing adjustments, and enterprise deals are seamlessly integrated into the platform. This is an opportunity to grow a rapidly expanding platform and contribute to the Tinker community.
$350k - $475k
Software Engineer, Research Acceleration
thinkingmachines
Thinking Machines Lab is seeking engineers to build the libraries and tools that accelerate research. You will own internal infrastructure, including evaluation libraries, RL training libraries, and experiment tracking platforms, and build systems that compound research velocity over time. This is a collaborative role where you will work directly with researchers to identify bottlenecks and pain points. Success means researchers trust your systems to just work and find them a delight to use.
$350k - $475k
Research Infrastructure Engineer, Research Acceleration
thinkingmachines
Thinking Machines Lab is seeking engineers to build the libraries and tools that accelerate research. You will own internal infrastructure, including evaluation libraries, RL training libraries, and experiment tracking platforms, to build systems that compound research velocity over time. This is a collaborative role where you will work directly with researchers to identify bottlenecks and pain points. Success means researchers trust your systems to just work and find them a delight to use.
$350k - $475k
Applied Scientist / Research Engineer
Mistral AI
Mistral AI is seeking Applied Scientists and Research Engineers to drive innovative research and collaborate with clients on complex research projects. You will develop state-of-the-art models across different modalities such as text, image, and speech. By developing novel methods and research ideas, you will apply these models across a diverse set of use cases and domains. Working cross-functionally with both external and internal science, engineering, and product teams, you will deliver high-impact AI solutions that turn the needle.
Senior/Staff Applied Scientist/Research Engineer
Mistral AI
Mistral AI is seeking Applied Scientists and Research Engineers to drive innovative research and collaborate with clients on complex research projects. You will develop state-of-the-art models across different modalities such as text, image, and speech, applying novel methods and research ideas to diverse use cases and domains. Working cross-functionally with science, engineering, and product teams, you will deliver high-impact AI solutions. We are a dynamic, collaborative team passionate about AI and its potential to transform society, with teams distributed across France, USA, UK, Germany, and Singapore.
Applied AI, Machine Learning Engineer
Mistral AI
Mistral AI is seeking an Applied AI Engineer to join their team and facilitate the adoption of their AI products among customers. This role involves collaborating with clients to address complex technical challenges, from pre-sale to post-implementation, ensuring Mistral's solutions meet and exceed expectations. The engineer will manage daily customer relations, act as a key resource in externalizing research into production settings, and work on state-of-the-art Generative AI applications across various industries. This position offers the opportunity to contribute to a pioneering company shaping the future of AI and make a meaningful impact.
AI research scientist
Writer
AI research at WRITER focuses on building the scientific foundation for ambitious enterprise AI deployments. As a staff AI research scientist, you will drive a high-impact research agenda centered on large language models, agentic reasoning, and system-level capabilities essential for enterprise-scale AI. This role offers a unique opportunity to advance the field while directly contributing to products used by hundreds of thousands daily. You will work on post-training, planning, multi-step reasoning, and agentic workflows, directly shaping the future of enterprise AI performance and scalability. The role provides resources, infrastructure, and cross-functional support to pursue and implement ambitious ideas rapidly.
Research Engineer, Frontier Speculative Decoding
Together AI
Together AI is building the Inference Platform that powers the world's most advanced generative AI models. This role will serve as a critical bridge between cutting-edge research and real-world applications, focusing on translating internal model training research into production-ready deployments for customers. The work involves a deep commitment to data-centric development, meticulous hyperparameter tuning, and rigorous checkpoint evaluation. You will transform general-purpose models into highly performant, specialized tools by fine-tuning them on customer-specific data and internal datasets, working with dedicated GPU clusters rather than training foundation models from scratch.
$190k - $270k
Research Intern, Model Shaping (Fall 2026)
Together AI
As a Research Intern in the Model Shaping team, you will work on advanced post-training methods, new techniques for efficient neural network training, and robust evaluation of foundation model capabilities. The Model Shaping team at Together AI focuses on tailoring open foundation models for downstream applications, building services for machine learning developers, and developing new methods for efficient model training and evaluation. This role offers the opportunity to contribute to cutting-edge research and potentially influence open-source projects.
Frontier Agents Intern (Fall 2026)
Together AI
The Agents team investigates how to build, align, and scale frontier AI systems capable of complex, multi-step tasks and workflows across text and speech, with a focus on agentic and scientific domains. This role sits at the intersection of agent capabilities, human-computer interaction, and infrastructure, exploring areas like post-training methods for agentic behavior and developing evaluation frameworks for open-ended tasks. As a research intern, you will tackle challenges in alignment, reliability, and scalability, potentially working on new training recipes for self-learning and long-horizon reasoning, curating datasets, studying failure modes, or building scalable agent infrastructure.
Staff Machine Learning Engineer, Voice AI
Together AI
Together AI is building the best inference infrastructure for voice applications, powering production-grade, real-time voice agents and applications. We are seeking a Staff ML Engineer to lead the model serving layer for voice workloads. This role involves hands-on optimization of inference engines and models like Whisper, Parakeet, Orpheus, and Kokoro, focusing on pushing latency and throughput boundaries. You will address unique challenges in voice inference, such as streaming audio and real-time latency, and shape the future of how voice models are served as the industry shifts towards end-to-end speech-to-speech systems. This is a foundational hire on a small, high-impact team.
$220k - $280k
Senior Machine Learning Engineer, Voice AI
Together AI
Together AI is building the best inference infrastructure for voice applications, powering production-grade, real-time voice agents and applications. We are seeking a Senior ML Engineer to lead the model serving layer for voice workloads. This role involves hands-on optimization of inference engines and models like Whisper, Parakeet, Orpheus, and Kokoro to achieve frontier-level latency and throughput. You will focus on unique voice inference challenges such as streaming audio, tokenization, and real-time latency budgets, shaping how voice models are served as the industry shifts towards end-to-end speech-to-speech systems. This is a foundational hire on a small, high-impact team.
$200k - $260k
Machine Learning, Platform Engineer
Together AI
Together AI is a research-driven artificial intelligence company focused on lowering the cost of modern AI systems. This role is part of a team dedicated to enabling custom models and dedicated inference on Together's platform. The team is responsible for building a container platform, optimizing autoscaling, minimizing cold starts, achieving the best end-to-end model performance, and providing a best-in-class developer experience with great tooling. The work often involves video or audio generation across the stack, including CUDA kernels, PyTorch optimization, inference engines, container orchestration, and queueing theory.
$160k - $250k
Machine Learning Engineer - Inference
Together AI
Together AI is seeking a Machine Learning Engineer to join our Inference Engine team, focusing on optimizing and enhancing the performance of our AI inference systems. This role involves working with state-of-the-art large language models and ensuring they run efficiently and effectively at scale. You will collaborate closely with AI researchers and engineers to create cutting-edge AI solutions and shape the future of AI inference.
$160k - $230k
LLM Inference Frameworks and Optimization Engineer
Together AI
Together.ai is building state-of-the-art infrastructure for efficient and scalable inference of large language models (LLMs). The company's mission is to optimize inference frameworks, algorithms, and infrastructure to push the boundaries of performance, scalability, and cost-efficiency. They are seeking an Inference Frameworks and Optimization Engineer to design, develop, and optimize distributed inference engines for multimodal and language models at scale. This role will focus on low-latency, high-throughput inference, GPU/accelerator optimizations, and software-hardware co-design, ensuring efficient large-scale deployment of LLMs and vision models. This position offers a unique opportunity to shape the future of LLM inference infrastructure and ensure scalable, high-performance AI deployment across diverse applications.
$160k - $230k
Machine Learning Platform Engineer
Scale AI
We are seeking a Machine Learning Platform Engineer to develop internal tooling for fine-tuning and evaluating large language models. This role is crucial for advancing our capabilities in working with cutting-edge AI technologies.
$140k - $190k
Applied AI Researcher
Databricks
We are seeking an Applied AI Researcher to join our team. This role will focus on researching and implementing state-of-the-art techniques in fine-tuning and Retrieval-Augmented Generation (RAG). You will be at the forefront of developing and deploying advanced AI solutions.
$175k - $245k
Applied AI, Technical Lead, Forward Deployed AI Engineer - Munich
Mistral AI
Mistral AI is seeking a Technical Lead, Applied AI to drive the technical strategy, execution, and delivery of complex AI solutions for enterprise customers. This role involves leading project teams of Applied AI Engineers to ensure successful deployment of Mistral AI products and the development of high-impact, scalable AI use cases. You will serve as the primary technical point of contact for strategic customers, guiding them through the entire lifecycle from pre-sales to post-implementation, and collaborating with research, product, and engineering teams to shape future offerings. The position bridges the gap between AI research and real-world enterprise applications, ensuring solutions are robust, scalable, and aligned with customer needs and Mistral's vision.