PyTorch Jobs
223 open roles mentioning PyTorch
AI research scientist
Writer
AI research at WRITER focuses on building the scientific foundation for ambitious enterprise AI deployments. As a staff AI research scientist, you will drive a high-impact research agenda centered on large language models, agentic reasoning, and system-level capabilities essential for enterprise-scale AI. This role offers a unique opportunity to advance the field while directly contributing to products used by hundreds of thousands daily. You will work on post-training, planning, multi-step reasoning, and agentic workflows, directly shaping the future of enterprise AI performance and scalability. The role provides resources, infrastructure, and cross-functional support to pursue and implement ambitious ideas rapidly.
Research Engineer, Frontier Speculative Decoding
Together AI
Together AI is building the Inference Platform that powers the world's most advanced generative AI models. This role will serve as a critical bridge between cutting-edge research and real-world applications, focusing on translating internal model training research into production-ready deployments for customers. The work involves a deep commitment to data-centric development, meticulous hyperparameter tuning, and rigorous checkpoint evaluation. You will transform general-purpose models into highly performant, specialized tools by fine-tuning them on customer-specific data and internal datasets, working with dedicated GPU clusters rather than training foundation models from scratch.
$190k - $270k
Research Intern, Model Shaping (Fall 2026)
Together AI
As a Research Intern in the Model Shaping team, you will work on advanced post-training methods, new techniques for efficient neural network training, and robust evaluation of foundation model capabilities. The Model Shaping team at Together AI focuses on tailoring open foundation models for downstream applications, building services for machine learning developers, and developing new methods for efficient model training and evaluation. This role offers the opportunity to contribute to cutting-edge research and potentially influence open-source projects.
Frontier Agents Intern (Fall 2026)
Together AI
The Agents team investigates how to build, align, and scale frontier AI systems capable of complex, multi-step tasks and workflows across text and speech, with a focus on agentic and scientific domains. This role sits at the intersection of agent capabilities, human-computer interaction, and infrastructure, exploring areas like post-training methods for agentic behavior and developing evaluation frameworks for open-ended tasks. As a research intern, you will tackle challenges in alignment, reliability, and scalability, potentially working on new training recipes for self-learning and long-horizon reasoning, curating datasets, studying failure modes, or building scalable agent infrastructure.
Staff Machine Learning Engineer, Voice AI
Together AI
Together AI is building the best inference infrastructure for voice applications, powering production-grade, real-time voice agents and applications. We are seeking a Staff ML Engineer to lead the model serving layer for voice workloads. This role involves hands-on optimization of inference engines and models like Whisper, Parakeet, Orpheus, and Kokoro, focusing on pushing latency and throughput boundaries. You will address unique challenges in voice inference, such as streaming audio and real-time latency, and shape the future of how voice models are served as the industry shifts towards end-to-end speech-to-speech systems. This is a foundational hire on a small, high-impact team.
$220k - $280k
Senior Machine Learning Engineer, Voice AI
Together AI
Together AI is building the best inference infrastructure for voice applications, powering production-grade, real-time voice agents and applications. We are seeking a Senior ML Engineer to lead the model serving layer for voice workloads. This role involves hands-on optimization of inference engines and models like Whisper, Parakeet, Orpheus, and Kokoro to achieve frontier-level latency and throughput. You will focus on unique voice inference challenges such as streaming audio, tokenization, and real-time latency budgets, shaping how voice models are served as the industry shifts towards end-to-end speech-to-speech systems. This is a foundational hire on a small, high-impact team.
$200k - $260k
Machine Learning, Platform Engineer
Together AI
Together AI is a research-driven artificial intelligence company focused on lowering the cost of modern AI systems. This role is part of a team dedicated to enabling custom models and dedicated inference on Together's platform. The team is responsible for building a container platform, optimizing autoscaling, minimizing cold starts, achieving the best end-to-end model performance, and providing a best-in-class developer experience with great tooling. The work often involves video or audio generation across the stack, including CUDA kernels, PyTorch optimization, inference engines, container orchestration, and queueing theory.
$160k - $250k
Machine Learning Engineer - Inference
Together AI
Together AI is seeking a Machine Learning Engineer to join our Inference Engine team, focusing on optimizing and enhancing the performance of our AI inference systems. This role involves working with state-of-the-art large language models and ensuring they run efficiently and effectively at scale. You will collaborate closely with AI researchers and engineers to create cutting-edge AI solutions and shape the future of AI inference.
$160k - $230k
LLM Inference Frameworks and Optimization Engineer
Together AI
Together.ai is building state-of-the-art infrastructure for efficient and scalable inference of large language models (LLMs). The company's mission is to optimize inference frameworks, algorithms, and infrastructure to push the boundaries of performance, scalability, and cost-efficiency. They are seeking an Inference Frameworks and Optimization Engineer to design, develop, and optimize distributed inference engines for multimodal and language models at scale. This role will focus on low-latency, high-throughput inference, GPU/accelerator optimizations, and software-hardware co-design, ensuring efficient large-scale deployment of LLMs and vision models. This position offers a unique opportunity to shape the future of LLM inference infrastructure and ensure scalable, high-performance AI deployment across diverse applications.
$160k - $230k
AI Field Engineer - Enterprise
fireworks ai
Fireworks is seeking an AI Field Engineer to join their team. This role is crucial for embedding with ambitious customers and technology partners to rapidly transform complex AI challenges into production systems. You will operate at the intersection of engineering, product, and customer delivery, taking a hands-on approach to building Proofs of Concept (POCs), Minimum Viable Products (MVPs), and production integrations. Simultaneously, you will engage in executive-level discussions regarding architecture, strategy, and business outcomes. The position involves significant coding, running benchmarks, debugging production issues, and architecting deployments, alongside leading discovery conversations, aligning stakeholders, and translating customer needs into product improvements. This role requires comfort working on-site with customers to build relationships and trust in person.
AI Field Engineer - Strategic Partnerships
fireworks ai
Fireworks is seeking an AI Field Engineer for Strategic Partnerships to be a technical owner for their most strategic partnerships. This role involves working closely with partner field teams, ISVs, and SIs to establish Fireworks as the primary inference and fine-tuning layer within partner AI architectures. You will bridge engineering, partner development, and customer delivery by building reference architectures, conducting benchmarks, resolving integration issues, and co-developing Proofs of Concept (POCs). The position requires strong technical skills, the ability to manage executive-level discussions on strategy and business outcomes, and a proactive approach to building and enabling solutions. You will also be responsible for leading discovery conversations, aligning stakeholders, and translating field feedback into product improvements to accelerate the development cycle.
$200k - $260k
Software Engineer, AI Training and Infrastructure
Skild
Skild AI is seeking a Software Engineer to develop and optimize the software infrastructure and tools for training cutting-edge AI models. This role involves building scalable and efficient training pipelines and frameworks that support the entire machine learning lifecycle, from data preparation to model deployment. You will collaborate with researchers and machine learning engineers to integrate new algorithms and techniques, pushing the boundaries of AI in real-world robotics. The position also focuses on exploring efficient data utilization within training pipelines.
Software Engineer, AI Inference
Skild
Skild AI is seeking a Software Engineer to join their team and work on deploying cutting-edge AI models for embodied systems. This role focuses on optimizing AI inference processes for models ranging from lightweight to billion-parameter scale, ensuring robots operate efficiently and intelligently in real-world scenarios. You will bridge the gap between systems and machine learning, directly enhancing AI model power and adaptability by maintaining consistent performance under varying compute and hardware constraints.
AI Field Engineer - AI Natives
fireworks ai
Fireworks is seeking an AI Field Engineer to join their team. This role is at the forefront of technical engagement, embedding with ambitious customers and technology partners to rapidly transform complex AI challenges into production systems. You will operate at the intersection of engineering, product, and customer delivery, taking a hands-on approach to building Proofs of Concept (POCs), Minimum Viable Products (MVPs), and production integrations. Simultaneously, you will engage in executive-level discussions on architecture, strategy, and business outcomes. The position requires a significant amount of hands-on work, including shipping code, running benchmarks, debugging production issues, and architecting deployments, alongside leading discovery conversations, aligning stakeholders, and translating customer needs into product improvements.
Machine Learning Platform Engineer
Scale AI
We are seeking a Machine Learning Platform Engineer to develop internal tooling for fine-tuning and evaluating large language models. This role is crucial for advancing our capabilities in working with cutting-edge AI technologies.
$140k - $190k
Member of Technical Staff- Full Stack
fireworks ai
Fireworks is seeking a Full Stack Software Engineer to architect, build, and scale its developer-focused web application and shape its end-to-end technical architecture. This role combines hands-on engineering with technical leadership, offering opportunities to mentor engineers, collaborate cross-functionally, and drive innovation in a fast-paced, AI-focused environment. You will be instrumental in delivering scalable, elegant, and highly performant solutions that directly impact hundreds of thousands of developers and the company's product vision. The position involves shaping and shipping the optimal developer journey, encompassing intuitive front-end interfaces, robust APIs, and efficient backend orchestration, integrated with a wide offering of open-source generative AI models.
Applied AI Researcher
Databricks
We are seeking an Applied AI Researcher to join our team. This role will focus on researching and implementing state-of-the-art techniques in fine-tuning and Retrieval-Augmented Generation (RAG). You will be at the forefront of developing and deploying advanced AI solutions.
$175k - $245k
Applied AI, Technical Lead, Forward Deployed AI Engineer - Munich
Mistral AI
Mistral AI is seeking a Technical Lead, Applied AI to drive the technical strategy, execution, and delivery of complex AI solutions for enterprise customers. This role involves leading project teams of Applied AI Engineers to ensure successful deployment of Mistral AI products and the development of high-impact, scalable AI use cases. You will serve as the primary technical point of contact for strategic customers, guiding them through the entire lifecycle from pre-sales to post-implementation, and collaborating with research, product, and engineering teams to shape future offerings. The position bridges the gap between AI research and real-world enterprise applications, ensuring solutions are robust, scalable, and aligned with customer needs and Mistral's vision.
Applied AI, Forward Deployed Machine Learning Engineer - Munich
Mistral AI
Mistral AI is looking for an Applied AI Engineer to join their customer-facing technical team. This role involves working directly with enterprise clients to deploy cutting-edge AI solutions and address complex technical challenges. You will bridge the gap between AI research and real-world applications, ensuring solutions are robust, scalable, and aligned with customer needs and Mistral's vision. The team operates with a focus on impact and collaboration, valuing direct communication and the best ideas regardless of seniority.
Computer Vision AI & ML Engineer
Skild
Skild AI is seeking a Computer Vision AI & ML Engineer to design, build, and deploy advanced perception systems for real-world robotics and automation. You will work across the full machine learning lifecycle—model development, data strategy, evaluation, and production integration—to deliver robust, high-performance vision capabilities. This role combines applied research with hands-on engineering and offers the opportunity to influence both architecture and roadmap decisions.