PyTorch Jobs
223 open roles mentioning PyTorch
Full-Stack Engineer
Nabla
Nabla is seeking a Full-Stack Engineer to join one of their cross-functional squads, contributing end-to-end to specific product areas. This role involves building and improving product features across web and desktop platforms, contributing to back-end APIs and services, and collaborating closely with Product, ML, and Design teams to bring projects from conception to production. The engineering team is lean, fast-moving, and deeply technical, focused on delivering AI into clinical settings reliably and at scale. Nabla is backed by a $70M Series C funding round and is at the forefront of developing AI assistants for healthcare, aiming to restore the human connection in medicine.
Staff+ Software Engineer, ML Inference Path
Anthropic
The Safeguards ML Inference Path team designs, builds, and operates the production infrastructure that powers Claude's ML based safety systems. We collaborate closely with safety researchers and inference engineers to bring new classifiers and novel classes of ML defenses to production. We own the research → production transfer of new safety technologies that is on the critical path for every Claude model launch. We build for scale, serving thousands of ML classifiers for all requests on the token generation path and for every platform Claude runs on. We are looking for engineers who have deep expertise in productionizing ML systems, working at the intersection of machine learning, large-scale distributed systems, and AI safety, developing the platforms and tools that enable our safeguards to operate reliably at scale.
Member of Technical Staff, New Grad (BS/MS)
fireworks ai
This role is designed for new graduates who are eager to work on real-world AI systems in production from day one. As a Member of Technical Staff at Fireworks, you will write and deploy code that operates on one of the world's most active inference platforms, serving hundreds of models to developers and enterprises at a massive scale. You will be matched to a team based on your strengths and interests, with potential work spanning inference and performance, model training and fine-tuning, distributed systems and cloud infrastructure, developer platform and APIs, or product and full-stack application engineering. You will be supported by a senior engineer mentor and a structured ramp-up program, starting with well-scoped problems and progressing to owning features and systems independently. The company emphasizes a fast-paced environment with a high bar and a commitment to employee development.
Member of Technical Staff, Systems Infrastructure (2026 PhD New Grad)
fireworks ai
Fireworks is seeking PhD graduates to join their Systems Infrastructure team. This role is designed for individuals finishing their PhD in Computer Science, Computer Engineering, Electrical Engineering, or a similar field, who are interested in applying their research to large-scale, real-world AI infrastructure. You will be responsible for designing and building the core systems that power Fireworks, including schedulers, storage systems, and networks, to ensure efficient operation of tens of thousands of accelerators and low inference latency. You will tackle complex problems such as optimizing job placement on heterogeneous hardware, high-speed data movement for model weights, and maintaining saturated datacenter networks. You will be paired with a senior engineer mentor and work on impactful projects from day one, with start dates flexible around thesis defense.
Member of Technical Staff, Research (2026 PhD New Grad)
fireworks ai
Fireworks is seeking a Member of Technical Staff on the Research team, designed for PhD candidates who want their work to reach production. This role involves pushing the boundaries of generative AI, advancing LLMs and multimodal systems through foundational research. You will enhance model efficiency, accuracy, and scalability, directly shaping high-performance AI infrastructure. You will be paired with a senior researcher as a mentor and given a real research problem from day one, collaborating with top experts in deep learning, distributed systems, and optimization. The work you contribute will be deployed by leading companies, often within weeks.
Member of Technical Staff - Multimodal Understanding
xAI
SpaceXAI is seeking a Member of Technical Staff to join their multimodal team and advance the understanding and generation of AI across image, video, audio, and text. This role involves working across the full stack, from data curation and pre-training to alignment, infrastructure, and end-to-end product experiences. You will collaborate with various teams to deliver cutting-edge multimodal reasoning, world modeling, tool use, agentic behaviors, and human-AI collaboration capabilities. The goal is to build models that can perceive, reason about, and interact with the world in real-time at an unprecedented level.
$180k - $440k
Member of Technical Staff - Imagine Model
xAI
SpaceXAI is seeking a multimodal engineer for the Imagine Model Team to develop cutting-edge AI experiences beyond text, focusing on high-fidelity understanding and generation across image and video modalities, with audio integration where it enhances visual content. Responsibilities cover data curation, modeling, training, inference serving, and product integration across pretraining and post-training phases. The role involves close collaboration with product teams to advance model frontiers and deliver exceptional end-to-end user experiences.
$180k - $440k
Member of Technical Staff - RL Inference
xAI
SpaceXAI's mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. The RL infrastructure team is looking for an engineer to help with low precision RL training and inference. This role involves designing and optimizing the inference stack for all shapes of RL workloads, analyzing and addressing performance bottlenecks in large-scale RL systems, and working closely with the modeling team to efficiently implement novel RL techniques and algorithms.
$180k - $440k
Software Engineer - Voice Model
xAI
SpaceXAI's mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. The Grok Voice Model team is building the world's best voice AI, delivering smooth, natural, low-latency spoken interactions that are expressive, multilingual, and reliable across devices and real-time scenarios. The team owns the full training pipeline, from massive data curation and premium audio processing to frontier speech-language pre-training and intensive post-training to push quality, speed, and stability to the limit. The goal is to make talking to AI feel like conversing with a charming, kind, and knowledgeable person, and exceptionally smart, execution-oriented engineers are sought to achieve this.
$150k - $450k
Member of Technical Staff (EMEA)
fireworks ai
Fireworks is seeking a Member of Technical Staff for their EMEA field team. This role involves building on top of one of the world's largest inference platforms, spanning capabilities on inference and serving, tooling for fine-tuning and post-training of LLMs, and developing product features specific to the region. The ideal candidate is an excellent software engineer who prefers building product, with a strong emphasis on first principles thinking and the ability to reason about unfamiliar problems from the ground up. This is a flexible role designed for exceptional builders with 3 to 10 years of experience.
Compensation Partner
fireworks ai
Fireworks is seeking its first dedicated Compensation Partner to translate the company's compensation philosophy into practical application. This role will build the essential systems and processes for benchmarking, reviewing, and adjusting compensation as the organization grows rapidly. The Compensation Partner will be responsible for hands-on execution, including market research, managing review cycles, ensuring data accuracy, and serving as the primary resource for compensation-related inquiries from HRBPs and leaders. Additionally, this role will collaborate closely with Finance to develop and refine the company's equity framework, bringing a compensation perspective to financial planning.
Research Engineer, Audio and Speech
decagon
Decagon is seeking a Research Engineer focused on Audio and Speech to build and deploy the next generation of AI voice agents. This role involves developing models and agent harnesses for real-time, full-duplex conversational systems that can listen, reason, speak, and respond naturally. You will own projects end-to-end, from idea to production, making high-impact technical decisions and shipping real improvements in AI voice technology.
$200k - $400k
Senior Machine Learning Engineer
Cloudflare
You will help define how machine learning models run across Cloudflare’s global network, from frontier open LLMs and real-time voice models to customer-deployed models served on heterogeneous GPUs and next-generation accelerators. You’ll work with systems engineers, product teams, hardware partners, and AI/ML engineers to bring models into production with low latency, strong reliability, and efficient resource use. This role combines applied ML, inference optimization, evaluation, and production engineering, with a focus on benchmarking models, improving serving performance, validating quality, and building tooling that helps Cloudflare and its customers ship AI applications at Internet scale.
Senior Machine Learning Engineer
Cloudflare
We are seeking a visionary and hands-on Lead Machine Learning Engineer to join our Austin team. In this role, you will be the principal architect behind the next generation of our unified AI/ML platform, designing and building the infrastructure that powers everything from traditional predictive models to generative AI, large language models (LLMs), and autonomous agent frameworks. You will own the end-to-end technical strategy, blueprint, and execution of scalable backend services and data pipelines that support AI-driven applications across go-to-market, engineering, and product teams. Because our products are initiated and owned entirely within the team, you will drive the vision from initial requirements and system design to global deployment, optimization, and long-term evolutionary ownership.
Solutions Architect, Customer Success - US (Remote)
fiddler-ai
Fiddler is seeking a Solutions Architect, Customer Success to ensure clients achieve significant, measurable outcomes from their AI observability investments. In this role, you will act as both a technical expert and a trusted advisor, connecting complex Machine Learning (ML) and Large Language Model (LLM) systems with tangible business value. By guiding customers through the onboarding process, building integrations, and advocating for their needs internally, you will help them deploy trustworthy AI at scale, contributing to Fiddler's growth through successful adoption, renewals, and expansion.
$160k - $200k
Senior Software Engineer, GPU Infrastructure (HPC)
Cohere
Cohere is seeking a Staff Software Engineer to join our internal infrastructure team, responsible for building and operating world-class infrastructure and tools for training, evaluating, and serving Cohere's foundational AI models. You will work closely with AI researchers to support their AI workload needs on cutting-edge systems, focusing on stability, scalability, and observability. This role involves building and operating superclusters across multiple clouds, directly accelerating the development of industry-leading AI models. Participation in a 24x7 on-call rotation is required and compensated.
AI Deployment Strategist
fireworks ai
Fireworks AI is seeking an AI Deployment Strategist to serve as the technical backbone for customer relationships on their fast inference platform. This hybrid technical and commercial role involves bridging customer engineering teams with Fireworks' product, engineering, and applied AI teams to ensure successful onboarding, deep technical adoption, and long-term account health for companies running large-scale production AI workloads. You will be responsible for understanding customer needs, defining success criteria, managing the end-to-end relationship from technical evaluation through production, and acting as a "product manager for the customer's problem."
Senior Manager, Indirect Tax
fireworks ai
Fireworks is seeking an experienced Senior Manager, Indirect Tax to lead and manage the organization’s indirect tax function. This role will be responsible for indirect tax compliance, reporting, planning, audits, and risk management, while serving as a key advisor to Finance, Accounting, Legal, Operations, and other business stakeholders. The successful candidate will bring strong technical expertise in indirect taxation, excellent judgment, and the ability to translate complex tax requirements into practical business solutions. This role is well suited to a tax professional who can operate strategically while maintaining strong oversight of day-to-day compliance and execution.
Applied Machine Learning Engineer, EMEA
fireworks ai
Fireworks is seeking an Applied Machine Learning Engineer for the EMEA region. In this role, you will be the technical owner of customer engagements, embedding within client teams to understand their specific needs and challenges. You will be responsible for the entire lifecycle of a customer's deployment on the Fireworks platform, from initial scoping and model selection to ensuring production readiness, performance, and cost-efficiency. This position emphasizes first principles thinking and requires a deep understanding of software engineering, machine learning techniques, and infrastructure optimization.
Internship - Machine Learning Research Engineer
Perplexity AI
We are seeking motivated interns to join our Machine Learning Research Engineering team in Berlin for a 12-24 week full-time, in-person program. You will have the opportunity to significantly improve search quality by developing and optimizing large-scale deep learning models. This role involves conducting cutting-edge research in representation learning and building advanced RAG pipelines.
$12k - $24k