PyTorch Jobs

223 open roles mentioning PyTorch

AI Product Engineer - Nexus

1mo ago
f

fireworks ai

Fireworks is seeking a product-minded engineer to join their team and own the surfaces that make their AI platform, Nexus, a reality. Nexus is a drop-in platform designed to help engineering organizations reduce AI spend by routing workloads to high-performing open models. This role involves owning features end-to-end, from understanding problems and designing solutions to shipping and observing developer usage. You will be responsible for the intelligent router, cost-observability and enterprise-control layers, integrations, and core platform components like inference APIs and fine-tuning workflows. This is a high-autonomy individual contributor role on a small team, offering significant leverage and direct impact on a product that ships to production at scale.

San Mateo onsite FullTime
OpenAIAnthropicPython +5 more

Member of Technical Staff, AI Training Infrastructure

1mo ago
f

fireworks ai

Fireworks is seeking a Training Infrastructure Engineer to design, build, and optimize the infrastructure that powers large-scale model training operations. This role is crucial for developing high-performance AI training infrastructure, requiring collaboration with AI researchers and engineers to create robust training pipelines, optimize distributed training workloads, and ensure reliable model development. The position offers the opportunity to solve hard problems at the forefront of AI infrastructure, build what's next with bleeding-edge technology, and have a direct impact on the future of AI within a fast-growing, passionate team.

San Mateo hybrid FullTime
AWSAzureDocker +5 more

Technical Support Engineer (L2)

1mo ago
R

Runpod

Runpod is seeking a Technical Support Analyst (L2) to provide advanced technical assistance and resolve complex customer issues. This role is crucial for supporting customers on the AI Developer Cloud platform, which is used by over a million developers for AI experimentation, training, fine-tuning, and deployment. The ideal candidate thrives in a dynamic, remote-first environment and is passionate about delivering exceptional customer service. This position requires weekend availability as per business needs.

$70k - $93k

Remote - USA remote FullTime
DockerPythonJavaScript +5 more

Technical Support Engineer (L2)

1mo ago
R

Runpod

Runpod is seeking a Technical Support Analyst (L2) to provide advanced technical assistance and resolve complex customer issues. This role is crucial for supporting customers on the Runpod AI Developer Cloud platform, which is used by over a million developers for AI experimentation, training, fine-tuning, and deployment. The ideal candidate thrives in a dynamic, remote-first environment and is passionate about delivering exceptional customer service. This position requires availability on weekends and the candidate must be located in Ireland, the United Kingdom, Spain, or Poland.

€70k - €83k

Remote - EMEA remote FullTime
DockerPythonJavaScript +5 more

Endpoint Engineer, IT

1mo ago
thinkingmachines

thinkingmachines

Thinking Machines Lab is seeking an Endpoint Engineer to join their IT team. This role will focus on building secure infrastructure and efficient processes for an all-Mac environment, managing the endpoint fleet as a distributed platform with production-engineering practices. The team utilizes version-controlled workflows for endpoint configurations, security policies, scripts, and software deployments, incorporating testing, review, staged rollouts, and rollback capabilities. The Endpoint Engineer will collaborate closely with IT, Security, and Identity teams to ensure a secure and reliable employee computing experience.

$200k - $325k

San Francisco | New York onsite
OpenAIMistralAWS +4 more

Safety Operations Lead

1mo ago
thinkingmachines

thinkingmachines

Thinking Machines Lab is seeking an Operations Analyst to focus on ensuring product safety while supporting rapid product iteration. This role involves close collaboration with product engineers, security, researchers, and designers to integrate safety and integrity into the entire product lifecycle. Daily tasks will include managing the moderation queue and developing tools and policies to enhance the speed and accuracy of future moderation efforts.

$190k - $300k

San Francisco onsite
OpenAIMistralPyTorch +1 more

Tech Lead Manager, Inference

1mo ago
l

lumalabs

Luma is seeking a Tech Lead Manager for its Inference team to own the entire inference serving stack, encompassing routing, scheduling, and fleet-wide orchestration across thousands of GPUs, multiple clouds, and hardware vendors. This is a hands-on role where at least half of your time will be dedicated to architecting and building core platform components, making critical design decisions, and debugging complex incidents. You will also be responsible for leading, growing, and developing the inference engineering team, including hiring, coaching, and managing on-call rotations. The role involves setting the technical roadmap for serving infrastructure, owning platform SLOs and economics, and partnering with research to deploy new architectures and integrate serving into online RL and evaluation loops. The ideal candidate has extensive experience operating large-scale inference fleets and a genuine desire to remain hands-on in building and improving the serving stack.

$30k - $60k

Redwood City, CA hybrid FullTime
KubernetesPythonRust +3 more

Software Engineer, Inference

1mo ago
l

lumalabs

Luma is seeking a Software Engineer to own the serving of their models. This role involves integrating new architectures into the inference engine, scaling deployments across thousands of machines, and optimizing GPU fleet utilization while meeting internal service level objectives (SLOs). The work focuses on large-scale inference systems, including scheduling, fleet management, deployment pipelines, and reliability across various clusters and hardware providers. This position is ideal for a strong systems engineer experienced with model serving and Kubernetes at scale, rather than pure modeling.

$30k - $60k

Redwood City, CA hybrid FullTime
DockerKubernetesPython +4 more

Research Scientist / Engineer – Reinforcement Learning Infrastructure

1mo ago
l

lumalabs

Luma is seeking a Research Scientist / Engineer to build the systems that enable reinforcement learning (RL) at frontier scale. This role involves coupling policy optimization with large fleets of inference workers, agentic environments, and reward/verification systems to transform model behavior into learning signals. RL is crucial for Luma's models to evolve from capable to useful. Operating RL at scale is a complex systems challenge, encompassing training, rollout generation, environment execution, and reward computation across thousands of GPUs, demanding speed, stability, and correctness. This position is ideal for someone with hands-on experience in post-training LLMs with RL, building environments and verifiers, and debugging large-scale asynchronous rollout pipelines.

$30k - $60k

Redwood City, CA hybrid FullTime
KubernetesGoPyTorch +3 more

Research Scientist / Engineer — Foundation Model (Agent)

1mo ago
l

lumalabs

Luma is seeking a Research Scientist/Engineer to build and train large-scale multimodal agentic models capable of reasoning, planning, coding, and tool utilization for complex, multi-step tasks involving pixels. This core research role spans modeling, data, systems, and evaluation, tackling novel problems where both scientific and engineering expertise are equally crucial. The ideal candidate will have a strong foundation in large-scale model training and agentic systems, and is comfortable working across multiple layers of the technology stack.

$30k - $60k

Redwood City, CA hybrid FullTime
PyTorch

Applied Machine Learning Engineer, Singapore

1mo ago
f

fireworks ai

As an Applied Machine Learning Engineer, you will serve as a vital bridge between cutting-edge AI research and practical, real-world applications. Your work will focus on developing, fine-tuning, and operationalizing machine learning models that drive business value and enhance user experiences. This is a hands-on engineering role that combines deep technical expertise with a strong customer focus to deliver scalable AI solutions.

Singapore onsite FullTime
PythonFine-TuningPyTorch +4 more

AI Field Engineer, Singapore

1mo ago
f

fireworks ai

Fireworks is seeking an AI Field Engineer to join their team in Singapore. This role is at the forefront of technical engagement, embedding with key customers and partners to transform complex AI challenges into production-ready systems. You will operate at the nexus of engineering, product development, and customer delivery, actively building proofs-of-concept (POCs), minimum viable products (MVPs), and production integrations. Simultaneously, you will engage in high-level discussions regarding architecture, strategy, and business impact. The position requires a blend of hands-on coding, performance benchmarking, debugging, and deployment architecture, alongside leading customer discovery, aligning stakeholders, and translating customer needs into product enhancements.

Singapore onsite FullTime
AWSAzureKubernetes +5 more

AI Field Engineer, EMEA

1mo ago
f

fireworks ai

Fireworks is seeking an AI Field Engineer to join our team in EMEA. This role is at the forefront of our technical engagement with ambitious customers and technology partners, focusing on transforming complex AI challenges into production-ready systems. You will operate at the intersection of engineering, product, and customer delivery, actively building Proofs of Concepts (POCs), Minimum Viable Products (MVPs), and production integrations. This position requires a blend of hands-on technical execution and the ability to engage in executive-level discussions regarding architecture, strategy, and business outcomes. The role emphasizes building strong relationships and trust through in-person interactions with clients, particularly within large organizations and digital-native companies adopting GenAI.

London onsite FullTime
AWSAzureKubernetes +5 more

Software Engineer, Full Stack

1mo ago
thinkingmachines

thinkingmachines

Thinking Machines Lab is seeking a full stack engineer to build and ship products from prototype to scale, and to maintain tools that accelerate research and product teams. This role involves working across frontend and backend components, and contributing to the reliability, observability, and security of production systems. The company's mission is to empower humanity through advancing collaborative general intelligence, building a future where everyone has access to AI tools for their unique needs.

$350k - $475k

San Francisco onsite
OpenAIMistralPython +3 more

Software Engineer, Full Stack, Tinker

1mo ago
thinkingmachines

thinkingmachines

Thinking Machines Lab is seeking a full-stack engineer to develop and deploy the products and services that Tinker users engage with daily. This role involves working across frontend, backend, and infrastructure to build the Tinker console, developer tools, and other essential components for the platform. Tinker is a fine-tuning API that enables researchers and developers to customize frontier AI models using their own data and algorithms, managing the underlying infrastructure to provide flexibility and access to advanced capabilities.

$350k - $475k

San Francisco onsite
OpenAIMistralPython +5 more

Software Engineer, Developer Productivity, AI Tools

1mo ago
thinkingmachines

thinkingmachines

Thinking Machines Lab is seeking a developer productivity engineer to enhance internal software development processes, focusing on safety, speed, and user experience. This role will concentrate on AI tools and coding agents, collaborating with platform, security, and product engineers to build cutting-edge tooling for AI-assisted software development and significantly accelerate the inner development loop. The position involves both establishing company-wide platforms and assisting individual developers in optimizing their workflows.

$350k - $475k

San Francisco onsite
OpenAIMistralDocker +5 more

Software Engineer, Data Infrastructure

1mo ago
thinkingmachines

thinkingmachines

Thinking Machines Lab is seeking an engineer to join a high-impact team focused on data infrastructure. This role is crucial for architecting and scaling the core systems that power distributed training pipelines, multimodal data catalogs, and intelligent processing of petabytes of data. You will work directly with researchers to accelerate experiments, develop new datasets, enhance infrastructure efficiency, and derive key insights from our data assets. If you are passionate about distributed systems, large-scale data mining, and building foundational tools from the ground up, we encourage you to apply.

$350k - $475k

San Francisco onsite
OpenAIMistralPython +5 more

Site Reliability Engineer (SRE)

1mo ago
thinkingmachines

thinkingmachines

Thinking Machines Lab is seeking a Site Reliability Engineer (SRE) to ensure the end-to-end reliability of their Tinker platform. This role involves working closely with engineers and research teams to enhance the robustness and resilience of every system layer. The SRE will be instrumental in maintaining and improving the infrastructure that supports custom AI model fine-tuning, ensuring a seamless experience for researchers and developers.

$350k - $475k

San Francisco onsite
OpenAIMistralKubernetes +3 more

Staff Research Engineer - Multimodal Generative Modelling

1mo ago
s

synthesia.io

Synthesia is seeking a Staff Research Engineer to join their Voice team and contribute to the company's long-term vision of building the best human-interactive models. These advanced systems will go beyond simple conversation to perceive, respond to, and react to user actions and emotions in real-time. The role involves defining and driving this broader vision across teams, proposing ambitious research directions, and taking ownership of critical component design and implementation. You will collaborate closely with the voice lead, video teams, and other senior members to develop models that combine text, audio, and video for a seamless, interactive experience, moving beyond traditional turn-taking in speech models.

Europe remote FullTime
Fine-TuningPyTorchDeep Learning +1 more

Research Engineer, Infrastructure, Inference

1mo ago
thinkingmachines

thinkingmachines

Thinking Machines Lab is seeking an infrastructure research engineer to design, optimize, and scale the systems that power large AI models. The goal is to make inference faster, more cost-effective, more reliable, and more reproducible, enabling research teams to focus on advancing model capabilities. This role is crucial for ensuring that every experiment, evaluation, and deployment runs smoothly at scale, with a focus on performant and efficient model inference for both real-world applications and research acceleration.

$350k - $475k

San Francisco onsite
OpenAIMistralKubernetes +3 more

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.