Kubernetes Jobs

410 open roles mentioning Kubernetes

Software Engineer, Compute Infrastructure

1mo ago
OpenAI

OpenAI

We are seeking engineers to build the compute platform that powers OpenAI's research and products. This role involves designing, provisioning, scheduling, operating, and optimizing systems that connect accelerators, CPUs, networks, storage, data centers, and orchestration software into a cohesive experience for researchers and product teams. You will work across the entire stack, from capacity planning and cluster lifecycle to deep system optimization and developer experience, aiming to improve research velocity through enhancements in communication, scheduling, hardware efficiency, and debugging workflows. The specific focus will be matched to your strengths and interests, whether that's close to hardware, users, CaaS, agent infrastructure, or the control and data planes in between.

San Francisco hybrid FullTime
OpenAIKubernetesGo

Senior Manager, Engineering

1mo ago
Harvey

Harvey

Harvey is seeking an Engineering Manager to lead its Core Infrastructure team in Bengaluru. This role is crucial for scaling the platform, team, and future Harvey products, focusing on building new systems and hardening existing ones with an emphasis on architecture, scale, and long-term platform leverage. The Core Infrastructure team is responsible for delivering the foundational systems that power all user interactions across Harvey's legal AI platform, processing billions of prompt tokens and millions of daily requests. This is a high-leverage position for a builder-leader who can grow engineers, drive technical execution, and enhance reliability, performance, security, and developer velocity.

Bengaluru hybrid FullTime
AWSAzureKubernetes +3 more

Data Engineer - Axion

1mo ago
A

Applied Intuition

Applied Intuition is a rapidly growing company focused on powering the future of physical AI for various industries including automotive, defense, and construction. We are building the digital infrastructure to bring intelligence to moving machines. Our solutions are trusted by leading global automakers and the U.S. military. We are seeking a Data Engineer to join our team and contribute to the development of our ML infrastructure and data pipelines.

$60k - $300k

Sunnyvale onsite FullTime
DockerKubernetesPython +3 more

Engineering Manager, AI Platform

1mo ago
c

crusoe

As an Engineering Manager on the Managed AI team, you will lead and scale a team of engineers building our next-generation platform for the full lifecycle of Large Language Models (LLMs). You will guide the team through the design and implementation of highly scalable, fault-tolerant infrastructure, combining technical expertise with strong people leadership. This role is central to a fast-growing, strategically important organization where you will shape the engineering roadmap and drive the execution of key projects. Success requires close partnership with product, business, and platform stakeholders to deliver a performant and reliable platform that powers AI for customers globally.

$215k - $260k

San Francisco, CA - US onsite FullTime
KubernetesPythonGo +2 more

Solutions Engineer (Texas)

1mo ago
L

Langchain

LangChain is seeking a Solutions Engineer to join our Deployed Engineering team. This role is a hands-on, highly technical position focused on partnering with account executives to guide customers from initial technical conversations through production rollout of AI agents. You will own the technical win by scoping evaluations, designing proof-of-concepts (POCs) that mirror real-world workloads, and answering complex architecture questions. The role sits at the intersection of engineering, product, and sales, with feedback from the field directly shaping customer adoption and future product development. You will work on challenging applied AI problems in a fast-paced environment, with a direct impact on closed deals and live deployments.

$200k - $250k

Dallas, TX remote FullTime
LangGraphLangChainAWS +5 more

Member of Technical Staff, Agentic Environments

1mo ago
Cohere

Cohere

Cohere is seeking a senior engineer to join our team, focusing on the practical challenges of deploying AI systems at scale in production environments. This hands-on, engineering-driven role involves working with frontier AI models, building scalable solutions, and bridging research concepts with real-world implementations. You will contribute to both engineering and research efforts, designing and writing high-performing software for model training, developing new tools to support LLM research, and collaborating with various engineering and scientific teams. We provide access to world-class compute resources, data, and talent to enable you to do your best work.

Europe remote FullTime
CohereKubernetesPython +3 more

Senior Staff Software Engineer, DC Infrastructure

1mo ago
c

crusoe

Crusoe is seeking a highly skilled Software Engineer to join their Data Center Infrastructure Engineering team. This role focuses on developing software for managing a fleet of GPU servers and the data centers that house them. The position involves creating and implementing advanced diagnostic, observability, automation, and repair tools for high-performance GPU compute clusters. The ideal candidate will be a hands-on problem solver, comfortable working independently, and will play a crucial role in maintaining the health and scalability of Crusoe's rapidly expanding GPU fleet.

$250k - $300k

San Francisco, CA - US onsite FullTime
KubernetesPythonGo +5 more

Staff Software Engineer, DC Infrastructure

1mo ago
c

crusoe

Crusoe is seeking a highly skilled and motivated Software Engineer to join its Data Center Infrastructure Engineering team. This role focuses on developing software for managing a fleet of GPU servers and the data centers that house them. The position involves creating and implementing advanced diagnostic, observability, automation, and repair tools for high-performance GPU compute clusters. The ideal candidate will be a hands-on problem solver, comfortable working independently, and will play a crucial role in maintaining the health and scalability of Crusoe's rapidly expanding GPU fleet.

$215k - $260k

San Francisco, CA - US onsite FullTime
KubernetesPythonGo +5 more

AI Performance Engineer

1mo ago
A

Applied Intuition

Applied Intuition is seeking a performance engineer to specialize in making large-scale machine learning workloads fast and cost-efficient within the datacenter. This role focuses on optimizing distributed training runs across multiple nodes and high-throughput batch inference for processing vast amounts of real-world autonomy logs. The primary goal is to improve throughput, cluster efficiency, and reduce cost per unit of data processed, directly impacting the company's iteration speed. You will be responsible for identifying and resolving performance bottlenecks across the entire stack, from accelerators to ML frameworks and data infrastructure, working collaboratively with various engineering teams to achieve significant improvements in training time and processing costs.

$60k - $300k

Sunnyvale onsite FullTime
KubernetesPythonGo +4 more

Forward Deployed Engineer, Infrastructure Specialist (Singapore)

1mo ago
Cohere

Cohere

Cohere is a leading enterprise AI company building cutting-edge foundation AI models and end-to-end products. We are seeking engineers passionate about Agentic AI to join our team and shape how enterprises harness AI in real-world applications. As a Forward Deployed Engineer, you will act as a bridge between our core North product and client engineering teams, solving complex problems and securely integrating AI into critical sectors. This role involves working at the forefront of AI deployment, ensuring client success and driving the adoption of advanced AI solutions.

$20k - $40k

Singapore remote FullTime
CohereAWSAzure +3 more

Engineering Site Lead

1mo ago
P

Perplexity AI

Perplexity is seeking an exceptional Site Lead to establish and scale its London office, a strategic presence in one of the world's leading tech hubs. This role involves building teams and culture from the ground up while driving technical excellence in infrastructure and AI systems. The Site Lead will serve as the face of Perplexity in London, responsible for building the technical organization, fostering a world-class engineering culture, and directly managing infrastructure teams. This position reports to senior leadership and collaborates cross-functionally with global teams.

$15k - $30k

London onsite FullTime
AWSAzureKubernetes +3 more

Software Engineer - AI Developer Productivity

1mo ago
B

Baseten

Baseten is seeking an AI Developer Productivity Engineer to build the internal platform that empowers engineers to work in an AI-first manner. This role involves creating the agent configurations, tooling, and evaluation frameworks that make AI agents competent and easy to use within our codebase. You will be responsible for making the "good path" for AI adoption the easiest path, focusing on shipping infrastructure and measuring its impact. The goal is to enable teams to adopt these AI tools because they are superior to self-assembled solutions, not due to mandates. You will write the playbook for an AI-first Software Development Lifecycle (SDLC) at Baseten, driving adoption through excellent developer experience, clear documentation, and low friction.

San Francisco hybrid FullTime
DockerKubernetesPython +3 more

Forward Deployed Engineer, Infrastructure Specialist (South Korea)

1mo ago
Cohere

Cohere

Cohere is seeking a Forward Deployed Engineer, Infrastructure Specialist to join our team in South Korea. This role is crucial for shaping how enterprises leverage AI through our cutting-edge North platform. You will act as a bridge between our core product and client engineering teams, tackling complex problems and integrating AI securely into critical sectors. We are looking for engineers passionate about customer success and working at the forefront of Agentic AI, with an anticipated 20-40% travel requirement.

$20k - $40k

Korea remote FullTime
CohereAWSAzure +3 more

Senior Frontend Engineer - Arya

1mo ago
s

sarvam

Sarvam.ai is at the forefront of India’s AI revolution, dedicated to building transformative technologies that empower users and redefine digital interactions. Our mission is to develop intelligent systems that seamlessly integrate into everyday workflows, enhancing productivity and user experience. We are looking for a Senior Frontend Engineer to join our team building cutting-edge AI agent interfaces. You will architect and deliver production-grade web applications using Next, React, and TypeScript in a modern monorepo architecture. This role requires high autonomy, the ability to ship fast without compromising quality, and thriving in ambiguity while building AI-powered products at scale.

Bengaluru onsite FullTime
DockerKubernetesTypeScript +1 more

Member of Technical Staff, Enterprise Foundations

1mo ago
f

fireworks ai

Enterprise Foundations is responsible for building the core capabilities that large enterprises require to operate their business on the Fireworks platform. This involves addressing complex needs related to organizational structure, user access control, usage metering, billing, data security, and deployment within customer-owned cloud environments. The role offers end-to-end ownership of features, starting from customer conversations, evolving into changes in core data models and APIs, and extending through the control plane, training and inference stacks, SDK, CLI, and console. The work is foundational, often involving new primitives that impact other teams' code and must be rolled out to production without disruption. This position is closely tied to revenue, with requirements stemming directly from enterprise deals and customer interactions, offering a unique opportunity to influence product direction and build solutions for scaled companies.

New York hybrid FullTime
AWSKubernetesPython +5 more

Software Engineer, Infrastructure

1mo ago
C

Cognition

Cognition is an applied AI lab building end-to-end software agents, known for creating Devin, the first AI software engineer, and Windsurf, an AI-native IDE. Our vision is AI that acts as a genuine teammate to engineers. We are a small, talent-dense team with a background in competitive programming, founding startups, and leading AI research from top companies. We are seeking Infrastructure Engineers to build and maintain the critical systems that power our AI products, ensuring reliability, scale, and a seamless developer experience.

$260k - $300k

San Francisco onsite FullTime
AWSAzureKubernetes +3 more

Deployed Engineer (Chicago)

1mo ago
L

Langchain

LangChain is seeking a Deployed Engineer in Chicago to join their Deployed Engineering team. This role focuses on working directly with companies to build and run AI agents in production, transforming prototypes into reliable systems. You will partner closely with customer engineers throughout the entire lifecycle, from pre-sales evaluations to post-deployment advisory. The position involves co-designing agent architectures, helping customers operate agents at scale using the LangChain suite, and contributing to the product's evolution based on real-world insights. This is a hands-on, highly technical role at the intersection of engineering, product, and go-to-market.

$200k - $250k

Chicago, IL remote FullTime
LangGraphLangChainAWS +5 more

Principal Systems Software Engineer

1mo ago
c

crusoe

Crusoe is seeking a Principal Systems Software Engineer to lead the vision for its next-generation AI infrastructure. This role is for an industry expert with hyperscale experience, tasked with redefining the I/O path for generative AI. You will design the core fabric that unifies Bare-Metal-as-a-Service, Intelligent IaaS, and Elastic CaaS into a high-performance pool of intelligence. This position involves bridging silicon and software, advising executive leadership on hardware/software co-design, and leading R&D teams in shipping production-grade kernel and orchestration code. The ideal candidate is a master of the I/O path, capable of pushing massive-scale training workloads to hardware limits.

$260k - $340k

San Francisco, CA - US onsite FullTime
AWSAzureKubernetes +1 more

Senior Software Engineer (DCIE)

1mo ago
c

crusoe

Crusoe is seeking a highly skilled Software Engineer to join their Data Center Infrastructure Engineering team. This role focuses on developing software for managing a fleet of GPU servers and the data centers that house them. The position involves creating and implementing advanced diagnostic, observability, automation, and repair tools for high-performance GPU compute clusters. The ideal candidate is a hands-on problem solver who can work independently and play a critical role in maintaining the health and scalability of Crusoe's rapidly growing GPU fleet.

$170k - $205k

San Francisco, CA - US onsite FullTime
KubernetesPythonGo +5 more

Senior Solution Engineer

1mo ago
L

Lambda

Lambda, a leader in AI cloud infrastructure, is seeking a Senior Solution Engineer to join their growing team. This role is crucial in enabling customers to achieve their business goals with AI infrastructure by partnering with leading AI researchers and enterprise engineering teams. The ideal candidate will design, scale, and optimize high-performance GPU cloud solutions, turning complex compute challenges into seamless, production-ready AI infrastructure. This position requires a customer-first mindset and technical mastery to drive growth and customer success.

San Francisco Office (Second St) hybrid FullTime
DockerKubernetesPython +5 more

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.