Kubernetes Jobs

411 open roles mentioning Kubernetes

Security Operations Lead

2mo ago
Replit

Replit

Replit is seeking a Security Operations Lead (SOC Lead) to establish, enhance, and manage a 24/7 detection and response capability within a modern, cloud-native, and AI-driven environment. This leadership role will oversee global SOC functions including monitoring, SIEM management, detection engineering, alert triage, and operational readiness. A key aspect of this position involves evaluating and integrating emerging AI-based SOC products and autonomous response platforms. The role requires monitoring across multi-cloud environments (primarily GCP, with AWS/Azure secondary), Kubernetes, SaaS services, endpoints, developer tools, and AI workloads. Collaboration with Cloud Security, Compliance/GRC, SRE, Platform Engineering, IT/Endpoint teams, and AI Infrastructure is essential to ensure the detection strategy scales effectively against evolving threats. This is a hands-on leadership opportunity to shape the future of SOC operations while addressing complex challenges in a high-scale AI setting.

Foster City, CA hybrid FullTime
AWSAzureKubernetes +3 more

Director of Infrastructure Engineering

2mo ago
R

Runpod

Runpod is seeking a Director of Infrastructure Engineering to lead and scale its core cloud and bare-metal environments. This pivotal role will oversee critical foundational layers including Site Reliability Engineering (SRE), global networking, High-Performance Computing (HPC) networks, and distributed storage engines. The ideal candidate will establish the operating rhythm, culture, and technical direction necessary to ensure Runpod's platform remains highly available, performant, and capable of meeting massive GPU computing demands. This position requires close partnership with Product Engineering, Product, and GTM leadership to support enterprise customers, focusing on delivering the fastest, most reliable, and lowest-latency infrastructure for large-scale AI workloads.

$225k - $325k

Remote - USA remote FullTime
KubernetesDeep LearningTerraform

Senior Software Engineer, Cloud Monitoring Service

2mo ago
c

crusoe

Crusoe is seeking a Senior Software Engineer to build and own customer-facing observability products within Crusoe Cloud's Cloud Monitoring Service. This team delivers managed logs and self-service observability capabilities, enabling customers to collect, search, and act on their telemetry efficiently. You will be responsible for end-to-end ownership of product surfaces, including APIs, backend services, and the user onboarding and configuration experience, with a strong emphasis on reliability and ease of use.

$175k - $210k

San Francisco, CA - US onsite FullTime
DockerKubernetesGo +4 more

Senior Customer Success Manager, Managed Inference

2mo ago
c

crusoe

We are seeking a highly motivated and skilled Senior Customer Success Manager with a strong background in customer engagement and a deep technical understanding of cloud computing, AI, and ML. The ideal candidate has experience supporting customers running production AI inference workloads and understands the operational, technical, and business challenges associated with deploying and scaling AI applications. Experience supporting Managed Inference, model serving platforms, LLM deployments, AI agents, or GPU-based inference environments is highly preferred. This role is pivotal in ensuring that our clients maximize the value of our solutions, guiding them through the technical complexities and empowering them with the tools and knowledge to achieve their business and sustainability goals.

$190k - $215k

Denver, CO - US onsite FullTime
KubernetesRAGAI Agents +1 more

Senior Customer Success Manager, Managed Inference

2mo ago
c

crusoe

We are seeking a highly motivated and skilled Senior Customer Success Manager with a strong background in customer engagement and a deep technical understanding of cloud computing, AI, and ML. The ideal candidate has experience supporting customers running production AI inference workloads and understands the operational, technical, and business challenges associated with deploying and scaling AI applications. Experience supporting Managed Inference, model serving platforms, LLM deployments, AI agents, or GPU-based inference environments is highly preferred. This role is pivotal in ensuring that our clients maximize the value of our solutions, guiding them through the technical complexities and empowering them with the tools and knowledge to achieve their business and sustainability goals.

$190k - $215k

San Francisco, CA - US onsite FullTime
KubernetesRAGAI Agents +1 more

Staff Engineer, Agentic

2mo ago
Inflection AI

Inflection AI

Inflection AI is seeking a Staff Engineer, Agentic to own the platforms, systems, and services that bring conversational AI to life at scale. This role involves collaborating across research, product, and infrastructure teams to enable rapid iteration, high reliability, and secure delivery of novel AI features to millions of users. Your work will directly impact both the pace of product development and the stability of our production systems, focusing on human-centered, emotionally intelligent AI.

$10k - $15k

Palo Alto, California, United States onsite
OpenAIAnthropicLangGraph +5 more

Software Engineer, Product

2mo ago
pika

pika

We are seeking a Staff Software Engineer, Product to help shape and scale the user-facing systems at Pika. In this key role, you will bridge advanced engineering and product experience, taking ownership of the architecture and delivery of Pika’s product features—from intuitive user interfaces and collaboration tools to real-time interactions and generative AI-powered workflows. As a Staff-level engineer, you will lead the design and implementation of robust, user-centric systems that enable millions of users to create and collaborate seamlessly with our AI-powered platform. Your high-impact architectural and product decisions will directly shape the end-user experience. You will also play a critical role in elevating the technical and product bar through architectural proposals, code reviews, technical mentorship, and cross-functional collaboration with product, design, and engineering teams.

Palo Alto HQ onsite FullTime
AWSDockerKubernetes +5 more

Software Engineer, Machine Learning

2mo ago
M

Mercor

As a Machine Learning Engineer on the Marketplace team, you will build the models and decision systems that power Mercor's hiring engine. This includes search and ranking, candidate-job matching, marketplace recommendations, personalization, and allocation decisions across a rapidly growing talent network. This is an applied ML role with direct product and revenue impact, working on problems shaped by real marketplace constraints such as sparse and delayed labels, cold start, noisy feedback, heterogeneous supply and demand, and the need to optimize across speed, quality, and conversion simultaneously.

San Francisco onsite FullTime
KubernetesPythonGo +5 more

Software Engineering Manager, Database (SmithDB)

2mo ago
L

Langchain

LangChain is seeking a hands-on Engineering Manager to lead the SmithDB team, responsible for building a storage and query layer purpose-built for AI observability and evaluation. This role is a blend of technical leadership and people management, requiring the individual to write production code, review PRs at the systems level, and make architectural decisions, while also owning team health, hiring, and the technical roadmap. The ideal candidate possesses deep systems or database engineering experience and a passion for staying technical while growing a team, contributing to a greenfield system with significant production load and ambitious engineering goals.

$240k - $300k

San Francisco, CA onsite FullTime
LangGraphLangChainAzure +3 more

Platform Engineer - AI Infrastructure

2mo ago
s

sarvam

Sarvam is building the bedrock of Sovereign AI for India, developing a full-stack AI platform across research, models, infrastructure, and applications. This role focuses on building the platform that sits on top of a large, multi-vendor GPU fleet, serving demanding training and inference workloads. You will design and ship control-plane services, scheduler integrations, autoscaling controllers, inference-serving platforms, RBAC and quota systems, observability and cost tooling, and the CLI and APIs that ML engineers use daily. The work is heavy software engineering, treating the platform as a product with internal users as customers.

Bengaluru onsite FullTime
KubernetesPythonGo +2 more

Designated Technical Support Engineer

2mo ago
G

Glean

Glean is seeking a Designated Technical Support Engineer to join our rapidly expanding startup. We are building a modern knowledge assistant personalized to every employee in an organization, making all information within the company accessible, contextual, and fresh. This role is customer-obsessed, providing both proactive and reactive support to our growing customer base to ensure the best customer experience in the industry. This position requires dedication to select customers, including additional background screenings, clearances, training, certification, and carrying customer-provided equipment. Extended on-call shifts may be required based on customer contractual obligations.

Bangalore, India onsite
AWSAzureKubernetes +4 more

Machine Learning Engineer, Reliability

2mo ago
Fal

Fal

fal is building the generative media ecosystem for the next generation of AI products, providing the infrastructure, tools, and model access needed to scale from idea to production. As generative media reshapes industries, fal is becoming the foundation for ambitious teams. This hybrid ML Engineering / Site Reliability Engineering role will own the reliability, security, and safety of fal's generative media model APIs, ensuring they remain available, performant, secure, and safe for thousands of developers and enterprises. You will address model-specific failure modes, such as degraded output quality, drift, unsafe generations, and abuse patterns, as critical reliability concerns alongside uptime and latency.

Remote - APAC remote FullTime
KubernetesPythonTransformers +1 more

Sr. Staff Software Engineer, Managed Platform Services

2mo ago
c

crusoe

Crusoe is seeking Sr. Staff Software Engineers to act as senior floating technical leaders within the Managed Platform Services (MAPS) organization. This role involves deploying expertise across various critical areas, including infrastructure scale-out, performance and reliability enhancements, new product exploration, customer-facing platform development, and cross-functional initiatives. The objective is to transform ambiguity into clarity, prototypes into products, and individual team solutions into organization-wide leverage, driving the company's mission to accelerate the abundance of energy and intelligence.

$250k - $300k

San Francisco, CA - US onsite FullTime
KubernetesGo

Senior Software Engineer - Agentic Tooling & Productivity

2mo ago
Scale AI

Scale AI

Scale AI is seeking a Senior Software Engineer to design, build, and operate secure, scalable infrastructure that empowers employees. You will join a team focused on automating identity and access management, endpoint management, and the broader SaaS stack. The role involves leveraging internal and external models to create applications, slackbots, and dashboards for internal users, streamlining complex workflows through automation, and significantly enhancing the velocity of G&A teams. As a full-stack engineer, you will contribute to building internal and external tools, developing cloud-native distributed systems, integrating with SaaS platforms, and ensuring product quality through testing and debugging.

$216k - $270k

San Francisco, CA remote
AWSAzureDocker +12 more

Senior Cloud Support Engineer

2mo ago
c

crusoe

Crusoe Cloud is seeking a Senior Cloud Support Engineer to empower customers utilizing their sustainable, low-cost GPU compute power for AI/ML, physics simulations, and computational biology. As the primary technical support contact, you will ensure customers can seamlessly leverage Crusoe Cloud, directly impacting the company's mission to accelerate AI research and development. This role involves working with cutting-edge technologies and collaborating with a talented team to solve complex challenges, making it an ideal opportunity for a motivated and experienced technical professional passionate about customer success and Crusoe's values.

Dallas, TX - US onsite FullTime
AWSAzureKubernetes +2 more

Senior Cloud Support Engineer

2mo ago
c

crusoe

Crusoe Cloud is seeking a Senior Cloud Support Engineer to empower customers utilizing their sustainable, low-cost GPU compute power for AI/ML, physics simulations, and computational biology. As the primary technical support contact, you will ensure customers can seamlessly leverage Crusoe Cloud, directly impacting the company's mission to accelerate AI research and development. This role involves working with cutting-edge technologies and collaborating with a talented team to solve complex challenges, making it an ideal opportunity for a motivated and experienced technical professional passionate about customer success and Crusoe's values.

Denver, CO - US onsite FullTime
AWSAzureKubernetes +2 more

Deployment Engineering Manager, Enterprise

2mo ago
Scale AI

Scale AI

Scale AI is seeking an Infrastructure Engineering Manager to lead a team focused on the full deployment lifecycle of AI applications in live customer environments. This role bridges the gap between AI research and production, transforming innovative prototypes into scalable, high-performance enterprise solutions. The team works on interactive AI applications, enterprise SaaS products, and platform capabilities, ensuring deployments are production-ready, secure, and built to last. The manager will drive improvements through customer feedback, automation, and scalable playbooks, focusing on complex technical problems at the intersection of customer impact and platform excellence.

$216k - $270k

San Francisco, CA; New York, NY onsite
AWSAzureKubernetes +7 more

Staff Front-End (React) Engineer

2mo ago
N

Nabla

Nabla is seeking a Staff Front-End Engineer to lead the development of high-impact user experiences across our web, desktop, browser extension, and mobile interfaces. You will be instrumental in setting the standard for UI quality and performance, and will play a central role in evolving our design system, architecture, and tooling. This role offers the opportunity to contribute end-to-end to a specific product area within a cross-functional squad, working closely with Product, ML, and Back-End teams to ship full-stack features. We are a fast-moving, deeply technical engineering team leveraging AI to revolutionize healthcare delivery.

Paris office onsite FullTime
KubernetesPythonTypeScript +5 more

Forward Deployed Engineer, Infrastructure Specialist

2mo ago
Cohere

Cohere

Cohere is seeking a Forward Deployed Engineer, Infrastructure Specialist to join their team. This role focuses on shaping how enterprises harness AI in real-world applications, acting as a bridge between Cohere's North AI workspace platform and client engineering teams. You will be at the forefront of solving complex problems and securely integrating AI into critical sectors like finance, healthcare, and telecommunications. The ideal candidate will be passionate about working at the cutting edge of Agentic AI and deeply care about customers.

$20k - $40k

Tokyo remote FullTime
CohereAWSAzure +3 more

Designated Technical Support Engineer

2mo ago
G

Glean

Glean is seeking a Designated Technical Support Engineer to join its growing startup. Glean offers a Work AI platform that enhances productivity through intelligent search, an AI assistant, and AI agents. The platform provides enterprise-grade infrastructure for managing AI across businesses, with extensive connectors and LLM flexibility. The company is recognized for its innovation in AI and is expanding globally. The Designated Technical Support Engineer will act as a key technical resource for customers, ensuring a high level of service and customer satisfaction.

$120k - $190k

New York, NY onsite
AWSAzureKubernetes +4 more

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.