Kubernetes Jobs

410 open roles mentioning Kubernetes

Staff / Principal Machine Learning Engineer, Serving - Switzerland

5mo ago
i

inworld

Inworld is seeking a Staff/Principal Machine Learning Engineer to join their research lab focused on building top-ranked real-time voice models. These models power large consumer-facing AI applications across various sectors. The role involves optimizing real-time inference, developing best-in-class APIs and products, and contributing to the research and development of state-of-the-art models. The ideal candidate is a fast learner who thrives in ambiguity and can demonstrate a strong portfolio of built, broken, and understood systems. This position emphasizes impact, shipping stable code, and a deep understanding of the underlying logic behind engineering decisions.

Switzerland remote FullTime
KubernetesPythonGo +2 more

Staff / Principal Machine Learning Engineer, Serving - UK

5mo ago
i

inworld

Inworld is seeking a Staff/Principal Machine Learning Engineer to join their research lab focused on building top-ranked real-time voice models. These models power large consumer-facing AI applications across various sectors, reaching hundreds of millions of end-users. The role involves optimizing real-time inference, developing state-of-the-art models, and creating best-in-class APIs and products. The ideal candidate thrives in ambiguity, learns quickly, and can demonstrate a strong track record of building, breaking, and understanding complex systems. This position emphasizes impact, shipping stable and reliable systems, and a deep understanding of the underlying logic behind engineering decisions.

£140k - £200k

UK onsite FullTime
KubernetesPythonGo +2 more

Senior / Lead Machine Learning Engineer, Serving - Serbia

5mo ago
i

inworld

Inworld is a research lab building top-ranked real-time voice models used in large consumer-facing AI applications. We are seeking a Senior/Lead Machine Learning Engineer to optimize real-time inference and create best-in-class APIs and products. This role involves tackling complex, ambiguous problems and driving impact through shipped, stable systems. We value engineers who are comfortable with ambiguity, have a bias for action, and obsess over performance, latency, and reliability as core product features.

Serbia onsite FullTime
KubernetesPythonGo +2 more

Senior / Lead Machine Learning Engineer, Serving - Germany

5mo ago
i

inworld

Inworld is seeking a Senior / Lead Machine Learning Engineer to join our team in Germany. We are a research lab building top-ranked real-time voice models used in large-scale AI applications across various sectors. Our work involves research, development, optimizing real-time inference, and creating best-in-class APIs and products. We are looking for fast learners with a strong ability to build, break, and understand systems, who thrive in ambiguity and can demonstrate their past work. This role emphasizes full-cycle ownership, taking models from research to production-ready serving.

Germany onsite FullTime
KubernetesPythonGo +2 more

Staff / Principal Machine Learning Engineer, Serving - USA

5mo ago
i

inworld

Inworld is seeking a Staff/Principal Machine Learning Engineer to join their team in Mountain View, USA. This role focuses on optimizing real-time inference for state-of-the-art AI models, which power large consumer-facing applications. The ideal candidate will have a strong background in systems programming and a proven ability to take models from research to production, ensuring reliability and performance. The company values engineers who can navigate ambiguity, drive impact, and deeply understand the underlying logic of their work, fostering a collaborative and fast-paced environment.

$270k - $500k

Mountain View, California, USA hybrid FullTime
KubernetesPythonGo +2 more

Staff Security Engineer, Infrastructure

5mo ago
Fal

Fal

We are seeking a Security Engineer, Infrastructure to safeguard the core systems powering fal.ai's platform, including GPU compute, multi-cloud environments, networking, and data pipelines. This role involves hands-on, systems-oriented work across the full stack, from cloud and Kubernetes to identity, networking, and secrets. You will be responsible for designing and implementing scalable security controls at the intersection of security, infrastructure, and distributed systems, ensuring the integrity and safety of our high-performance AI platform.

San Francisco onsite FullTime
AWSAzureDocker +5 more

Product Manager, Inference Platform

5mo ago
B

Baseten

Baseten is seeking a Product Manager to define and build the product function for its inference platform, which powers mission-critical AI inference for leading AI companies. This role offers a unique opportunity to shape the future of AI infrastructure, working directly with founders and top engineers. You will own the product surface responsible for making AI model inference fast, reliable, and economical at scale, including autoscaling, traffic routing, failover, and workload scaling across clusters and regions. The ideal candidate is deeply technical, customer-obsessed, and enjoys tackling foundational infrastructure challenges to make AI systems scalable and reliable.

San Francisco hybrid FullTime
KubernetesSpark

Software Engineer, Systems Generalist

5mo ago
thinkingmachines

thinkingmachines

Thinking Machines Lab is seeking generalist infrastructure and systems engineers to build the core systems powering their foundation models and support internal research and product development teams. This high-impact role involves architecting and scaling critical infrastructure across the full technical stack, solving complex distributed systems problems, and building robust, scalable platforms. You will work directly with researchers to accelerate experiments, improve infrastructure efficiency, and enable key insights across models, products, and data assets.

$350k - $475k

San Francisco onsite
OpenAIMistralKubernetes +5 more

Member of Technical Staff (Offensive Security Engineer)

5mo ago
P

Perplexity AI

Perplexity is looking for a skilled Offensive Security Engineer to join its security team. This role involves taking an adversarial approach to protect Perplexity's infrastructure, applications, and AI systems. You will be responsible for planning and executing red team operations, penetration tests, and attack simulations across various environments, including cloud infrastructure, web and mobile applications, and the AI/ML pipeline. The goal is to identify vulnerabilities before malicious actors do and collaborate with engineering teams to ensure effective remediation.

San Francisco hybrid FullTime
AWSAzureKubernetes +3 more

CyberSecurity Engineer, Offensive Security

5mo ago
Mistral AI

Mistral AI

Mistral AI is a pioneering company at the forefront of AI innovation, developing cutting-edge models, products, and solutions that democratize AI. We are building products like Mistral Studio and Mistral Vibe that redefine user interaction with AI. Our mission is to integrate AI seamlessly into daily life, enhancing learning, creativity, and efficiency. We are a dynamic, collaborative, and diverse team passionate about AI's potential to transform society, with a global presence across France, the USA, the UK, Germany, and Singapore. We are seeking a Security Researcher to safeguard our innovations by anticipating, identifying, and mitigating risks, embedding an attacker's mindset into our product development.

Paris remote Full-time
MistralAWSAzure +8 more

Staff Site Reliability Engineer

5mo ago
Harvey

Harvey

Harvey is transforming legal and professional services by combining frontier agentic AI, an enterprise-grade platform, and deep domain expertise. This is a rare opportunity to help build a generational company at an inflection point, with strong product-market fit and world-class investor support. The team moves fast, takes ownership, and is deeply committed to the mission, operating with intensity and pushing for excellence. As a Staff Software Engineer on the Site Reliability team, you will ensure the reliability, scalability, and performance of our legal AI platform, owning the systems that keep our platform fast, secure, and always on. Your work will be critical in scaling across 50+ regions and automating mission-critical operations to ensure Harvey remains resilient as we grow. If you are passionate about building robust systems and reducing complexity through automation, we encourage you to apply.

Bengaluru hybrid FullTime
AWSAzureKubernetes +4 more

Software Engineer, Workload Enablement

5mo ago
OpenAI

OpenAI

OpenAI is seeking a Software Engineer to join the Scaling team, which builds the architectural and engineering backbone for the company's infrastructure. This role focuses on enabling production workloads and end-to-end testing on new platforms. You will be responsible for creating test harnesses, developing platform stress benchmarks, and porting existing AI model training and inference workloads to new systems and hardware. A key aspect of this position involves analyzing performance, identifying bottlenecks, and characterizing the behavior of new compute, communication, storage, and control plane systems, including their failure modes.

San Francisco hybrid FullTime
OpenAIKubernetesPython +3 more

Staff Software Engineer, Core Infrastructure

5mo ago
Harvey

Harvey

Harvey is transforming legal and professional services by combining frontier agentic AI, an enterprise-grade platform, and deep domain expertise. This is a rare opportunity to help build a generational company at an inflection point, with strong product-market fit and significant investor backing. The team moves fast, takes ownership, and is deeply committed to the mission, operating with intensity and pushing for excellence. As a Staff Software Engineer on the Core Infrastructure team, you will play a critical role in designing, building, scaling, and strengthening the infrastructure that powers every user interaction with Harvey's AI platform. You will work in an environment balanced between innovation and operational excellence, ensuring the platform remains resilient and efficient as it scales.

Bengaluru hybrid FullTime
AWSAzureKubernetes +5 more

CyberSecurity Engineer, DevSecOps

5mo ago
Mistral AI

Mistral AI

Mistral AI is seeking a DevSecOps Engineer to architect and maintain the security of its rapidly scaling AI infrastructure and application lifecycle. The role involves treating security as a seamless enabler for research and engineering teams, embedding robust security controls into CI/CD pipelines, infrastructure, and developer workflows without compromising deployment velocity. The company is a dynamic, collaborative team passionate about AI, democratizing it through open-source models and products, and aims to make a meaningful impact by shaping the future of AI.

Paris remote Full-time
MistralKubernetesPython +8 more

Staff Software Engineer, Enterprise Platform

5mo ago
Replit

Replit

Join our Enterprise Platform team and build the infrastructure foundations that enable the world's largest organizations to run Replit within their security and compliance boundaries. As a Software Engineer on this team, you'll design and implement the deployment flexibility, networking capabilities, authorization systems, and data controls that enterprises require, from single-tenant architectures and private connectivity to custom policy enforcement and customer-managed encryption. You'll work at the intersection of cloud infrastructure and enterprise requirements, partnering with Platform Engineering, Security, and Sales to ship capabilities that unlock adoption at demanding organizations.

Foster City, CA hybrid FullTime
KubernetesPythonTypeScript +3 more

Generative AI Inference Engineer

5mo ago
Stability AI

Stability AI

We are seeking passionate Machine Learning Engineers to join our Inference team, focusing on the creative applications of generative AI models. The ideal candidate will have substantial experience developing and running inference for multi-modal models. A deep understanding of diffusion model architectures and familiarity with workflow tools like ComfyUI are a big plus. You will be expected to leverage and push the boundaries of state-of-the-art inference optimization techniques for multi-modal generative models. This role offers the opportunity to work alongside top researchers and engineers, utilizing cutting-edge high-performance computing resources to make a significant impact in the rapidly evolving field of generative AI.

United States remote
AWSAzureDocker +10 more

Solutions Architect - Public Sector

5mo ago
Cohere

Cohere

Cohere is seeking a Solutions Architect for its Public Sector business, offering significant autonomy in technical pre-sales and post-sales activities. This role is crucial for growing Cohere's presence in the public sector and requires a deep understanding of customer challenges to map them to Cohere's AI solutions. You will act as a trusted technical advisor, fostering strong relationships with stakeholders and driving the adoption of Cohere products. This position demands a blend of strategic thinking and hands-on execution, including building customer Proof of Concepts (PoCs) and translating business objectives into technical solutions. You will also serve as the voice of the customer, liaising with the product team, providing guidance on best practices, and identifying areas for platform improvement.

Washington, DC hybrid FullTime
CohereKubernetesPython +1 more

Solutions Architect - Public Sector

5mo ago
Cohere

Cohere

Cohere is seeking a Solutions Architect to play a significant role in growing the company's Defence and National Security business. This dynamic role requires a blend of strategic thinking and hands-on execution, focusing on building customer demos and proof-of-concepts that highlight the business value of Cohere's AI platform. As the technical relationship owner, you will collaborate with stakeholders to understand their objectives and translate them into technical solutions. You will also serve as the voice of the customer, liaising between clients and the product team, providing guidance on best practices, identifying platform improvements, and cultivating technical champions within customer organizations to drive adoption and gather feedback.

Ottawa hybrid FullTime
CohereAWSAzure +5 more

Technical Support Engineer - Use Cases

5mo ago
Mistral AI

Mistral AI

Mistral AI is seeking a Technical Support Engineer - Use Cases to join our Support team. This role is ideal for someone who excels at technical troubleshooting, incident investigation, and customer communication in a B2B environment. You will be responsible for maintaining and ensuring the run of customer applications, handling escalated technical issues from enterprise clients, reproducing complex problems, and collaborating with engineering, data, and product teams for swift resolution. This is a unique opportunity to work at the intersection of AI infrastructure, customer success, and technical problem-solving, reporting directly to the Head of Support.

Marseille hybrid Full-time
MistralKubernetesPython +8 more

Software Engineer, Localization

5mo ago
OpenAI

OpenAI

OpenAI's Internationalization team is seeking a Senior Software Engineer to build the systems that power localization and international product launches. This role involves working on the platform that manages product content, translation workflows, and localization infrastructure across all OpenAI products. It sits at the intersection of AI systems, developer platforms, and product infrastructure, aiming to ensure OpenAI products work seamlessly across languages, regions, and cultures.

San Francisco onsite FullTime
OpenAIKubernetesJava

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.